خانه هوش ایران
خانه هوش ایران

Speech AI Engineer (ASR/TTS)

Tehran/Meydan Valiasr
Full Time
Saturday to Wednesday from 8 to 17
-
Health insurance -Lunch

این فرصت شغلی چقدر برای من مناسب است؟

11 - 50 employees
Technology and Innovation / VC / Accelerator
توضیحات بیشتر

key Requirements

2 years experience in similar position
Bachelor Computer and IT or Electrical Engineering
Python - Intermediate
GIT - Intermediate
Linux - Intermediate

Job Description

About the Role:
We are looking for a talented Speech AI Engineer to join our AI team and develop state-of-the-art speech technologies. You will be responsible for designing, training, optimizing, and deploying Automatic Speech Recognition (ASR) and Text-to-Speech (TTS) systems for production environments.

Responsibilities:
Design, train, and fine-tune ASR and TTS models.
Build and maintain speech data processing pipelines.
Optimize models for low-latency and high-throughput inference.
Deploy speech models to production environments.
Benchmark different model architectures and inference engines.
Collaborate with AI, Platform, and Product teams to deliver scalable speech solutions.
Stay up to date with the latest advances in Speech AI and Large Language Models.

Required Qualifications:
Bachelor's or Master's degree in Computer Science, Artificial Intelligence, Electrical Engineering, or a related field.
At least 2 years of professional experience in Artificial Intelligence or Machine Learning.
Experience with Automatic Speech Recognition (ASR), Text-to-Speech (TTS), or related speech technologies.
Strong programming skills in Python and familiarity with C/C++.
Strong understanding of Deep Learning fundamentals and hands-on experience with PyTorch.
Good understanding of Transformer-based architectures and modern sequence models.
Familiarity with Linux, Git, and software development best practices.
Familiarity with High Performance Computing (HPC) concepts, including distributed training, GPU computing, and parallel processing.

Nice to Have:
Experience with NVIDIA NeMo, Whisper, wav2vec 2.0, Parakeet, or similar speech frameworks.
Familiarity with inference engines such as GGML, CTranslate2, ONNX Runtime, TensorRT, or TensorRT-LLM.
Experience with model quantization, pruning, distillation, or efficient inference.
Experience with CUDA, NCCL, Triton Inference Server, or multi-GPU environments.
Experience with Docker, Kubernetes, or cloud-based ML deployment.
Contributions to open-source AI or speech-related projects.

What We Offer:
Opportunity to work on cutting-edge Speech AI technologies.
Access to high-performance GPU infrastructure.
Research-driven engineering culture.
Competitive compensation and opportunities for professional growth.
Flexible and collaborative working environment.

Job Requirements

Gender
Men / Women
Education
Bachelor| Computer and IT Bachelor| Electrical Engineering
Software
Python| Intermediate Linux| Intermediate GIT| Intermediate

ثبت مشکل و تخلف آگهی

ارسال رزومه برای خانه هوش ایران