Type
Full-time
Work mode
On-site
Level
Staff
Industry
Engineering / Mechanical
Salary
Thương lượng
Location
Quận Thanh Xuân, Hà Nội, Hà Nội
Overview
- Explore and prototype new approaches for speech and audio AI problems.
- Develop, fine-tune, and adapt speech models for real-world use cases.
- Evaluate and iterate on different modeling approaches to improve model performance.
- Translate promising research findings into solutions that can be integrated into our products.
- Automatic Speech Recognition (Speech-to-Text)
- Text-to-Speech (Speech Synthesis)
- Audio processing and speech enhancement
- Speech understanding and voice interaction pipelines
- Audio preprocessing and feature extraction
- Noise suppression and speech enhancement
- Real-time audio streaming and processing
- Model serving and inference pipelines
- Integration of AI models into cloud and mobile environments
- Voice recording and transcription
- AI-powered voice assistants
- Voice-driven workflows
- Natural conversational experiences
- AI capabilities embedded into mobile and business workflows
- Improving speech recognition accuracy and robustness
- Improving speech synthesis quality and naturalness
- Reducing inference latency and resource consumption
- Optimizing models for different deployment environments
- Building scalable and reliable AI services
- Monitoring real-world model performance and using production feedback to drive further improvements
- 3+ years of experience building AI/ML systems or AI-powered products.
- Strong Python programming skills.
- Experience with deep learning frameworks such as PyTorch or TensorFlow.
- Hands-on experience deploying AI models into production.
- Solid understanding of speech processing and audio machine learning.
- Experience working on one or more of the following:• Automatic Speech Recognition (ASR)
- Text-to-Speech (TTS)
- Audio AI or Voice AI applications
- Ability to design and implement end-to-end AI workflows.
- Strong problem-solving skills with a product-oriented mindset.
- Research, reproduce, and evaluate state-of-the-art Speech AI methods, and translate promising ideas into real-world products.
- Design rigorous experiments, benchmarks, and systematic error analyses to understand model limitations and identify opportunities for improvement.
- Fine-tune and adapt speech models, and explore new modeling approaches when existing solutions are insufficient.
- Build and improve high-quality training and evaluation datasets, including data collection, synthetic data generation, augmentation, and quality validation.
- Continuously improve models and data based on experimental results and real-world production feedback.
- AI deployment on mobile devices
Benefits
BonusTrainingStock options / ESOP
Summary of facts from the official posting. View original ↗
Interested in this role?
You'll be taken to the employer's official application page.
Apply on official site ↗
Is this your business?
Claim this page, request edits or removal
→