Skip to content
Tuyển Dụng
← Base Inc.

AI Engineer Voice Conversational AI

Base Inc. · Hà Nội
Apply on official site ↗
Type
Full-time
Work mode
On-site
Level
Staff
Industry
Engineering / Mechanical
Salary
Thương lượng
Location
Quận Thanh Xuân, Hà Nội, Hà Nội

Overview

  • Explore and prototype new approaches for speech and audio AI problems.
  • Develop, fine-tune, and adapt speech models for real-world use cases.
  • Evaluate and iterate on different modeling approaches to improve model performance.
  • Translate promising research findings into solutions that can be integrated into our products.
  • Automatic Speech Recognition (Speech-to-Text)
  • Text-to-Speech (Speech Synthesis)
  • Audio processing and speech enhancement
  • Speech understanding and voice interaction pipelines
  • Audio preprocessing and feature extraction
  • Noise suppression and speech enhancement
  • Real-time audio streaming and processing
  • Model serving and inference pipelines
  • Integration of AI models into cloud and mobile environments
  • Voice recording and transcription
  • AI-powered voice assistants
  • Voice-driven workflows
  • Natural conversational experiences
  • AI capabilities embedded into mobile and business workflows
  • Improving speech recognition accuracy and robustness
  • Improving speech synthesis quality and naturalness
  • Reducing inference latency and resource consumption
  • Optimizing models for different deployment environments
  • Building scalable and reliable AI services
  • Monitoring real-world model performance and using production feedback to drive further improvements
  • 3+ years of experience building AI/ML systems or AI-powered products.
  • Strong Python programming skills.
  • Experience with deep learning frameworks such as PyTorch or TensorFlow.
  • Hands-on experience deploying AI models into production.
  • Solid understanding of speech processing and audio machine learning.
  • Experience working on one or more of the following:• Automatic Speech Recognition (ASR)
  • Text-to-Speech (TTS)
  • Audio AI or Voice AI applications
  • Ability to design and implement end-to-end AI workflows.
  • Strong problem-solving skills with a product-oriented mindset.
  • Research, reproduce, and evaluate state-of-the-art Speech AI methods, and translate promising ideas into real-world products.
  • Design rigorous experiments, benchmarks, and systematic error analyses to understand model limitations and identify opportunities for improvement.
  • Fine-tune and adapt speech models, and explore new modeling approaches when existing solutions are insufficient.
  • Build and improve high-quality training and evaluation datasets, including data collection, synthetic data generation, augmentation, and quality validation.
  • Continuously improve models and data based on experimental results and real-world production feedback.
  • AI deployment on mobile devices

Benefits

BonusTrainingStock options / ESOP

Summary of facts from the official posting. View original ↗

Interested in this role?

You'll be taken to the employer's official application page.

Apply on official site ↗
Is this your business? Claim this page, request edits or removal