Job Description
About the job
Company Description
Eastgate Software is a Vietnam-based software development company with consulting teams in Germany, Japan, and Vietnam, serving clients worldwide through remote cooperation models. The company builds world-class distributed teams and delivers custom software solutions for SMEs and Fortune 500 companies, with a strong emphasis on reliability and quality. Eastgate Software is a long-term strategic partner of SIEMENS Mobility, reflecting its track record in complex, high-stakes projects. Team members join a diverse pool of 250+ experienced professionals, including Software Developers, Architects, QA Engineers, AI/ML Scientists, Business Analysts, and Project Managers. Cooperation models range from project-based engagements to dedicated outsourcing teams, offering variety in domains and technologies.
Role Description
We are looking for a Middle - Senior AI Engineer with experience in Voice AI to develop AI-powered audio analysis and singing evaluation features. You will research, integrate, optimize, and deploy machine learning models into production while working closely with backend and frontend engineers.
- Level: Middle +
- Working model: Remote (with EGS Watch installed)
- Candidate preference: Vietnamese
- Duration: 3 months
- Expected start date: August 2026
- Working hours: Part-time (4 hours/day), preferably within Vietnam business hours, 9:30 AM - 3:00 PM UTC+7
Responsibilities
- Research and evaluate voice AI models.
- Develop and optimize audio processing pipelines.
- Build real-time voice analysis and scoring systems.
- Deploy AI models to cloud infrastructure.
- Improve model accuracy, latency, and scalability.
- Collaborate with cross-functional teams to deliver production-ready AI solutions.
Requirements
- Bachelor's degree in Computer Science, AI, or related fields.
- 3+ years of experience in AI/ML development.
- Strong Python programming skills.
- Experience building and deploying production AI models.
- Experience with PyTorch or TensorFlow.
- Experience with cloud platforms (Azure, AWS, or GCP).
Voice AI Experience: Candidates should have experience in several of the following areas:
- Speech Recognition (ASR)
- Audio Signal Processing (DSP)
- Source Separation
- Pitch Detection
- Beat & Tempo Detection
- Forced Alignment
- Music Information Retrieval (MIR)
- Real-time Audio Processing
- Audio Streaming
Nice to Have
- Experience with karaoke or singing evaluation systems.
- Experience with multilingual speech processing.
- Experience with ONNX Runtime, TensorRT, or CUDA optimization.
- Experience with AI-powered learning or recommendation systems.