Machine Learning Engineer – Voice AI, Generative Music

Posted 2ds ago

Employment Information

Education
Salary
Experience
Job Type

Report this job

Job expired or something wrong with this job?

Job Description

Applied ML Engineer building voice-cloning, text-to-speech, and Spanish lyrics-generation systems. Deploying production AI models for an AI-native music label.

Responsibilities:

  • Own the end-to-end machine-learning workflow from audio data preparation and training experiments through optimized inference and production deployment
  • Fine-tune and improve a singing-voice conversion and cloning model to achieve at least 90% fidelity in blind listening tests
  • Reproduce expressive vocal characteristics including vibrato, falsetto, dynamics, spoken delivery, and sung delivery
  • Evaluate dataset expansion, alternative base models, and enterprise APIs with no-training guarantees
  • Own an existing Chatterbox Multilingual LoRA fine-tune for a Spanish-speaking voice
  • Ensure accurate Mexican-accent reproduction and pronunciation of J, Ñ, and X
  • Containerize and deploy the text-to-speech model as a serverless inference endpoint using RunPod or a similar platform
  • Integrate the inference endpoint with the existing web platform through an API
  • Train a second text-to-speech version using clean studio recordings to improve output stability
  • Develop a proprietary Spanish-language lyrics-generation model using an open-weight LLM, LoRA fine-tuning, DPO, and RAG
  • Establish evaluation criteria and organize quality assessment with native Spanish speakers
  • Integrate the completed lyrics model into the existing frontend

Requirements:

  • 3+ years of experience training and deploying deep learning models in production, preferably within audio or NLP
  • Hands-on experience in at least two of: Voice cloning or Singing Voice Conversion (SVC); Text-to-Speech model fine-tuning; LLM fine-tuning using LoRA, DPO, and RAG
  • Strong proficiency in PyTorch
  • Practical experience with GPU cloud infrastructure such as RunPod or AWS, Docker, and serverless inference
  • Evidence-based model evaluation, including benchmarks, ablation studies, and blind testing
  • Native or strong professional proficiency in Spanish, required for lyrics evaluation and voice quality assurance
  • Nice to have: music background or experience with stems, MIDI, and DAWs
  • Nice to have: experience with ACE Studio, ACE-Step, RVC, so-vits-svc, or similar singing-voice synthesis technologies
  • Nice to have: experience with licensed celebrity or artist voices and consent-based voice AI

Benefits:

  • Security through client vetting to minimize risks and ensure reliability and timely payments
  • Career support and help finding new opportunities if a project is not the right fit
  • Legal assistance with independent contractor or sole proprietorship status, taxes, and related processes
  • English courses
  • Professional growth opportunities
  • Team-building events
  • Flexible working hours
  • 29 days of PTO (18 working days per year plus all national holidays)
  • 10 paid recovery days
  • Full financial and legal support for independent contractors
  • Free English classes with native speakers or Ukrainian teachers
  • Dedicated HR support