Machine Learning Engineer – Voice AI, Generative Music

Job not on LinkedIn

🔥 16 hours ago

🇲🇽 Mexico – Remote

⏱ Part Time

🟡 Mid-level

🟠 Senior

🤖 Machine Learning Engineer

👻 Ghost score 25%

infoinfo

🗣️🇪🇸 Spanish Required

Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of MWDN

MWDN

51 - 200 employees

💼 Consulting

📦 Logistics

🏥 Healthcare

Consulting • Logistics • Healthcare

MWDN is a comprehensive IT services company that specializes in software development, web design, mobile development (iOS and Android), and desktop development. With over 200 in-house software engineers, MWDN provides a range of services including staff augmentation, dedicated teams, and global IT recruitment. Operating globally with delivery centers in Central and Eastern Europe, MWDN effectively bridges businesses in the USA, Europe, and Israel to top-tier tech talent while focusing on cost-effective solutions and personalized services tailored to client needs.

📋 Description

• Fine-tune and improve a singing-voice conversion and cloning model to achieve at least 90% quality in blind listening tests • Reproduce expressive vocal characteristics including vibrato, falsetto, dynamics, and spoken and sung delivery • Evaluate training dataset expansion, alternative base models, and enterprise APIs with no-training guarantees • Own and improve a Spanish-speaking Chatterbox Multilingual LoRA fine-tune • Ensure accurate Mexican accent reproduction and pronunciation of J, Ñ, and X • Containerize the text-to-speech model and deploy it as a serverless inference endpoint using RunPod or a similar platform • Integrate the inference endpoint with the existing web platform through an API • Train a second text-to-speech version using clean studio recordings • Develop a proprietary Spanish-language lyrics-generation model using an open-weight LLM, LoRA fine-tuning, DPO, and RAG • Establish evaluation criteria and organize quality assessment with native Spanish speakers • Integrate the completed lyrics model into the existing frontend

🎯 Requirements

• 3+ years of experience training and deploying deep learning models in production, preferably within audio or NLP • Hands-on experience in at least two of: voice cloning or Singing Voice Conversion (SVC); Text-to-Speech model fine-tuning; LLM fine-tuning using LoRA, DPO, and RAG • Strong proficiency in PyTorch • Practical experience with GPU cloud infrastructure such as RunPod or AWS • Experience with Docker and serverless inference • Evidence-based model evaluation, including benchmarks, ablation studies, and blind testing • Native or strong professional proficiency in Spanish, required for lyrics evaluation and voice quality assurance • Nice to have: music background or experience with stems, MIDI, and DAWs • Nice to have: singing-voice synthesis experience with ACE Studio, ACE-Step, RVC, so-vits-svc, or similar technologies • Nice to have: experience working with licensed celebrity or artist voices and consent-based voice AI

🏖️ Benefits

• Security through client vetting to minimize risks and ensure reliable, timely payments • Career support and help finding new opportunities if a project is not the right fit • Legal assistance with independent contractor or sole proprietorship status, taxes, and related processes • English courses • Professional growth opportunities • Team-building events • Flexible working hours • 29 days of PTO (18 working days per year plus all national holidays) • 10 paid recovery days • Full financial and legal support for independent contractors • Free English classes with native speakers or Ukrainian teachers • Dedicated HR support

Apply Now