# Voice AI & Generative Music (Part-time)

[Mwdn](https://gurify.com/jobs?q=Mwdn) · Kyiv, Ukraine · Posted today

Part-time

[AI & ML](https://gurify.com/jobs/ai-ml)

[Apply on the original posting → (opens in a new tab)](https://www.comeet.com/jobs/mwdn/61.005/machine-learning-engineer---voice-ai--generative-music-part-time/2C.270)

## Job description

MWDN is a global IT outstaffing company with 23+ years of experience that connects exceptional tech talent with leading companies across Israel, the USA, Great Britain, and Western Europe. We offer opportunities to work on international products in a stable and professional environment.

### Why does MWDN rock?

### Here’s what you can expect when you join MWDN:

- Security: We carefully vet our clients to minimize risks and ensure reliability and timely payments - no fraud or unpleasant surprises.

- Career support: If a project isn’t the right fit, we support you and actively help find new opportunities that match your skills and career goals.

- Legal assistance: We provide guidance on legal matters, including opening and managing your independent contractor or sole proprietorship status, taxes, and related processes.

- Professional development: We offer English courses and professional growth opportunities, as well as team-building events.

Why choose us? MWDN is ranked among the top 5 IT employers in our region according to DOU. We take pride in our transparency and strong commitment to our team. Curious to learn more? See what our employees say about working with us on DOU.

### What is your new project?

### Domain: Music Technology / Voice AI / Generative AI

An AI-native music label developing advanced tools for producing and releasing Spanish-language music. Artificial intelligence serves as the foundation of its production process, powering artist-consistent audio generation, vocal cloning, custom text-to-speech, and generative lyrics.

The company has built its own end-to-end music production pipeline covering audio data preparation, model training, vocal conditioning, inference, quality evaluation, and deployment. The environment is hands-on, fast-moving, and focused on turning sophisticated AI models into release-ready music products rather than experimental prototypes.

### What makes this project exciting?

Our client is the technology-driven music label where artificial intelligence powers the entire content-production process. The company develops artist-specific audio, creates custom vocal models, and operates a proprietary system designed to produce and release music faster than conventional labels.

Each track moves through an internally developed workflow that covers audio data preparation, model training, vocal conditioning, generation, quality control, and production deployment.

We are looking for an Applied Machine Learning Engineer specializing in audio and generative AI to take ownership of this workflow. This is a practical engineering position rather than a research, analytics, or prompt-only role. The ideal candidate has already built and launched complete ML systems, taking them from raw datasets and training experiments through optimized inference and production use.

### What makes you a great fit

- 3+ years of experience training and deploying deep learning models in production, preferably within audio or NLP.

- Hands-on experience in at least two of the following areas: Voice cloning or Singing Voice Conversion (SVC), Text-to-Speech model fine-tuning, LLM fine-tuning using techniques such as LoRA, DPO, and RAG.

- Strong proficiency in PyTorch.

- Practical experience with GPU cloud infrastructure, such as RunPod or AWS, as well as Docker and serverless inference.

- A rigorous, evidence-based approach to model evaluation, including benchmarks, ablation studies, and blind testing.

- Native or strong professional proficiency in Spanish, required for lyrics evaluation and voice quality assurance.

### Nice to Have

- A music background or experience with music-production tools and workflows, including stems, MIDI, and DAWs.

- Experience with singing-voice synthesis solutions such as ACE Studio, ACE-Step, RVC, so-vits-svc, or similar technologies.

- Previous experience working with licensed celebrity or artist voices and consent-based voice AI.

### Your day-to-day in this position

### 1.Custom Singing Voice Model

- Fine-tune and improve an existing singing-voice conversion and cloning model, increasing its current 70–75% fidelity baseline to production-level quality of 90% or higher in blind listening tests.

- Reproduce expressive vocal characteristics beyond basic timbre, including vibrato, falsetto, dynamics, and both spoken and sung delivery.

- Evaluate different improvement strategies, including expanding the training dataset, testing alternative base models, and assessing enterprise APIs that offer no-training guarantees.

### 2. Text-to-Speech Model

- Take ownership of an existing validated LoRA fine-tune based on Chatterbox Multilingual for a Spanish-speaking voice.

- Ensure accurate reproduction of a Mexican accent and correct pronunciation of sounds and letters such as J, Ñ, and X.

- Containerize the model and deploy it as a serverless inference endpoint using RunPod or a similar platform.

- Integrate the endpoint with the existing web platform through an API.

- Train a second version using clean studio recordings to eliminate the remaining output instability.

### 3. ALMA — Spanish Lyrics Model

- Develop a proprietary Spanish-language lyrics-generation model, with experience in regional Mexican genres considered a strong advantage.

- Build the solution using an open-weight LLM, LoRA fine-tuning, preference optimization through DPO, and RAG based on a curated content corpus.

- Establish clear evaluation criteria and organize quality assessment with native Spanish speakers.

- Integrate the completed model into the existing frontend.

### Why work with us?

- People-first management with minimal bureaucracy

- A friendly company culture, proven by employees who choose to return

- Flexible working hours

- 29 days of PTO (18 working days per year pluse all national holidays)

- 10 paid recovery days

- Full financial and legal support for independent contractors

- Free English classes, with native speakers or Ukrainian teachers

- Dedicated HR support

### Our next steps

✅ Intro call with a Recruiter — ✅ Intro call with a client — ✅ Final client interview— ✅ Offer

**Live in Mwdn’s hiring system.** Read from the company's own applicant tracking system, not reposted from a job board — so it's a real, open requisition rather than an ad that outlived the role.

We remove it as soon as it disappears at source.

## More jobs like this

- 5B [Senior Full Stack AI-first Saas Engineer](https://gurify.com/job/senior-full-stack-ai-first-saas-engineer-at-5bluesoftware-0e4809e85383) 5Bluesoftware · Ukraine · last week
- MW [AI Video Creator](https://gurify.com/job/ai-video-creator-at-mwdn-423088efc7c6) Mwdn · Kyiv, Ukraine · 2 months ago
- IS [Senior AI Product Engineer](https://gurify.com/job/senior-ai-product-engineer-at-ispeedtolead-b782cc4c6b74) Ispeedtolead · Ukraine · 2 months ago
- VO [AI Product Manager - Paper.io 2](https://gurify.com/job/ai-product-manager-paper-io-2-at-voodoo-7f396f5e303e) Voodoo · Paris · last week
- DA [Solutions Architect (Presales, Data & AI, Digital Native Business)](https://gurify.com/job/solutions-architect-presales-data-ai-digital-native-business-at-2b32bb859961) Databricks · London, United Kingdom · last week
- RI [Software Engineer, GPU Kernels](https://gurify.com/job/software-engineer-gpu-kernels-at-riverai-d38198190289) Riverai · Palo Alto, CA · 4 days ago

```json
{"@context":"https://schema.org/","@type":"JobPosting","title":"Voice AI \u0026 Generative Music (Part-time)","description":"\u003Cp\u003EMWDN is a global IT outstaffing company with 23\u002B years of experience that connects exceptional tech talent with leading companies across Israel, the USA, Great Britain, and Western Europe. We offer opportunities to work on international products in a stable and professional environment.\u003C/p\u003E\u003Ch3\u003EWhy does MWDN rock?\u003C/h3\u003E\u003Ch3\u003EHere\u2019s what you can expect when you join MWDN:\u003C/h3\u003E\u003Cul\u003E\u003Cli\u003ESecurity: We carefully vet our clients to minimize risks and ensure reliability and timely payments - no fraud or unpleasant surprises.\u003C/li\u003E\u003C/ul\u003E\u003Cul\u003E\u003Cli\u003ECareer support: If a project isn\u2019t the right fit, we support you and actively help find new opportunities that match your skills and career goals.\u003C/li\u003E\u003C/ul\u003E\u003Cul\u003E\u003Cli\u003ELegal assistance: We provide guidance on legal matters, including opening and managing your independent contractor or sole proprietorship status, taxes, and related processes.\u003C/li\u003E\u003C/ul\u003E\u003Cul\u003E\u003Cli\u003EProfessional development: We offer English courses and professional growth opportunities, as well as team-building events.\u003C/li\u003E\u003C/ul\u003E\u003Cp\u003EWhy choose us? MWDN is ranked among the top 5 IT employers in our region according to DOU. We take pride in our transparency and strong commitment to our team. Curious to learn more? See what our employees say about working with us on DOU.\u003C/p\u003E\u003Ch3\u003EWhat is your new project?\u003C/h3\u003E\u003Ch3\u003EDomain: Music Technology / Voice AI / Generative AI\u003C/h3\u003E\u003Cp\u003EAn AI-native music label developing advanced tools for producing and releasing Spanish-language music. Artificial intelligence serves as the foundation of its production process, powering artist-consistent audio generation, vocal cloning, custom text-to-speech, and generative lyrics.\u003C/p\u003E\u003Cp\u003EThe company has built its own end-to-end music production pipeline covering audio data preparation, model training, vocal conditioning, inference, quality evaluation, and deployment. The environment is hands-on, fast-moving, and focused on turning sophisticated AI models into release-ready music products rather than experimental prototypes.\u003C/p\u003E\u003Ch3\u003EWhat makes this project exciting?\u003C/h3\u003E\u003Cp\u003EOur client is the technology-driven music label where artificial intelligence powers the entire content-production process. The company develops artist-specific audio, creates custom vocal models, and operates a proprietary system designed to produce and release music faster than conventional labels.\u003C/p\u003E\u003Cp\u003EEach track moves through an internally developed workflow that covers audio data preparation, model training, vocal conditioning, generation, quality control, and production deployment.\u003C/p\u003E\u003Cp\u003EWe are looking for an Applied Machine Learning Engineer specializing in audio and generative AI to take ownership of this workflow. This is a practical engineering position rather than a research, analytics, or prompt-only role. The ideal candidate has already built and launched complete ML systems, taking them from raw datasets and training experiments through optimized inference and production use.\u003C/p\u003E\u003Ch3\u003EWhat makes you a great fit\u003C/h3\u003E\u003Cul\u003E\u003Cli\u003E3\u002B years of experience training and deploying deep learning models in production, preferably within audio or NLP.\u003C/li\u003E\u003C/ul\u003E\u003Cul\u003E\u003Cli\u003EHands-on experience in at least two of the following areas: Voice cloning or Singing Voice Conversion (SVC), Text-to-Speech model fine-tuning, LLM fine-tuning using techniques such as LoRA, DPO, and RAG.\u003C/li\u003E\u003C/ul\u003E\u003Cul\u003E\u003Cli\u003EStrong proficiency in PyTorch.\u003C/li\u003E\u003C/ul\u003E\u003Cul\u003E\u003Cli\u003EPractical experience with GPU cloud infrastructure, such as RunPod or AWS, as well as Docker and serverless inference.\u003C/li\u003E\u003C/ul\u003E\u003Cul\u003E\u003Cli\u003EA rigorous, evidence-based approach to model evaluation, including benchmarks, ablation studies, and blind testing.\u003C/li\u003E\u003C/ul\u003E\u003Cul\u003E\u003Cli\u003ENative or strong professional proficiency in Spanish, required for lyrics evaluation and voice quality assurance.\u003C/li\u003E\u003C/ul\u003E\u003Ch3\u003ENice to Have\u003C/h3\u003E\u003Cul\u003E\u003Cli\u003EA music background or experience with music-production tools and workflows, including stems, MIDI, and DAWs.\u003C/li\u003E\u003C/ul\u003E\u003Cul\u003E\u003Cli\u003EExperience with singing-voice synthesis solutions such as ACE Studio, ACE-Step, RVC, so-vits-svc, or similar technologies.\u003C/li\u003E\u003C/ul\u003E\u003Cul\u003E\u003Cli\u003EPrevious experience working with licensed celebrity or artist voices and consent-based voice AI.\u003C/li\u003E\u003C/ul\u003E\u003Ch3\u003EYour day-to-day in this position\u003C/h3\u003E\u003Ch3\u003E1.Custom Singing Voice Model\u003C/h3\u003E\u003Cul\u003E\u003Cli\u003EFine-tune and improve an existing singing-voice conversion and cloning model, increasing its current 70\u201375% fidelity baseline to production-level quality of 90% or higher in blind listening tests.\u003C/li\u003E\u003C/ul\u003E\u003Cul\u003E\u003Cli\u003EReproduce expressive vocal characteristics beyond basic timbre, including vibrato, falsetto, dynamics, and both spoken and sung delivery.\u003C/li\u003E\u003C/ul\u003E\u003Cul\u003E\u003Cli\u003EEvaluate different improvement strategies, including expanding the training dataset, testing alternative base models, and assessing enterprise APIs that offer no-training guarantees.\u003C/li\u003E\u003C/ul\u003E\u003Ch3\u003E2. Text-to-Speech Model\u003C/h3\u003E\u003Cul\u003E\u003Cli\u003ETake ownership of an existing validated LoRA fine-tune based on Chatterbox Multilingual for a Spanish-speaking voice.\u003C/li\u003E\u003C/ul\u003E\u003Cul\u003E\u003Cli\u003EEnsure accurate reproduction of a Mexican accent and correct pronunciation of sounds and letters such as J, \u0026#209;, and X.\u003C/li\u003E\u003C/ul\u003E\u003Cul\u003E\u003Cli\u003EContainerize the model and deploy it as a serverless inference endpoint using RunPod or a similar platform.\u003C/li\u003E\u003C/ul\u003E\u003Cul\u003E\u003Cli\u003EIntegrate the endpoint with the existing web platform through an API.\u003C/li\u003E\u003C/ul\u003E\u003Cul\u003E\u003Cli\u003ETrain a second version using clean studio recordings to eliminate the remaining output instability.\u003C/li\u003E\u003C/ul\u003E\u003Ch3\u003E3. ALMA \u2014 Spanish Lyrics Model\u003C/h3\u003E\u003Cul\u003E\u003Cli\u003EDevelop a proprietary Spanish-language lyrics-generation model, with experience in regional Mexican genres considered a strong advantage.\u003C/li\u003E\u003C/ul\u003E\u003Cul\u003E\u003Cli\u003EBuild the solution using an open-weight LLM, LoRA fine-tuning, preference optimization through DPO, and RAG based on a curated content corpus.\u003C/li\u003E\u003C/ul\u003E\u003Cul\u003E\u003Cli\u003EEstablish clear evaluation criteria and organize quality assessment with native Spanish speakers.\u003C/li\u003E\u003C/ul\u003E\u003Cul\u003E\u003Cli\u003EIntegrate the completed model into the existing frontend.\u003C/li\u003E\u003C/ul\u003E\u003Ch3\u003EWhy work with us?\u003C/h3\u003E\u003Cul\u003E\u003Cli\u003EPeople-first management with minimal bureaucracy\u003C/li\u003E\u003C/ul\u003E\u003Cul\u003E\u003Cli\u003EA friendly company culture, proven by employees who choose to return\u003C/li\u003E\u003C/ul\u003E\u003Cul\u003E\u003Cli\u003EFlexible working hours\u003C/li\u003E\u003C/ul\u003E\u003Cul\u003E\u003Cli\u003E29 days of PTO (18 working days per year pluse all national holidays)\u003C/li\u003E\u003C/ul\u003E\u003Cul\u003E\u003Cli\u003E10 paid recovery days\u003C/li\u003E\u003C/ul\u003E\u003Cul\u003E\u003Cli\u003EFull financial and legal support for independent contractors\u003C/li\u003E\u003C/ul\u003E\u003Cul\u003E\u003Cli\u003EFree English classes, with native speakers or Ukrainian teachers\u003C/li\u003E\u003C/ul\u003E\u003Cul\u003E\u003Cli\u003EDedicated HR support\u003C/li\u003E\u003C/ul\u003E\u003Ch3\u003EOur next steps\u003C/h3\u003E\u003Cp\u003E\u2705 Intro call with a Recruiter \u2014 \u2705 Intro call with a client \u2014 \u2705 Final client interview\u2014 \u2705 Offer\u003C/p\u003E","identifier":{"@type":"PropertyValue","name":"Gurify","value":"voice-ai-generative-music-part-time-at-mwdn-609b2bd3566c"},"url":"https://gurify.com/job/voice-ai-generative-music-part-time-at-mwdn-609b2bd3566c","datePosted":"2026-09-16","validThrough":"2026-10-31T23:59:59Z","hiringOrganization":{"@type":"Organization","name":"Mwdn","sameAs":"https://www.comeet.com/jobs/mwdn"},"directApply":false,"jobLocation":{"@type":"Place","address":{"@type":"PostalAddress","addressCountry":"UA","addressLocality":"Kyiv"}},"employmentType":"PART_TIME"}
```

```json
{"@context":"https://schema.org/","@type":"BreadcrumbList","itemListElement":[{"@type":"ListItem","position":1,"name":"Jobs","item":"https://gurify.com/jobs"},{"@type":"ListItem","position":2,"name":"Ukraine","item":"https://gurify.com/jobs/ukraine"},{"@type":"ListItem","position":3,"name":"Voice AI \u0026 Generative Music (Part-time)","item":"https://gurify.com/job/voice-ai-generative-music-part-time-at-mwdn-609b2bd3566c"}]}
```
