
Closed
Posted
I want to take a single recorded narration and transform it—through AI vocal conversion—into distinct, believable voices that fit a range of book characters. The goal is to breathe life into motivational and emotional works as well as conventional fiction and non-fiction titles, so each character sounds unique while remaining consistent across chapters. Here’s what I need from you: • An end-to-end workflow (model selection, training or fine-tuning, and post-processing) that converts one clean reference voice into multiple character voices suitable for long-form listening. • Output in standard audiobook formats (44.1 kHz WAV or 192 kbps MP3), ready for direct mastering or distribution. • Settings or presets I can reuse for new projects so I’m not locked into a single production cycle. Acceptance criteria: • At least four clearly differentiated voices delivered from the same base narration: two suited to fiction, one for non-fiction, and one adaptable for motivational or emotionally driven passages. • Consistent pronunciation, pacing, and volume within each generated voice. • Minimal artifacts—no robotic or metallic resonance when played through consumer headphones. If you already work with tools such as ElevenLabs, Voice Conversion GANs, or similar neural TTS engines, mention it briefly so I know we’re speaking the same language. Once the process is proven on one chapter, we can expand to entire books.
Project ID: 40688048
12 proposals
Remote project
Active 6 days ago
Set your budget and timeframe
Get paid for your work
Outline your proposal
It's free to sign up and bid on jobs
12 freelancers are bidding on average ₹988 INR/hour for this job

My dear prospective client, As a freelancer well-versed in numerous arenas, I pride myself on tackling diverse challenges head-on. While my expertise may not lie in neural TTS engines per se, my broad skill set in communication, content creation and email marketing shows my agility to quickly adapt and grasp new technologies. My arsenal includes proficiency with multiple audio editing software such as Adobe Audition, Pro Tools, and Audacity, which enhance my ability to deliver clean, crisp output devoid of glitches or robotic resonances. When it comes to your AI Audiobook Voice Conversion project, I understand the paramount importance of breathing life into each character and making them distinct while preserving consistency throughout all chapters. My extensive experience in video editing has taught me how to create different voices suitable for specific contexts and characters while keeping pronunciation, pacing, and volume consistent - skills that easily extend to an audio project like yours. Furthermore, I offer a unique proposition: rather than just delivering the final creation, I can also share with you detailed settings and presets for future projects—providing you with the autonomy to continue producing top-notch audiobooks even if our partnership ends. So allow me to leverage my considerable skills not only to meet your acceptance criteria but exceed your expectations at every step. Let's work together, dear client
₹1,000 INR in 40 days
3.7
3.7

Hi I can definitely do this job as I have 6+ years of experience in AI automation, AI integrations and audio processing workflows. I can build an end-to-end voice conversion workflow that takes your clean narration and creates multiple consistent character voices, including model selection, voice conversion, post-processing and reusable presets for future books. I’m familiar with AI voice tools like ElevenLabs and similar neural TTS/voice conversion solutions, and I can optimize the output for long-form listening with clean pronunciation, pacing, volume and minimal artifacts. I can also deliver the final voices in WAV or MP3 format and first test the workflow on one chapter before scaling it to the full book. Many thanks. Akif S
₹1,000 INR in 30 days
0.0
0.0

You recorded one voice. Your books need four that stay distinct, natural, and consistent from chapter to chapter. I can start right now. In 24 to 48 hours you get a live sample: your narration as four listening-ready voices, two fiction, one non-fiction, one motivational, as WAV or MP3. The hard part is no metallic edge on headphones and steady pace and volume on long chapters. I will save reusable presets so the next title is a repeat, not a rebuild. Share a short clean clip of the base narration so I can convert it first and you can hear the four characters?
₹850 INR in 3 days
0.0
0.0

Hello, I have 5+ years of experience in AI development, machine learning, audio processing and neural voice technologies, and I can build a reusable AI voice-conversion workflow for long-form audiobook production. I can work with ElevenLabs and similar neural TTS/voice-conversion solutions, selecting the approach based on voice quality, consistency, licensing and production requirements. I’ll create at least four distinct character voices from the same narration: two fiction voices, one professional non-fiction voice and one expressive motivational/emotional voice. The pipeline will include clean audio preprocessing, reference-voice conditioning, character/style controls, pronunciation and pacing consistency, loudness normalization, artifact reduction and final mastering. Deliverables will include 44.1 kHz WAV and 192 kbps MP3 outputs along with reusable voice presets and configuration. I’ll first validate everything on one chapter against the acceptance criteria, then document the complete workflow, model/settings, inference process and production steps. I’ll provide the source repository, installation guide and short walkthrough, with the architecture ready to scale across complete books. Best Regards, Shailender
₹760 INR in 48 days
0.0
0.0

Hi, I can build an end-to-end AI voice conversion workflow for audiobook production, creating distinct and consistent character voices from a single clean narration. My Approach 1. Voice Pipeline – Evaluate suitable tools/models such as ElevenLabs, neural voice conversion, or similar TTS/VC systems based on quality and control. 2. Character Voices – Create 4 differentiated voices: 2 fiction, 1 non-fiction, and 1 motivational/emotional style. 3. Consistency – Tune pronunciation, pacing, tone, loudness and character identity for long-form narration. 4. Post-Processing – Apply noise cleanup, EQ, compression and mastering for natural headphone playback. 5. Audiobook Output – Deliver 44.1 kHz WAV and/or 192 kbps MP3, ready for mastering/distribution. 6. Reusable Workflow – Provide presets, settings and documentation so the same process can be reused for future books. I’ll first prove the workflow on one chapter, refine the voices based on your feedback, and then make it scalable for full books. Let’s have a quick chat to discuss the sample narration, desired character styles, and preferred AI voice platform.
₹1,100 INR in 45 days
0.0
0.0

Hi there, One narration, four distinct characters, zero robotic artifacts on consumer headphones that's the real bar, and it's a mastering problem as much as an AI one. Voice conversion gets you 70% of the way; the remaining 30% is where most projects fail. My workflow: reference voice cleaned and normalized first noise floor, room tone, and mic inconsistencies removed before any conversion, since artifacts in the source only compound downstream. I'd use ElevenLabs' voice design and cloning pipeline as the core engine, layered with targeted fine-tuning per character to shift pitch, pacing, and tonal texture distinctly enough that a listener never confuses one voice for another, even across chapters recorded separately. Post-processing is where the metallic resonance problem actually gets solved spectral repair, de-essing, and consistent EQ profiling per voice, then a final normalization pass so volume and pacing stay stable within each character across the full runtime, not just the sample clip. Deliverables: four differentiated voices as specified two fiction, one non-fiction, one motivational/emotional register output in 44.1kHz WAV or 192kbps MP3, plus reusable presets so your next book doesn't restart the calibration process from zero. I'd propose proving this on one chapter first, exactly as you suggested, before scaling to a full book. Best, Rajat Trivedi
₹1,000 INR in 40 days
0.0
0.0

Hi — I build voice conversion pipelines you own outright, not a per-book vendor subscription. I can prove out one chapter fast so you judge quality before committing to the full book. RATE & TIMELINE — $10/hr, two phases: - Phase 1, proof chapter: ~5 hrs ($50). All 4 voices, one chapter, delivered with quality scores so you can judge before anything else is built. Ready in 3-4 days. - Phase 2, full audiobook: ~25-30 hrs ($250-300), once you approve Phase 1. Done in ~2 weeks. (Final numbers confirmed once I see your source audio.) QUALITY BAR — each voice ships only after clearing fixed thresholds: clarity scores that catch robotic artifacts, consistent loudness/pacing across chapters, and — for the motivational voice, the hardest — at least 80% of your original emotional range preserved. You get the numbers, not just my word. WHY SELF-HOSTED OVER ELEVENLABS — costs ~$2-5 compute per 10-hr book vs. ~$60-123 + monthly license, keeps your narration's original emotion/pacing (ElevenLabs re-generates from scratch), and you own the voices after — no vendor lock-in. If a voice fails a blind listening test on my setup, I fall back to ElevenLabs for that voice only. DELIVERABLES: 4 distinct voices (2 fiction, 1 non-fiction, 1 motivational), mastered WAV/MP3, reusable presets so your next book is a near-zero-cost re-run. Send over your source narration and I'll start on the proof chapter right away. I'm available to start immediately and happy to share prior VC samples on request.
₹1,500 INR in 40 days
0.0
0.0

You're looking for AI vocal conversion, and my demonstrated ability to deliver AI development through diligently comprehensive and top tier quality makes me the ideal candidate, just as I did for Grobrix Inc. I'm passionate about AI Model Development and bring self-management and proactive communication to ensure smooth execution. As a result, I'm more than just a task-doer; I'm a reliable team member.
₹800 INR in 40 days
0.0
0.0

Voice conversion from one narration into several character voices is very doable now. Two things decide whether it sounds believable or slightly wrong in a way listeners cannot name. The first is that conversion carries the original delivery. It changes timbre, not timing, so the pacing, pauses and emphasis of your narrator stay underneath every character. For fiction with dialogue that becomes noticeable, because two characters end up phrasing lines identically. The fix is practical: record the dialogue lines with the intended pacing per character, then convert. Same narrator, different performance, and the result stops sounding like one person wearing masks. The second is consistency across a whole book. A character has to sound the same in chapter one and chapter twenty, so the voice profiles get fixed and versioned up front rather than regenerated per session, and long files are processed in consistent chunks to avoid drift. On rights, worth confirming the voices are either synthetic or ones you have permission to use, since cloning a real narrator without consent is the one thing that causes trouble later. How long is the book, and roughly how many distinct character voices? Ronak
₹850 INR in 20 days
0.0
0.0

India
Member since Sep 3, 2026
₹1500-12500 INR
$10-30 USD
$15-25 USD / hour
₹600-1500 INR
₹600-1500 INR
₹750-1250 INR / hour
$10-30 USD
₹1500-12500 INR
£18-36 GBP / hour
$10-30 USD
₹750-1250 INR / hour
₹1500-12500 INR
$250-750 USD
₹1500-12500 INR
₹12500-37500 INR
£2-5 GBP / hour
$250-750 USD
min $50 USD / hour
$2-8 USD / hour
₹600-1500 INR