
In Progress
Posted
I'm seeking an expert in Gemini TTS to create engaging and casual voiceovers for museum audio guide clips. Requirements: - Use our proprietary audio guide production platform with its Gemini TTS API to generate high-quality, engaging, and casual voiceovers. - Create audio guide clips for museum exhibits. Ideal Skills and Experience: - Experience with Gemini TTS or similar platforms. - Strong understanding of voice modulation and tone for casual engagement. - Background in audio production is a plus. - Ability to meet deadlines and deliver high-quality audio files. Please provide samples of previous work and estimated delivery time.
Project ID: 40562220
12 proposals
Remote project
Active 5 days ago
Set your budget and timeframe
Get paid for your work
Outline your proposal
It's free to sign up and bid on jobs

Hello, I’d be happy to help with your museum audio guide project. I have hands-on experience working with ElevenLabs and creating natural, engaging AI voiceovers. I can produce speech with the tone, pacing, and voice style you’re looking for, making the narration feel warm, conversational, and enjoyable for museum visitors. One of my AI voice projects is already available in my portfolio. If you’d like to hear more examples, I’d be happy to send additional samples that match your requirements. As an audio engineer, I don’t just generate the voice—I also ensure the final audio is clean, balanced, and polished for a professional listening experience. My estimated delivery time is within 1 day, or even sooner depending on the number of clips and the overall project size. I also prefer an interactive workflow, so we can review the first samples together and make any adjustments to the voice style or tone before completing the full project. I look forward to working with you! Best regards, Pouya Azadi
$20 USD in 7 days
0.0
0.0
12 freelancers are bidding on average $23 USD/hour for this job

Hi, — this is a narrow TTS production problem, and the important part is not just generating audio through Gemini TTS but getting repeatable, natural reads across many exhibit clips. The real engineering risk is consistency: tone, pacing, pronunciation, and revision control tend to drift when output is produced inside a proprietary workflow. I’ve built production AI pipelines that combine model output with real-world media handling, including TTS-driven systems where audio quality and routing had to hold up in live use. For your case, I’d treat this as a controlled generation workflow rather than simple text-to-audio conversion. The closest match is TikTok AI Livestream Setup, where I configured an end-to-end TTS pipeline with synchronized voice output and low-latency audio behavior. Custom Feature Development & Integration is also relevant because this sounds like work inside an existing platform with its own constraints and review flow. I usually structure this by separating script preparation, voice generation, and quality review so pronunciation fixes and tone adjustments are easy to apply without reworking the whole batch. I’d also look closely at how your platform handles retries, versioning, and clip-level approvals. If useful, I can review one sample exhibit script and sketch the generation and QA flow before broader production. Thanks, Hercules
$50 USD in 40 days
3.4
3.4

Hi, I'm a great fit for this project. I regularly work with Gemini TTS and AI voice production, but for the most natural, engaging museum audio guides, I highly recommend ElevenLabs, as it consistently delivers more lifelike narration, better emotional expression, and superior pronunciation. I can generate professional, casual voiceovers, refine every clip, and ensure they sound polished and ready for your platform. I also have relevant samples that I can share via direct message and can deliver within your required timeline. Let's create museum guides that truly engage your visitors.
$20 USD in 40 days
5.3
5.3

Hello, Drawing from my expertise in AI and Automation, I am uniquely positioned to offer you top-notch voiceover services for your museum audio guide clips. Throughout my decade-long career, I've developed robust systems and solutions that have significantly reduced costs, streamlined processes, and scaled businesses seamlessly. With a strong background in digital marketing and a deep understanding of AI-driven systems, I not only possess the proficiency in Gemini TTS but also have an ear for engaging and casual voice modulation- which would be a key requirement for your project. Moreover, the scope of my skills go beyond just AI Voice Assistance. My exposure to n8n automation would come in handy as it emphasizes on meeting project deadlines without sacrificing quality. And since audio production is something I’m respective about, you could trust that the audio files delivered would meet your standards. To further reinforce my capacity for the project at hand, please find enclosed samples of previous voiceover work for your reference. Rest assured, if selected, not only will I complete the task skillfully but also with creativity and class that upholds the reputation of your museum exhibitions. Let's get started on making those audio guide clips captivating! Thanks!
$15 USD in 26 days
1.2
1.2

Hi, I have experience working with AI voice technologies and TTS APIs to create natural, engaging voiceovers. I can use your Gemini TTS production platform to generate high-quality museum audio guides with a warm, conversational tone while ensuring consistency across all exhibits. I focus on clear pronunciation, natural pacing, and timely delivery, and I am available to begin immediately. Best regards Muhammad
$20 USD in 40 days
0.5
0.5

Hi, this is Kris from McKinney, Texas, I’ve reviewed your requirement for generating museum audio guide voiceovers using Gemini TTS through your proprietary production platform. The core objective is to produce natural, engaging, and lightly conversational narration that feels informative without sounding robotic or overly formal, while maintaining consistency across multiple exhibit clips. My approach would be to integrate directly with your Gemini TTS API pipeline and tune voice outputs using structured prompt engineering for tone, pacing, and emotional neutrality appropriate for museum environments. I would focus on refining prosody control (pause placement, emphasis, and cadence) to ensure each exhibit narration feels immersive and easy to follow. If needed, I can also implement a lightweight pre-processing layer to standardize scripts before TTS generation so output quality remains consistent across all clips. Final audio would be quality-checked for clarity, pacing, and tonal consistency before delivery in your required format. A few additional questions; Q1: Do you already have fixed scripts for each exhibit, or should part of the workflow include script refinement for better TTS output? Q2: How many voice profiles or tones do you want available (single narrator vs multiple characters/styles)? Q3: Do you require post-processing (noise shaping, EQ, normalization), or is raw TTS output sufficient? Regards, Kris
$25 USD in 40 days
0.0
0.0

❤️❤️❤️ Hello there! ❤️❤️❤️ I'm Elena, a full-stack developer with 9+ years of experience in Web, Mobile, and AI. I went through your project and it sounds like something I'd genuinely enjoy working on. If we decide to work together, you can expect honest communication, attention to detail, and someone who cares about getting the result right rather than just finishing the job. I'd love to hear more about your goals and answer any questions you may have. Warm regards, Elena
$20 USD in 40 days
0.0
0.0

Hello, I am an experienced AI/LLM Analyst, Prompt Engineer, and QA Specialist with hands-on experience in AI data annotation, prompt engineering, prompt evaluation, AI response rating, model evaluation, image annotation, content review, instruction following, and Text-to-Speech (TTS) projects. I focus on delivering accurate, high-quality work while meeting deadlines. I am detail-oriented, quick to learn new workflows, and committed to providing reliable results. I would be happy to discuss your project and start immediately.
$20 USD in 25 days
0.0
0.0

With my extensive background in software and AI development, specializing in text-to-speech (TTS) systems like Gemini, I bring a unique skill set to your project. Over the past two decades, I've stayed ahead of the curve, exploring various platforms from J2ME to SwiftUI to Flutter. My work within the industry ranges from casual gaming to sports analytics highlighting my adaptability and capacity to engage users in a variety of contexts. Having already worked with TTS platforms like Gemini, I intimately understand how to leverage this technology to create not just high-quality audio but engaging and casual voiceovers. My understanding of voice modulation and tone is a product of years working on various audio projects that benefited from these skills. In addition to these technical abilities, I also recognize that meeting timelines and delivering high-quality products are essential for any successful project, especially in immersive educational experiences like museum audio guides. And throughout it all, I'm not just your developer; I'm ready to be your strategic partner for ongoing project success. Stack your museum audio guide project alongside the many others I've successfully completed; together we take it to new auditory heights.
$20 USD in 40 days
0.0
0.0

Hi, I have experience in Text to Speech. I already have experience for about 150 dataset for Text to Speech for AI, and 30+ Voice acting for AI training. I have Condenser Mic that have an above good quality
$20 USD in 20 days
0.0
0.0

Stamford, United States
Member since Jul 5, 2026
$2-30 USD / hour
$10-30 USD
$30-250 USD
₹1500-12500 INR
₹1500-12500 INR
$15-25 USD / hour
₹600-1500 INR
$2-30 USD / hour
₹75000-150000 INR
$3-10 NZD / hour
$250-750 USD
$10-30 USD
₹600-1500 INR
₹12500-37500 INR
₹12500-37500 INR
₹600-1500 INR
₹1250-2500 INR / hour
₹600-1500 INR
₹1250-2500 INR / hour
$250-750 CAD