Today, we’re launching Voice Design.
Write a prompt, create a voice.
Describe the accent, age, gender, and pace your use case needs, and get new voices in seconds, ready to use.
Live and free in the Gradium API and Studio. gradium.link/voice-design
Welcome to our weekly community spotlight, featuring the top three teams from the AI Gaming Hack in Paris, all built with Gradium.
Here’s how each team used Gradium in their winning game 📷
Explore their projects: gradium.link/project-gallery
2nd Place - Last Alibi
Investigate a murder aboard a snowbound 1931 train by questioning suspects aloud.
Gradium Speech-to-Text transcribes your questions, and Text-to-Speech gives the suspects their voices.
last-alibi.kaisspace.workers…
3rd Place - AfterFlow
Ferry souls through the underworld and read scroll incantations aloud to calm the gods.
The team used Gradium Speech-to-Text for incantations and Text-to-Speech to create the gods’ voice lines.
afterflow.stream
Our default Text-to-Speech model now has around 50ms time to first audio.
The model scores highest on naturalness for any sub-100ms model on Speko public benchmarks.
How much of your turn budget does TTS take?
gradium.link/fastest-tts
Thank you to all who joined us at the AI Gaming Hackathon in Paris.
Check out the projects at gradium.link/project-gallery
See you at the next hackathon.
Our community page is live.
See apps built with Gradium, share yours in the showcase, and talk to our team on Discord.
Hear about upcoming hackathons and events.
Selected projects will be featured in the gallery and qualify for bonus credits.
gradium.ai/community
Join us in Paris and Zurich for the upcoming AI Gaming and Agentic AI Hack.
Every participant gets free Gradium credits to add voice to what they build. Registration links are below.
luma.com/par-hack
We measured Gradium TTS Beta on our TTS board: 49ms p50, 56ms p90 time to first audio from US East.
That is more than 4x faster than the previous Gradium TTS (217ms p50), with the same naturalness.
Full results: benchmarks.speko.ai/tts
Gradium TTS beta model is out now to define what fast means.
Less than 50ms TTFA, same quality. Most models are >100ms TTFA.
Try it now: studio.gradium.ai/?model_nam…
Gradium TTS beta model is out now to define what fast means.
Less than 50ms TTFA, same quality. Most models are >100ms TTFA.
Try it now: studio.gradium.ai/?model_nam…
H3 Max Lip Sync is now live on fal.
Upload one photo and one audio clip and get back a lip-synced video in seconds, with natural expression and movement, in any language.
In our evals, H3 Max Lip Sync ranks #1 for both quality and speed, with a median generation time of just 11 seconds.
Try it now on fal.