Behind every great app is a great API. We power the real-time APIs for live voice, video, and chat in the world's biggest apps. (NASDAQ: $API)

Santa Clara, CA
Pinned Tweet
𝗜𝗻𝘁𝗿𝗼𝗱𝘂𝗰𝗶𝗻𝗴 𝗔𝗴𝗼𝗿𝗮 𝗔𝗴𝗲𝗻𝘁𝘀 𝗦𝗗𝗞. Building a voice agent is easy in a demo. Production is where the plumbing shows up: auth, RTC channels, STT/LLM/TTS config, session lifecycle, recovery. Agora Agents SDK helps developers handle that layer faster. agora.io/en/blog/agora-agent… #AIAgents #VoiceAI
5
6
26
8,772
Some cool things developers have been building with Agora lately 👀 ☕️ Spidey is a friendly AI companion that helps you manage tasks by voice. By @HaimantikaM . ⚔️ Voice Arena brings two voice agents with different turn-taking and interruption settings into a live battle. By @HarishKotra . 🎙️ Jev VAD explores when a voice agent should listen or keep talking. By @SoCalJayF . 🎧 An interactive podcast where you can jump in and talk to two AI hosts. By @zicojzc.
1
1
10
475
Take a closer look at these projects in Agora Dev Digest👇 medium.com/agora-io/agora-de…
40
Excited to welcome @soniox_ai to Agora Conversational AI platform! More ways for developers to build responsive, multilingual voice experiences 🤝
We’re excited to announce that @AgoraIO has selected Soniox as a speech-to-text provider for its real-time Conversational AI platform. For over a decade, Agora has built the global network infrastructure that connects people through real-time voice, video, and messaging. Today, that same carrier-grade infrastructure is powering a new era of interactions: humans talking directly with AI. For voice AI, moving audio in real time is only part of the problem. The system must understand natural, conversational speech instantly—accurately capturing critical details despite accents and noisy environments. Soniox brings: • Native-speaker accuracy across 60+ languages • Ultra-low latency for real-time conversations • Seamless multilingual speech and mid-sentence language switching • High precision on names, numbers, IDs, and other critical details • Reliable performance on real-world audio at production scale Agora provides the orchestration for Conversational AI and the infrastructure for real-time communications. Soniox provides the intelligence for understanding spoken language. Together, we’re making it easier for developers to build voice AI that can communicate naturally with people around the world. Excited to partner with the Agora team and help power the next generation of conversational AI.
1
5
445
Getting Started with Agora Conversational AI x.lingyaoai.com/i/broadcasts/1pKdRDvrb…
3
2
326
If you're curious to see real applications of AI voice tech in action👇 Moderating a killer panel happening in London next week for Deep Tech Week!!!! Would love to see you there 🚀 Space is limited & sharing more deets below: luma.com/cjumnkfs
1
1
6
791
🎙️ Announcing the Agora Conversational AI LinkedIn Live series, starting this October. Grab your seat: linkedin.com/events/75104485… 𝗙𝗶𝗿𝘀𝘁 𝘀𝗲𝘀𝘀𝗶𝗼𝗻: 𝗚𝗲𝘁𝘁𝗶𝗻𝗴 𝗦𝘁𝗮𝗿𝘁𝗲𝗱 𝘄𝗶𝘁𝗵 𝗔𝗴𝗼𝗿𝗮 𝗖𝗼𝗻𝘃𝗲𝗿𝘀𝗮𝘁𝗶𝗼𝗻𝗮𝗹 𝗔𝗜 We'll go from "What exactly does Agora do in a Voice AI stack?" to getting something up and running. 📅 October 1 ⏱️ 30–40 mins 📍 LinkedIn Live No slides-only webinar. We'll actually open the tools and build.
1
2
8
766
Giving people time to finish their thoughts 💬 @SoCalJayF connected @typesafeai Jev to Agora ConvoAI to explore how live transcripts and conversation context can help voice AI recognize when someone has finished speaking. Take a look at his demo 👇
One really cool use case I’ve been exploring with Jev: a semantic VAD for real-time voice AI. I plugged it into @AgoraIO ConvoAI and used the live transcript + recent conversation as context to decide whether the user has actually finished speaking. That makes Jev surprisingly good at handling things like hesitation, unfinished thoughts, and those moments where you stop talking for a second because you’re still figuring out what you want to say. It’s a nice example of where Jev fits really well in a real-time AI loop. Here’s what it looks like in action 👇
1
8
825
A live podcast platform with AI hosts – built by @zicojzc using Gemini 3.8 Flash TTS, Gemini 3.5 Transcribe Live, and Agora RTC! It feels less like listening to an episode and more like stepping into the conversation: 🎙️ Everyone in the room hears the same show, and listeners can jump in and talk to the hosts 🫸 The joiner interrupts, and the AI hosts acknowledge the interruption immediately 🗣️ Gemini 3.8 Flash TTS streams a two-host answer, and then the podcast picks up where it left off We're stepping into a new era of real-time AI communication. What do you think @googledevs @GoogleDeepMind?
4
13
1,066
Tone and delivery shape how spoken communication is understood, and Google’s latest Gemini 3.8 Flash TTS and Flash-Lite TTS models address a persistent limitation in voice AI by giving developers finer control over how generated speech sounds and adapts throughout a conversation. We’re excited about the new models’ ability to direct emotion, pace, character, and accent – plus a library of 20,000+ voices! Our thoughts on what this opens up for real-time voice builders: agora.io/en/blog/a-new-gener…
Create and deploy custom audio with our new text-to-speech models: 🔵 Gemini 3.8 Flash TTS: Design unique voices with distinct accents and characteristics. 🔵 Gemini 3.8 Flash-Lite TTS: Built for efficiency and scale, choose from your created styles or our expansive production-ready library.
4
4
11
1,421
Your AI agent gets an empty response from a tool—then makes up the data it needs to keep going. What catches it? In the latest episode of Convo AI World podcast, host @hermes_f speaks with Dustin Allen from Trinitite about governing AI agents in production. Watch the episode 👉 podcast.convoai.world/episod…
1
6
706
Building voice AI with Gemini? Join us live on Discord with @thorwebdev from @GoogleDeepMind and @hermes_f from Agora. We’ll cover the latest real-time capabilities, natural turn-taking, interruptions, tool use, evals, and what it takes to ship dependable voice agents. 📅 Thursday, Sep 17 at 9 AM PT / 12 PM ET Link: bit.ly/4xAzPDQ
2
11
32
6,144
When it comes to voice-agent workloads, one size doesn’t fit all–and that’s what Google Gemini’s 3.8 Live update addresses. ⚡️ 3.8 Live for fast, direct interactions such as FAQs, routing, order status, and scheduling. 🧠 3.8 Live Extended Thinking for investigation, planning, and workflows spanning multiple tools. We looked at where each model fits in production voice AI: agora.io/en/blog/gemini-3-8-…
🗣️ Introducing Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking. These advanced audio models are built for natural conversation, featuring major upgrades in turn-taking and near real-time reasoning. They also significantly streamline how you build intelligent voice agents. Until now, advanced reasoning and reliability came at the cost of high latency and the complexity of cascaded, multi-model pipelines. Stitching together separate speech-to-text, reasoning, and text-to-speech models adds delay and inflates costs. By replacing that stack with a single multimodal API call, these two models make building highly responsive voice agents simpler and cost-effective. 🧵 Here’s what you can build with them:
1
1
8
6,740
Our dev team built a meeting demo using Agora RTC, Agora Conversational AI and GPT-Live-1 – and gave GPT‑Live‑1 a seat in our meeting. GPT‑Live‑1 listened to our team’s discussion, summarized it, and even added a task to the Kanban board.
4
4
14
1,513
Agora retweeted
Just gave GPT‑Live‑1 a seat in our meeting today. It felt so natural! It listened to our team’s discussion, summarized it, and added a task to the Kanban board. First time actually having a meeting with an AI. Pretty fun. Powered by @AgoraIO RTC, Conversational AI and @OpenAI GPT‑Live‑1 API
3
6
26
25,385
We’re excited to see this model’s full-duplex capabilities move the industry closer to human-to-AI conversations that flow more like human-to-human ones – much like our mission to make real-time video, voice and conversational AI ubiquitous. Our take: agora.io/en/blog/voice-ai-th…
GPT-Live-1 is now available in the API. Bring ChatGPT’s natural back-and-forth to your app, with voice agents that listen while they speak and work with the models and harness you choose.
2
4
7
808
AI Toys Come Alive! Real-Time Voice AI × Interactive Entertainment What happens when toys can talk, listen, remember, and respond naturally in real time? @AgoraIO, in collaboration with @awscloud, is bringing AI builders, robotics companies, voice AI innovators, and interactive entertainment creators together in Tokyo to explore what’s next for AI toys, AI companions, virtual characters, conversational AI, and connected devices. Featuring @kotoba_tech, INBREEZE, AnotherBall, @yukaikk, and @MiniMax_AI, plus live demos and networking. 📷 September 17 | 17:00–21:00 If you’re building the future of AI, robotics, gaming, toys, voice AI, or interactive entertainment, we’d love to have you join us! Register: luma.com/nts51f1c
1
2
508
AI makes a guess, then has to generate a video to prove it 😄 Built with Agora RTC, Interactive Whiteboard, and @reactorworld FastH3. Definitely more fun with friends!
Made a realtime Draw & Guess game with FastH3. AI joins as a player, it sees the sketch, makes a guess, then uses FastH3 to generate a matching video as proof. AI only wins if both the guess and the video are right. Played it with my friends, it got pretty funny. Built with @AgoraIO RTC + Interactive Whiteboard and @reactorworld’s FastH3. I’ll publish the demo and open-source it soon
1
7
782
Want to use Gemini 3.5 Transcribe Live in a voice agent without committing the rest of your stack to a single provider? Agora’s new TypeScript integration lets you add Gemini as the speech-to-text layer while retaining control over the LLM and voice. Using 𝘎𝘦𝘮𝘪𝘯𝘪𝘛𝘳𝘢𝘯𝘴𝘤𝘳𝘪𝘣𝘦𝘚𝘛𝘛 in the Agora Agents SDK, developers can: 🔌 Pair Gemini transcription with an OpenAI-compatible LLM and providers 🌍 Handle multilingual or code-switching conversations 📚 Add custom vocabulary and other domain-specific language 📡 Use Agora RTC for live audio while RTM delivers transcripts, agent state, metrics, and errors to the application 🔐 Keep the Google API key and Agora App Certificate safely on the server Start building with Gemini 3.5 Transcribe Live and Agora Conversational AI 👉 agora.io/en/blog/using-gemin…
4
2
10
12,875
@AgoraIO Joins @tib_tokyo as a Partner We're excited to announce that Agora has joined Tokyo Innovation Base (TIB) as a TIB Partner. Operated by the Tokyo Metropolitan Government, TIB brings together startups, entrepreneurs, corporations, investors, and ecosystem partners to accelerate innovation across Japan. Through this partnership, we're excited to contribute our expertise in real-time engagement infrastructure, Voice AI, and Conversational IoT—helping developers build immersive live experiences and natural conversational interactions. As we continue expanding in Japan, we look forward to collaborating with startups and innovators to shape the future of real-time experiences and conversational AI. #Agora #TokyoInnovationBase #VoiceAI #Japan
2
1
5
730
ConvoAI World Podcast: Dustin Allen, Co-founder @ Trinitite x.lingyaoai.com/i/broadcasts/1mxPaZZNy…
1
4
1,483