# inworld.ai > AI-optimized mirror of inworld.ai containing 50 pages totalling 29,416 words of clean markdown content, structured data, and semantic HTML. Original source: https://inworld.ai/. Last updated: 2026-05-30T12:21:13.762Z. Each page is available as HTML (with JSON-LD structured data) and Markdown (text-only, ideal for LLMs and RAG). ## Homepage - [The #1 ranked](/site-root.html): #1 ranked TTS with under 200ms latency, voice cloning, and 75% lower cost. Realtime agents built for scale. (1,625 words) ## Articles & Blog Posts - [blog/introducing-inworld-tts/index.html](/blog/introducing-inworld-tts/index.html) (1 words) - [models/openai-gpt-5-5/index.html](/models/openai-gpt-5-5/index.html) (1 words) - [resources/best-realtime-ai-api/index.html](/resources/best-realtime-ai-api/index.html) (1 words) - [models/google-ai-studio-gemini-3-1-pro-preview-customtools.html](/models/google-ai-studio-gemini-3-1-pro-preview-customtools.html) (1 words) - [resources/best-voice-ai-infrastructure-platform/index.html](/resources/best-voice-ai-infrastructure-platform/index.html) (1 words) - [resources/best-voice-ai-tts-apis-for-real-time-voice-agents-2026-benchmarks.html](/resources/best-voice-ai-tts-apis-for-real-time-voice-agents-2026-benchmarks.html) (1 words) - [resources/what-is-conversational-ai/index.html](/resources/what-is-conversational-ai/index.html) (1 words) - [resources/best-voice-agent-platforms/index.html](/resources/best-voice-agent-platforms/index.html) (1 words) - [resources/best-litellm-alternatives/index.html](/resources/best-litellm-alternatives/index.html) (1 words) - [terms/index.html](/terms/index.html) (1 words) - [resources/stt-voice-profiling-api/index.html](/resources/stt-voice-profiling-api/index.html) (1 words) - [resources/how-to-build-an-ai-voice-agent/index.html](/resources/how-to-build-an-ai-voice-agent/index.html) (1 words) - [resources/best-text-to-speech-apis/index.html](/resources/best-text-to-speech-apis/index.html) (1 words) - [resources/best-llm-router-ai-gateway/index.html](/resources/best-llm-router-ai-gateway/index.html) (1 words) - [blog/wishroll-status-cutting-ai-costs-by-95-percent/index.html](/blog/wishroll-status-cutting-ai-costs-by-95-percent/index.html) (1 words) - [resources/what-is-consumer-ai-infrastructure/index.html](/resources/what-is-consumer-ai-infrastructure/index.html) (1 words) - [models/deepseek-deepseek-v4-flash/index.html](/models/deepseek-deepseek-v4-flash/index.html) (1 words) - [blog/how-inworld-helped-the-ai-game-death-by-ai-with-20-million-players-reach-profitability.html](/blog/how-inworld-helped-the-ai-game-death-by-ai-with-20-million-players-reach-profitability.html) (1 words) - [resources/vapi-vs-pipecat-vs-livekit/index.html](/resources/vapi-vs-pipecat-vs-livekit/index.html) (1 words) - [Realtime TTS-2](/blog/realtime-tts-2/index.html): Make every user feel understood with the top-ranked text-to-speech, speech-to-speech and LLM routing. Production-grade APIs built for developers. (3,198 words) - [The new AI infrastructure for scaling games, media, and characters](/blog/new-ai-infrastructure-scaling-games-media-characters.html): Build scalable AI experiences and characters for games and media with Inworld Runtime. The AI infrastructure trusted by Xbox, Ubisoft, & NVIDIA. (947 words) - [Engage every user with the #1 ranked, _most natural_ Voice AI](/tts/index.html): Leading voice AI with sub-200ms latency, instant voice cloning, emotion and non-verbal controls, multilingual support, and radically accessible pricing. (2,542 words) - [Best Self-Hosted TTS: Open Source vs. On-Premise Voice AI (2026)](/resources/best-self-hosted-tts/index.html): Compare the best self-hosted text-to-speech options in 2026. Open-source models (Kokoro, Chatterbox-Turbo, Piper, Dia2, Fish Audio) and commercial on-premise (Realtime TTS, ElevenLabs Enterprise) ranked on quality, latency, and engineering ownership. (936 words) - [The 3 Engineering Challenges of Realtime Conversational AI](/blog/three-challenges-of-realtime-conversational-ai/index.html): Built from our learnings in latency and iteration speed, Inworld Runtime is the C++ engine for realtime AI agents. SDKs for Node.js, Unity, & Unreal. (1,009 words) - [Inworld AI Privacy Notice](/privacy/index.html): Privacy Policy (3,003 words) - [Talk to our team](/contact-sales/index.html): Get started with state-of-the-art TTS and conversational AI pipelines. Talk to our team about scaling to millions of users, custom solutions, and pricing. (172 words) - [Inworld TTS-1.5: Upgrading the #1 Ranked TTS Model with Production-Grade Latency, Expression and Stability](/blog/introducing-inworld-tts-1-5/index.html): TTS-1.5 Max hits P90 under 250ms. Mini under 130ms. 40% fewer word errors, 30% more expressive, 15 languages including Hindi. Built for production scale. (995 words) - [NVIDIA B200 GPU: Specs, Pricing, and Cloud Availability (2026)](/resources/nvidia-b200-gpu-cloud/index.html): NVIDIA B200 GPU specs, cloud rental pricing from 8 providers, availability status, and how B200 compares to H100 and A100 for LLM inference in 2026. (1,829 words) - [Build with confidence on secure AI infrastructure](/security/index.html): Enterprise-grade AI security and compliance. Built on a zero-trust framework with continuous monitoring, providing a secure foundation for teams (712 words) - [Grok 4.1 Fast](/models/xai-grok-4-1-fast/index.html): Grok 4.1 Fast API pricing: $1.25/1M input, $2.5/1M output. 2M context. Supports function calling, web search, reasoning. Access via Inworld Router with OpenAI SDK compatibility and failover. (564 words) - [Inworld + LiveKit: Unlocking studio-quality voice AI for real-time experiences at scale](/blog/unlocking-studio-quality-voice-ai/index.html): Create multiplayer games and AI agents with Inworld + LiveKit. Real-time voice AI, global edge infrastructure, accessible pricing. Build today. (648 words) - [GPT 5 Nano](/models/openai-gpt-5-nano/index.html): GPT 5 Nano API pricing: $0.05/1M input, $0.4/1M output. 272K context. Supports function calling, web search, reasoning. Access via Inworld Router with OpenAI SDK compatibility and failover. (557 words) - [Inworld meets Pipecat: Raising the bar for realtime voice AI](/blog/raising-bar-for-realtime-voice-ai-with-pipecat/index.html): Studio-quality voice AI at $15/1M characters (TTS 1.5-Mini), 75% cheaper than other providers. Real-time TTS in your Pipecat pipeline. (652 words) - [Staff / Principal Software Engineer - Canada](/careers/c9b39e18-0d33-4a2d-94f4-1b6b6abee429/index.html): Staff / Principal Software Engineer - Canada | Open position at Inworld (804 words) - [Your AI is boring: Boost engagement and drive immediate performance improvements with Inworld Runtime and custom Mistral AI models](/blog/inworld-mistral-collaboration/index.html): Partner with Mistral AI and Inworld to build consumer AI apps users love. Custom models, automated infrastructure, and instant A/B testing. Get started. (879 words) - [The complete guide to measuring and optimizing TTS latency](/blog/how-to-measure-tts-latency/index.html): Learn how to measure TTS latency accurately across providers. Understand how WebSocket, streaming, and audio post-processing affect latency, and try our interactive demo to benchmark latency across providers. (708 words) - [Staff / Principal Software Engineer - USA](/careers/bc788f3f-bab1-4f99-a068-22f59e818334/index.html): Staff / Principal Software Engineer - USA | Open position at Inworld (834 words) - [ElevenLabs v3 Is Now GA. Here's What Developers Should Know.](/resources/elevenlabs-v3-review/index.html): ElevenLabs v3 is best for pre-rendered audio, not real-time. Top alternatives for voice agents and conversational AI that need sub-200ms latency. (1,048 words) - [The Next Wave of AI Applications](/blog/the-next-wave-of-ai-applications/index.html): Make every user feel understood with the top-ranked text-to-speech, speech-to-speech and LLM routing. Production-grade APIs built for developers. (422 words) - [Talkpal AI scales to 5 million language learners with Realtime TTS](/blog/talkpal-ai-scales-to-5-million-language-learners-with-inworld-tts.html): How Talkpal cut TTS costs by 40% while improving retention and free-to-paid conversion (579 words) - [Best Voice-to-Text API for Developers (2026)](/resources/best-voice-to-text-api/index.html): Compare the leading voice-to-text APIs in 2026. Realtime STT, Deepgram, AssemblyAI, OpenAI Whisper, Speechmatics, and the cloud providers ranked on streaming latency, language coverage, and pipeline integration. (1,082 words) - [Unreal AI Runtime: The first unified interactive AI toolkit for game developers](/blog/introducing-unreal-ai-runtime-sdk/index.html): The first unified toolkit for building realtime interactive AI experiences in Unreal. With 100+ models, built-in observability, and an intuitive visual editor. (968 words) - [Realtime voice agents can now see, listen, and engage](/blog/inworld-tts-2-stream-vision-agents/index.html): Inworld TTS-2 + Stream Vision Agents build AI voice agents that read emotions, respond in real time — sub-200ms latency, 100+ languages, full context. (530 words) - [Build faster, smarter realtime agents - instant streaming, lower latency, and smart interruption handling](/blog/runtime-v0-8-faster-realtime-agents/index.html): Node.js Runtime SDK v0.8: Faster LLM/TTS streaming performance, production‑ready templates, and full trace and log visibility on the overview tab (337 words) - [Solve the way to evolve.](/careers/index.html): Solve the way software evolves. We build living AI systems that autonomously adapt to serve users. Technical depth is a requirement across every role. (266 words) - [Pricing that _scales with you_](/pricing/index.html): Flexible pricing for text-to-speech, speech-to-text, realtime voice agents, and LLM routing. Start free, scale with volume discounts. No hidden fees. (1,102 words) - [Something went wrong](/realtime-api/index.html) (13 words) - [Blog](/blog/index.html): Latest insights on realtime conversational AI, TTS, and runtime pipelines. Case studies, technical deep-dives, and product updates from Inworld. (265 words) - [Customer stories](/customers/index.html): How leading teams ship voice AI in production: real-time agents, games, language learning, and more. (171 words) ## Resources - [Full Page Index](/index.html): Browse all cached pages with rich metadata - [About This Cache](/content/about.html): Methodology, technical details, and usage guidelines - [XML Sitemap](/content/sitemap.xml): Machine-readable sitemap for crawler discovery - [Robots.txt](/content/robots.txt): Crawler directives