Nari Labs logo

Nari Labs

Nari Labs provides production-oriented multimodal inference services, initially focused on fast speech generation and transcription. Its Nari Qwen3-TTS endpoint offers streaming and non-streaming text-to-speech, while Nari Qwen3-ASR targets low-latency streaming speech recognition; the company also publishes open implementations and model-serving code for developers who want to run the stack themselves. The platform is aimed at voice-agent builders, realtime application teams and enterprises that need predictable latency without operating every optimization layer themselves. Official product pages describe free public beta endpoints, early-access pricing, managed and dedicated deployment, private infrastructure and fine-tuning options. The fresh Nari launch signal and official site/repository evidence make this a concrete tool listing rather than only a model announcement. Its combination of hosted APIs and open serving components is especially relevant for teams comparing speech infrastructure.

Reader rating

No ratings yet

Visit website

You might also like

Related tools

View all
Desert Ant Labs favicon
Desert Ant Labs
No ratings yet

Desert Ant Labs provides a family of small, specialized AI models and native SDKs designed to run on-device instead of through metered cloud inference. Its product line includes Voz for speech recognition, Clear for audio enhancement, Redact for privacy filtering, Clips for selecting video highlights, and other focused models that can be embedded into apps with Swift, Kotlin, or JavaScript. Voz is a concrete example: it transcribes 25 languages with word-level timestamps, runs on Apple platforms, and can process long audio locally with no login or per-token bill. The service is aimed at mobile and desktop developers building private, responsive AI features into their own products. Its official site offers free usage up to 100,000 monthly active devices per platform, while documentation and Hugging Face model pages provide installable evidence. This is an SDK/model platform, not merely a research announcement.

Demon favicon
Demon
No ratings yet

Demon (Diffusion Engine for Musical Orchestrated Noise) is an open-source real-time music generation system that runs locally on consumer GPUs at 25Hz. It is built for musicians, sound designers, music producers, and AI researchers who want to generate, iterate, and perform with AI music in real time without relying on cloud APIs. The system uses diffusion-based synthesis to produce musical audio streams with low latency, enabling live experimentation and performance workflows. Demon launched on Hacker News with 15 points and the project page at daydreamlive.github.io/DEMON describes a fully local, GPU-accelerated approach to music generation. What makes it notable is the combination of real-time performance with diffusion models — a technical achievement that opens up live music creation use cases that were previously impossible with slower batch-generation approaches.

Udio favicon
Udio
No ratings yet

Meet Udio, your AI-powered music creation companion. With Udio, you can effortlessly create and share music using cutting-edge AI technology. This free platform offers a range of tools to produce and refine audio content, from generating diverse music genres to creating vocals and instrumentals in seconds. Whether you're a music enthusiast, content creator, or just looking to add a unique touch to your projects, Udio's AI audio tools have you covered. From crafting melodies to experimenting with text-to-speech capabilities, Udio empowers you to explore the endless possibilities of AI-generated music and audio content. Unleash your creativity and dive into the world of AI music with Udio today!