Creative 4 events
Category RSS ↗Yesterday
-
Limited sourcesCreative
Launch HN: Speko (YC S26) – OpenRouter for Voice AI
Hi HN! I'm Bek, founder of Speko, a platform that finds an optimal combination of speech-to-text, LLM, and text-to-speech models, given your constraints, among all our public benchmarked options, and…
Hacker News Frontpage Aug 10 · Mon
-
ConfirmedOpen SourceAgents & Dev
Build Low-Latency Multilingual Voice Agents: Open Weights & Full Deployment Control with NVIDIA Magpie TTS
Every voice interaction has a latency budget. By the time a user hears your application respond, you've already spent precious milliseconds capturing audio, transcribing speech, running an LLM, retri…
Hugging Face Blog Aug 3 · Mon
-
ConfirmedCreative
How we built a realtime system for responsive voice AI in six months
GPT-Live enables continuous voice interaction with AI, using a turnless speech model and low-latency architecture for faster, more natural conversations.
OpenAI News Jul 23 · Thu
-
ConfirmedCreative
Bringing Nunchaku 4-bit Diffusion Inference to Diffusers
Large diffusion transformers can create stunning images (or even videos, audio snippets, and now text), but loading a modern text-to-image model in BF16 precision often requires 20-30 GB of VRAM, whi…
Hugging Face Blog