NVIDIA Company 4 related events
Today
-
ConfirmedOpen Source
llama.cpp b10481: CUDA: MMVQ nwarps=8 for bs=1 for dense models on DGX Spark (#26843)
CUDA: MMVQ nwarps=8 for bs=1 for dense models on DGX Spark Signed-off-by: ynankani ynankani@nvidia.com skip moe experts and allow others based on k geometry (allow only small idle tail) Signed-off-by…
GitHub Releases · ggml-org/llama.cpp Yesterday
-
Limited sourcesPolicy & SafetyInfra
Nvidia investing $1.5B in SoftBank data center developer behind OpenAI project
Tim De Chant Nvidia said on Monday that it will invest $1.5 billion in SB Energy, a data center linked to SoftBank and OpenAI. The investment ensures that Nvidia will be the sole supplier of compute…
TechCrunch AI Aug 10 · Mon
-
ConfirmedOpen SourceAgents & Dev
Build Low-Latency Multilingual Voice Agents: Open Weights & Full Deployment Control with NVIDIA Magpie TTS
Every voice interaction has a latency budget. By the time a user hears your application respond, you've already spent precious milliseconds capturing audio, transcribing speech, running an LLM, retri…
Hugging Face Blog Jul 27 · Mon
-
Confirmed
NVIDIA Cosmos-H-Dreams: Bringing Real-Time Generative Simulation to Surgical Robotics
Surgical robotics is moving quickly from teleoperation toward increasingly capable vision-language-action policies. But evaluating and training these systems remains difficult. Physical robotic platf…
Hugging Face Blog