Articles
NVIDIA Releases TensorRT Model Connect in Public Preview: Two Commands from Hugging Face Checkpoint to NativeNEW⭐7
NVIDIA launches TensorRT Model Connect in public preview, converting Hugging Face checkpoints to native C++ inference in two commands with no ONNX step.
Groq Secures $350M to Accelerate Neocloud Strategy⭐9
Groq raises $350M led by Disruptive with Nvidia backing to expand its neocloud AI infrastructure, post-talent deal with Nvidia, aiming to scale data centers.
Nvidia Invests $1.5B in SoftBank-Linked Data Center Developer Behind OpenAI Project⭐7
Nvidia invests $1.5B in SoftBank-linked SB Energy, securing sole compute supplier role at OpenAI's Ports-Pike data center with up to $105B in credit support.
Nvidia's $500B Plan: A Risky Yet Brilliant Move for an Aging GPU Era⭐9
Nvidia's $500B AI plan guarantees GPU value, a risky yet strategic move to create a secondary market for aging chips with major implications.
Self-Supervised Layout Generation to Fix Advanced-Node DRVs (Nvidia, Duke)⭐8
Self-supervised ML from Nvidia and Duke automatically fixes advanced-node DRVs, cutting design rule violations and speeding up physical design closure.
NVIDIA Unveils Nemotron 3.5 Lightning: A 30B Open MoE with 3B Active Parameters and NeMo Switchyard Model Router⭐8
NVIDIA launches Nemotron 3.5 Lightning, a 30B open MoE with 3B active params, plus NeMo Switchyard router for faster, efficient AI agents.
General Catalyst Leads $1.1B Investment in Two-Month-Old AI Startup River AI⭐9
General Catalyst leads $1.1B round in two-month-old AI startup River AI, founded by xAI co-founder, to rebuild AI stack for personal assistants.
The Download: The Next Big Thing in LLMs and How AI Academic Research Is Shifting⭐8
New LLM startups tackle transformer limits and AI research shifts, exploring faster architectures and academic evolution.
NVIDIA Releases NemotronLabs VoiceChat 11B: An Open Full-Duplex⭐7
NVIDIA releases NemotronLabs VoiceChat 11B, an open full-duplex speech model with 448ms turn-taking, live tool calling, and barge-in—ideal for research pilots.
Building a Multimodal RAG Pipeline with NVIDIA NeMo Retriever,⭐8
Build a multimodal RAG pipeline using NVIDIA NeMo Retriever, hosted NIMs, LanceDB, reranking, and grounded generation. Step-by-step tutorial included.
