NVIDIA AI AI News, Models and Product Updates
Track latest Nvidia AI AI news, launches, research, and ecosystem moves.
Nvidia Ai news, model releases, product launches, research updates, and major announcements in one place.
Latest Articles
- Heart of the Matter: How a Major Children’s Hospital Uses Open Source NVIDIA AI for Cardiac Care — Children's Hospital of Philadelphia is using open-source NVIDIA AI tools to generate 3D models of pediatric hearts in seconds, aiming to improve pre-surgical planning for children with congenital heart disease.
- mKernel: Fast Multi-GPU, Multi-Node Fused Kernels — mKernel achieves up to 1.88x speedup on Ring Attention across 16-GPU H200 clusters by fusing computation with NVLink and RDMA communication at tile granularity — targeting the communication bottleneck that limits distributed LLM training at scale.
- GGUF-Metadata Prediction of Single-Sequence llama.cpp Throughput Across Three Systems — llama.cpp inference throughput on Apple M4 Max can be predicted from GGUF file metadata with 13–14% mean error by counting active parameters rather than total parameters — 3–4x more accurate than the naive total-parameter baseline, with weaker results on RTX 5080 at 36% error.
- Jensen Huang took a call from Trump, and showed off something else, too — At the All-In conference, Jensen Huang took a live Trump phone call on Apple's unreleased $1,999 foldable iPhone — 5 weeks before it hits stores — while both men dismissed AI safety concerns as a 'hoax.' The moment also exposed Apple's renewed reliance on Nvidia to run Apple Intelligence, a decade after the two companies parted ways.
- Nvidia CEO Jensen Huang tells Trump ‘we’re not going to let [an AI slowdown] happen’ — Jensen Huang aligned with Trump live on stage at the All-In Summit, both opposing Dario Amodei's call to slow AI capabilities development — a position Musk and Altman had publicly endorsed. Trump called slowdown advocacy a hoax and a potential Chinese psyop; Huang replied 'we're not going to let that happen.'
- Jensen Huang puts Trump on speakerphone onstage to announce robots won’t take over the world — At the All-In Summit, Nvidia CEO Jensen Huang put President Trump on speakerphone before a large crowd, where Trump dismissed AI safety fears as a 'hoax' and insisted 'the robots will not be taking over' — directly contradicting Anthropic CEO Dario Amodei's weekend essay calling for a coordinated slowdown in AI development.
- Perplexity Portable Computer Is Now Available on Windows, Powered by NVIDIA RTX — Perplexity's Portable Computer AI agent is now available on Windows for NVIDIA GeForce RTX and RTX PRO GPUs with 24GB+ VRAM, running local agentic workflows on a post-trained Qwen 3.8 27B without sending data to the cloud or consuming Perplexity cloud credits.
- Presentation: Decision Models in Agentic Architectures: From Production to Agent Skills — QCon AI presentation from Aletyx CEO Alex Porcelli demonstrates how wrapping LLM agents with DMN (Decision Model and Notation) rules creates auditable, deterministic decision paths — letting business teams own governance logic while engineers maintain architectural control.
- NVIDIA Open-Sources OSMO: One YAML Orchestrates Physical AI Training, Simulation, and Robot Testing — NVIDIA open-sources OSMO under Apache-2.0—a Kubernetes-native orchestrator that lets robotics teams describe training (on GB200/H100), simulation (Isaac Sim on RTX), and hardware-in-the-loop testing (Jetson) in a single YAML file, eliminating cluster-specific glue scripts.
- RoofLang: Enabling AI-Driven Architecting of LLM Inference Systems — RoofLang is a domain-specific language for AI-driven LLM inference architecture search — revealing that DeepSeek V4 achieves 3.5-39.5x higher peak decode throughput than competing models, and enabling an agent to discover new architectures improving DeepSeek V4 Pro by 6.23-50.1% on NVIDIA B300.