NVIDIA AI News - Page 15 of 16
310 articles • Page 15 of 16
NY RAISE Act targets AI safety as Zhipu AI launches GLM 4.7, Nvidia buys Groq
The AI industry is moving faster than the laws meant to contain it. New York just answered with the RAISE Act, the second major U.S.
NVIDIA RTX PRO 4500 Blackwell GPUs Power New Amazon EC2 G7 Instances
Amazon just dropped new chips into its cloud. They’re banking you’ll pay for them.
Nvidia's 3B Nemotron-Cascade 2 wins math and coding gold; recipe open‑source
In the rush to build ever-larger AI models, Nvidia just proved a powerful counterpoint: smarter training beats more parameters.
ADaSci Introduces AI Agentic Bootcamp to Build Leadership Skills and Certification
Every tech leader is now expected to build systems that think and act on their own. Few have any idea how to start.
NVIDIA BioNeMo wraps CPU layer with DistributedTriangleMultiplication
The bottleneck has always been memory. In structural biology, protein complexes sprawl across thousands of residues, and single‑GPU limits have kept...
Nvidia unveils DGX Station supercomputer for trillion‑parameter AI at GTC 2026
Nvidia just made the impossible desktop-sized. The DGX Station, unveiled at GTC 2026, is a supercomputer that handles trillion-parameter AI models...
HPE AI Factory and NVIDIA unveil Vera, first CPU built for agents
HPE and NVIDIA just announced a new chip called Vera, which they are calling the first CPU built specifically for AI agents.
Mixture-of-Experts AI Models Run 10× Faster on NVIDIA Blackwell NVL72
Mixture-of-Experts models are compute gluttons. Running them is famously expensive.
TriAttention KV Cache Compression Matches Full Attention, 2.5× Faster
The key-value cache is the bottleneck of modern long-context inference, growing without bound as sequences lengthen, slowing throughput to a crawl.
You.com AI grounding guide, three-part method beating RAG, noted at Nvidia GTC
Nvidia’s big AI conference had everyone pitching their next miracle. You.com took a different angle.
NVIDIA TensorRT Enables Context Parallelism for Multi‑GPU AI Inference
AI is hitting a wall with long prompts, and the transformer is to blame. Its attention mechanism has a quadratic scaling problem: double the sequence...
Reviews suggest Nvidia DGX Spark mini-DGX copies DGX design, unveiled by Huang
Nvidia has built a business out of making the biggest, most expensive computers for AI. Now it's making a smaller one.
Universal Music partners with Nvidia on AI model for smarter song search
Universal Music Group spent two years suing AI firms. Now it's paying one. The world's largest music label has partnered with Nvidia to build a new...
Build Vision AI Pipelines with NVIDIA DeepStream and Custom Models
Every engineer dreams of slotting a custom vision model into a real-time pipeline. The nightmare?
Nvidia's Cosmos Reason 2 boosts robot reasoning for complex tasks
Nvidia’s Cosmos Reason 2 doesn’t just make robots smarter, it rewires how they think.
NVIDIA backs PyTorch with open-source contributions, leadership, event support
Nvidia sells chips. Very expensive ones. So why is a hardware giant pouring resources into free software like PyTorch?
Calibration uses NVIDIA Triton Llama-3-8B A10 and vLLM Qwen2.5-7B RTX 4090 data
Inference is a story of two systems, and the story begins with a single millisecond. Consider a transaction authorization.
NVIDIA XR AI equips AR glasses with LabOS co‑scientist to aid CRISPR work
CRISPR labs are where mistakes cost months. One mislabeled tube, one forgotten step in the protocol, and the whole experiment is trash.
NVIDIA cuts prices on Jetson edge-AI developer kits for holiday shoppers
This holiday season, NVIDIA is rewriting the wish list for robotics and AI developers.
RDMA Cuts CPU Use in S3-Compatible Storage, Boosting AI Performance
AI needs data constantly, and that movement is surprisingly expensive. Every chunk of training data pulled from an object store like S3 traditionally...