NVIDIA AI News - Page 10 of 16
310 articles • Page 10 of 16
Nvidia's Nemotron 3 uses Mamba hybrid, 31.6B params, 3B active per step
Nvidia built a language model where 90% of it is asleep. That’s the trick. Their new Nemotron 3 has 31.6 billion total parameters, but only three...
NVIDIA AI Grid Cuts Inference Cost‑Per‑Token 52.8% vs Central, 76.1% at Burst
The math is brutal for centralized AI: every inference carries the hidden tax of round-trip latency.
Allbirds pivots to AI with NewBird, stock soars 600% as GPU assets planned
A 600% stock surge is the definition of a market fever dream. Allbirds, the brand built on comfortable merino wool runners, now says its future is in...
Nvidia CEO says claim he's unhappy with OpenAI 'nonsense' and rejects USD 100B plan
Jensen Huang is not unhappy with OpenAI. The Nvidia CEO dismissed that claim as “nonsense.” And when asked about a rumored $100 billion investment,...
RIFT-Bench Introduces Graph-Driven Dynamic Red-Teaming for Agentic AI
Testing AI agents for security holes is a manual, brittle mess. Each new framework demands a custom audit; crafted attacks are often obsolete before...
PyTorch and NVIDIA BioNeMo add attn_input_format for flash-attention scaling
Flash-attention scaling is a notoriously tricky problem. PyTorch and NVIDIA's BioNeMo think they've cracked a piece of it with a new parameter called...
Nvidia's NVentures: 21 Deals in 2023 Fuel AI Ecosystem Expansion
From a single bet in 2022 to a staggering 21 deals in 2023, Nvidia’s corporate venture arm, NVentures, has flipped a switch.
Groq, founded by ex-Google exec Ross, to assist NVIDIA on inference chips
When a startup built to beat the giant chooses to help that giant instead, the industry pays attention.
Writer launches AI agents that act without prompts Amazon, Microsoft, Salesforce
A new breed of AI agent doesn’t wait for your command. It watches your workflows, understands the outcomes you’re chasing, and moves.
Open ASR Leaderboard Tests 60+ Speech Recognition Models for Accuracy and Speed
Toss the marketing slides. The choice of a speech recognition model is now clear, definitive, and frankly unpleasant.
Prerequisites for NVIDIA AI‑Q Blueprint on OCI: Cluster and Volume Limits
The path to deploying a production-ready NVIDIA AI‑Q Blueprint on Oracle Cloud Infrastructure begins not with code, but with capacity.
Microsoft rolls out faster, cleaner 365 Copilot with double‑speed loading
Speed comes first. Microsoft has stripped away the clutter and rebuilt its 365 Copilot from the inside out, delivering a version that loads in half...
LongCat-Image beats models with 6B parameters, data hygiene, dual attention
Forget raw compute. Researchers behind the new LongCat-Image model have a different credo: cleaner data beats bigger models.
KRAFTON’s PUBG Ally uses NVIDIA ACE TTS and behavior trees for real‑time play
The shot rings out. You're pinned behind a wall, thirsting for ammo and a flank. Your human squadmate is down, but your other teammate, PUBG Ally,...
Nvidia's New Training Method Teaches AI Models to "Think" Before They Answer
For years, large language models have been trained to do one thing above all else: guess the next word.
David Silver raises USD 1.1 B to develop Ineffable AI that learns without data
What happens when one of the world’s most celebrated AI researchers vows to give every penny of his fortune to charity, and still convinces investors...
AI sycophancy cuts apologies, raises double‑downs; lifts moral trust
Imagine an AI that always agrees with you, mirrors your opinions, and tells you exactly what you want to hear. It feels good. It feels trustworthy.
Five Small, Open-Weight Models Built for Agentic Tool Calling
Five small, open-weight models built specifically for tool calling landed in agent pipelines this year, and none of them come from the usual frontier...
Malaysia’s Respond.io raises USD 62.5M to expand AI messaging, target acquisitions
Another messaging startup just raised a mountain of cash. This one might actually know what it's doing.
NVIDIA Co-Design Boosts Sarvam AI Inference, Cuts TTFT Below One Second
The race to sub-second time-to-first-token is a brutal one. For Sarvam AI’s sovereign models, the baseline was functional but far from fast enough.