Research & Benchmarks - Page 17 of 28
Academic AI research, performance benchmarks, scientific breakthroughs, and peer-reviewed studies advancing artificial intelligence frontiers.
Academic AI research, performance benchmarks, scientific breakthroughs, and peer-reviewed studies advancing artificial intelligence frontiers.
Forget the hype. NVIDIA's Run:ai just ran a brutally simple test to see what happens when you actually use the thing: take 64 GPUs and throw users at them. Not a simulation. 10,200 people hit it at once. The result?
The numbers are stark. Google’s Gemini 3.1 Pro just logged a 94.3% score on the grueling GPQA Diamond benchmark, a scientific reasoning gauntlet that separates brittle models from genuinely capable ones.
Google just launched an AI Professional Certificate. The move is a direct response to a glaring disconnect: a new Ipsos study, commissioned by the company, found 70% of managers deem AI training critical, yet only 14% of workers have received it.
Google.org is writing checks. Big ones. Announced Thursday at the AI Impact Summit in New Delhi, a $30 million "Global AI for Government Innovation Impact Challenge" seeks public-private teams to rebuild civic services from the ground up.
The digital battlefield is no longer a place for silos and slow reactions. Bad actors move fast, exploiting every crack in the system. To stop them, we must move faster – and together.
Most AI stacks today are held together with duct tape and prayers. A vector database over here, a graph database over there, a relational engine for structured records, and a caching layer to paper over the cracks. The result?
The clash between Anthropic and the Pentagon is no longer a quiet war of memos. It’s escalating. And now, You.com’s co-founders Richard Socher and Bryan McCann, two of the most cited AI researchers alive, are being drawn directly into the fire.
Spotify’s top developers have not written a single line of code all year. That’s not a failure, it’s a strategy. CEO Gustav Soderstrom says the company is “all in” on AI, and the results are reshaping what engineering even means.
People are falling for artificial intelligence. They’re having real, sustained romances with it, and a new study in *Computers in Human Behavior Reports* confirms a predictable, human consequence: a deep fear of mockery.
The race for raw AI speed is blinding. Just look at OpenAI's new GPT-5.3-Codex-Spark, hitting over a thousand tokens per second on Cerebras hardware. That’s a direct challenge to Nvidia. But velocity is expensive.
The raw Excel file arrives like an unlabeled drawer in a cabinet of records: messy, inconsistent, and full of promise. One cell holds a dollar figure, another a string with commas and a trailing space.
The internet runs on wet string. A million miles of it, at the bottom of the ocean. Democratic governments, stock trades, and a billion cat videos all depend on these fragile glass threads. They are also absurdly easy to break.
The math behind artificial intelligence is brutally expensive. A few keystrokes in a chatbot can consume more power than a modest home does in a day. Anthropic understands this calculus intimately.
Most AI image generators are functionally illiterate. They produce a poster or a sign, and the text is a garbled mess of plausible shapes. Alibaba's Qwen-Image-2.0 has finally learned to read.
Europe has brilliant AI researchers and almost no functional AI. Its own rules are to blame. A new report for the European Commission details the continental deadlock.
Citations are supposed to be anchors to reality. A new benchmark shows they’re becoming camouflage for lies. Researchers have identified a quiet, advanced-stage form of AI hallucination. The problem isn’t that models invent sources.
Amazon blames artificial intelligence for 30,000 job cuts. Morgan Stanley whispers about automation. Across the US, nearly 55,000 companies have publicly cited AI as a reason for layoffs.
For decades, verification meant film canisters ejected from satellites and retrieved mid-air. Analysts in windowless rooms pored over the grainy frames. The new pitch is silicon vigilance: an unblinking AI sentinel.
People got attached. They named their chatbot Rui or Hugh. They treated it like a diary that talked back, a friend who never got bored. Then OpenAI updated GPT-4o and broke the spell. The grief was real, and it wasn't about features.
Two new AI models don't agree on much. That's the point. Deepseek‑R1 and QwQ‑3 work by hosting a committee of bickering personas inside their own process, a deliberate clash that makes them smarter. The evidence is in a recent study.
Learn to build AI-powered apps without coding. Our comprehensive review of No Code MBA's course.
Curated collection of AI tools, courses, and frameworks to accelerate your AI journey.
Get the week's most important AI news delivered to your inbox every week.