Research & Benchmarks - Page 23 of 35
Academic AI research, performance benchmarks, scientific breakthroughs, and peer-reviewed studies advancing artificial intelligence frontiers.
Academic AI research, performance benchmarks, scientific breakthroughs, and peer-reviewed studies advancing artificial intelligence frontiers.
Watermarking is often just security theater. A logo to crop. A tag to strip. A thin layer of metadata scraped off by the first social media upload. But DeepMind's SynthID has a different goal: to become a permanent ghost in the machine.
Google has a new habit of running tests that become permanent fixtures before anyone can ask why. Its latest involves rewriting the headlines of news articles directly in Search results.
Corporate America is in a full sprint to wire artificial intelligence into everything. Resumes are sorted by it. Burgers are grilled by it. The pressure to deploy is immense, the rhetoric breathless.
Generative AI is a Swiss Army knife, packed with dozens of functions. But most people pull out the blade and call it a day. Deep adopters do something different: they think like product managers.
NVIDIA’s DGX Spark can now connect four machines where it used to stop at two. This is not a subtle upgrade. It doubles the available memory from 256 gigabytes to 512, creating a small but potent cluster that fits in a server rack.
The barrier between machine and musician just crumbled. Google’s MusicFX DJ doesn’t just generate audio, it bends it to your will in real time. This is no longer about waiting for a model to spit out a track after a prompt.
AlphaGo conquered Go with moves that stunned the world. AlphaChess dismantled grandmasters at their own game. Yet both stumble on a game a child can learn in seconds: Nim, where players simply remove matchsticks from rows until none remain.
Teaching robots to navigate our world demands a torrent of visual data. Real footage is costly and scarce. So the industry relies on simulations, but their tell-tale artificiality is a problem—the AI models can spot it.
The Trump administration was done punishing Anthropic. Until it wasn't. At a federal court hearing this week, the government's lawyer was asked a simple question. Would the White House stop hitting the AI startup with new sanctions?
YouTube has never been able to decide if AI is its biggest problem or its best product. It purges junk channels filled with fake trailers one day and gives creators slick AI tools for making videos the next.
Andrej Karpathy built a robot that writes research papers. The former OpenAI and Tesla luminary quietly posted the code, called Autoresearch, on GitHub. It automates the scientific method—for code. You give it a problem.
The machines can spot the pattern. They just can’t tell you if it matters. A 12% swing in spend appears in the variance report, healthy growth or hidden collapse? The algorithm won’t know.
Bigger tiles should mean fewer memory accesses, faster attention, right? Not on NVIDIA GPUs. Across every sequence length tested, large CUDA tiles actually cratered Flash Attention throughput by 18 to 43 percent.
A research team slashed the memory load of large language models by 98 percent. Their secret? Compressing the model's key-value cache fiftyfold.
Anonymous posting just got harder, full stop. Researchers have crafted a method using large language models that can effectively unmask people.
The White House just secured a commitment from seven of the world’s largest tech companies, Amazon, Google, Microsoft, and others, to shoulder the financial burden of their own exploding energy appetite.
Bigger isn't smarter. While the AI industry obsesses over trillion-parameter behemoths, Microsoft just released a model that fits in your pocket. Phi-4-reasoning-vision-15B is a 15-billion parameter model. It is small. It is fast.
A coding agent is only as good as its environment. Give it the right tools, and it transforms. LangSmith CLI just gave agents three new portable skills, trace, dataset, and evaluator, that snap directly into any repo. These aren't abstract concepts.
Behind closed doors, 94% of attendees approved the least popular stance on AI resistance. That number alone should stop you cold. The dissenters? They didn’t matter. The partisanship that fractures public debate, Grok vs. Anthropic, “based” vs.
Forget Silicon Valley. The new front line for the AI boom is a frozen field in northern Sweden. Europe is running out of space and power for the vast server farms that make AI models work.
Learn to build AI-powered apps without coding. Our comprehensive review of No Code MBA's course.
Curated collection of AI tools, courses, and frameworks to accelerate your AI journey.
Get the week's most important AI news delivered to your inbox every week.