AI News Archive - Browse Page 28 of 170
Browse AI news articles covering LLMs, tools, research, and industry trends
Hotz warns AI coding agents could be costly despite 10x productivity boost
George Hotz once jailbroke the iPhone. Now, he's trying to break the spell of AI coding assistants.
Accurate source citations boost AI answer quality, study finds
A language model that cites its sources isn’t just being polite, it’s being smarter.
Google Antigravity 2.0 Retains Gemini CLI Features as Antigravity Plugins
Google Antigravity 2.0 doesn’t rebrand for the sake of rebranding. It keeps the CLI features developers actually use, Agent Skills, Hooks, Subagents,...
FuRA uses spectral preconditioning with full‑rank SVD for efficient fine‑tuning
Fine-tuning a big model has always been a choice between wasting money and settling for less.
Positional copying dominates answer readout in 1‑3B LMs on GSM8K
We’ve been told that chain-of-thought makes small language models reason. It doesn’t. It just tells them where to look.
Study Introduces Orchestration Overhead Index to Measure AI Energy Costs
The energy footprint of AI is usually tallied inference by inference, but that misses the hidden cost of coordination.
StepFun launches StepAudio 2.5 Realtime, evaluated via mobile app raters
The numbers tell a compelling story. StepAudio 2.5 Realtime didn’t just edge past competitors; it swept every benchmark dimension, from subjective...
Create a Claude Cowork‑Style Browser Agent with Playwright MCP and Claude Desktop
The official MCP documentation is a two-sentence trap. It hands you a JSON file path and a menu toggle. That’s it.
ByteDance study: LMMs answer questions better than full-page transcription
Teaching a multimodal model to read an entire document, word for word, might actually be holding it back.
Anthropic may keep supplying Claude to NSA despite Pentagon risk flag
The Pentagon flagged Anthropic as a supply chain threat. The NSA needs chips it doesn’t have. And yet, a deal is nearly done.
Claude Code auto‑creates AI scaling algorithms; new control allocates compute
The quest to scale AI has long been a human-driven art, tuning knobs, guessing heuristics, burning compute to find answers.
SuperClaude workflow ranks security issues, details attack vectors, gives fixes
Security reviews often feel like drinking from a firehose, vulnerabilities pour in, but prioritization remains fuzzy, attack vectors stay buried in...
Deepseek makes 75% discount permanent, output tokens priced over 34× below GPT‑5.5
Everyone expected the AI price war to ease up eventually. It hasn’t. Deepseek just dropped the discount hammer for good, making its 75% price cut...
Anthropic: Claude Mythos Preview finds ~3,900 high‑severity open‑source bugs
For one month, Anthropic turned its Claude Mythos Preview AI loose with roughly fifty partners.
Agent explores once, then compiles branch‑free recipe to bypass LLM thereafter
Everyone building AI agents knows the trick will eventually be making them stop thinking.
D&B rebuilds 642 million‑business database after AI agents hit limits
Why did D&B have to start from scratch? The answer lies in a data architecture that was never meant for autonomous agents.
Meta launches Forum: Reddit‑style advice within Facebook groups, AI‑assisted
Forget appending “Reddit” to your Google search. Stop copy-pasting your life crisis into ChatGPT.
CopilotKit launches AG-UI to bridge agent‑human interaction layer
Why does this matter? Because the tools that let autonomous agents talk to people have finally found a stable foundation.
AgentCo-op imports and refines searched workflows via component grounding
Forget building complex AI workflows from scratch. A team of researchers has a better idea: start with a blueprint.
LLM‑RL Agent Manages CAD, CAE and Geometry Revision for Closed‑Loop Optimization
Every engineering software demo promises a robot that designs, tests, and fixes its own work. They never deliver.