AI News Archive - Browse Page 40 of 182
Browse AI news articles covering LLMs, tools, research, and industry trends
StepFun launches StepAudio 2.5 Realtime, evaluated via mobile app raters
The numbers tell a compelling story. StepAudio 2.5 Realtime didn’t just edge past competitors; it swept every benchmark dimension, from subjective...
Create a Claude Cowork‑Style Browser Agent with Playwright MCP and Claude Desktop
The official MCP documentation is a two-sentence trap. It hands you a JSON file path and a menu toggle. That’s it.
ByteDance study: LMMs answer questions better than full-page transcription
Teaching a multimodal model to read an entire document, word for word, might actually be holding it back.
Anthropic may keep supplying Claude to NSA despite Pentagon risk flag
The Pentagon flagged Anthropic as a supply chain threat. The NSA needs chips it doesn’t have. And yet, a deal is nearly done.
Claude Code auto‑creates AI scaling algorithms; new control allocates compute
The quest to scale AI has long been a human-driven art, tuning knobs, guessing heuristics, burning compute to find answers.
SuperClaude workflow ranks security issues, details attack vectors, gives fixes
Security reviews often feel like drinking from a firehose, vulnerabilities pour in, but prioritization remains fuzzy, attack vectors stay buried in...
Deepseek makes 75% discount permanent, output tokens priced over 34× below GPT‑5.5
Everyone expected the AI price war to ease up eventually. It hasn’t. Deepseek just dropped the discount hammer for good, making its 75% price cut...
Anthropic: Claude Mythos Preview finds ~3,900 high‑severity open‑source bugs
For one month, Anthropic turned its Claude Mythos Preview AI loose with roughly fifty partners.
Agent explores once, then compiles branch‑free recipe to bypass LLM thereafter
Everyone building AI agents knows the trick will eventually be making them stop thinking.
D&B rebuilds 642 million‑business database after AI agents hit limits
Why did D&B have to start from scratch? The answer lies in a data architecture that was never meant for autonomous agents.
Meta launches Forum: Reddit‑style advice within Facebook groups, AI‑assisted
Forget appending “Reddit” to your Google search. Stop copy-pasting your life crisis into ChatGPT.
CopilotKit launches AG-UI to bridge agent‑human interaction layer
Why does this matter? Because the tools that let autonomous agents talk to people have finally found a stable foundation.
AgentCo-op imports and refines searched workflows via component grounding
Forget building complex AI workflows from scratch. A team of researchers has a better idea: start with a blueprint.
LLM‑RL Agent Manages CAD, CAE and Geometry Revision for Closed‑Loop Optimization
Every engineering software demo promises a robot that designs, tests, and fixes its own work. They never deliver.
SOLAR introduced as self‑optimizing autonomous agent for continual learning
Machine learning has a chronic case of amnesia. Show a model something new, and yesterday's lesson often vanishes.
Language Models Forecast Research Success Using 11,488 Comparative Idea Pairs
Forget raw intelligence. Predicting a good research idea is a job for a well-trained referee.
OpenAI’s Q1 2026 adjusted margin slips to –122%, burning USD 1.22 per USD 1 earned
OpenAI is losing $1.22 for every dollar it earns, even after excluding stock-based compensation.
VSAS‑Bench Introduces Standardized Real‑Time Evaluation for Visual Assistants
Building a visual assistant that keeps up with reality is hard. Measuring it is harder. Most tests are a slideshow. The real world is a live feed.
F_Call_Analysis_Planner forwards Parent_Instruction to generate Selection_Rule
The best part of any system is the small, stupid piece that does one job perfectly. This one is called F_Call_Analysis_Planner.
Temporal Contrastive Transformer embeddings boost financial crime detection
The promise of self-supervised learning in financial crime detection rests on a single, powerful idea: that a model can discover behavioral patterns...