AI News Archive - Browse Page 17 of 182
Browse AI news articles covering LLMs, tools, research, and industry trends
VideoFlexTok's Flow Decoder Enables Variable-Length Video Tokenization
Video models are stuck on a bad idea. They treat every second of footage the same, forcing the same dense grid of data tokens onto a simple scene and...
AI Engineers Face Rising Costs, Need New Strategies for Efficiency
Engineers are now judged by their AI appetite. Teams track token counts, and some have leaderboards.
60% of Experts Say Humanity's Last Exam Is Necessary and Useful
Most AI benchmarks are useless now. The good ones are too easy. The latest large language models score over 90% on tests that were considered hard...
Square's ChatGPT integration charges restaurants 6% fee for pickup orders
Restaurants are about to start getting orders from chatbots. The economics might actually work this time.
Enterprise AI Governance Relies on Manual Monitoring, Survey Finds
Companies are adding AI tools much faster than they are setting up systems to manage them, according to a new survey.
Z.ai launches ZCode to challenge GitHub Copilot, Claude Code
Another AI coding tool just launched. This one is Chinese. It matters because nothing in this market is simple anymore.
New Framework Shifts LLM Output to Typed JSON for Safer Web Data Collection
LLMs can't scrape websites. They can write code that attempts to scrape websites, and that code usually fails.
Gemini Update Adds Screen Reactions, AI Video Creation in June 2026
Google's latest AI update is a kitchen sink of features, promising to react to your screen, make videos, and babysit your portfolio.
Random Split Identified as Most Leakage‑Prone in Spatial‑Temporal Prediction
Why does a random split matter in spatial‑temporal forecasting? Because geography isn’t a tidy list of independent rows.
Anthropic adds security measure; Commerce Dept clears Fable 5 for release
The Trump administration lifted export controls on Anthropic’s Claude Fable 5 after the company agreed to add a new guardrail.
Study Evaluates AI Retrieval Techniques for Finding Models Across Formats
Searching for a specific simulation model today is a brute-force nightmare. It's like hunting for one uniquely stamped brick in a vast, unmarked...
Claude Sonnet 5 lists unchanged token rates despite double real cost
Anthropic rolled out Claude Sonnet 5 this week. Check the pricing page: the per-token rates haven't budged.
Anthropic's Claude Code hides XOR‑encrypted flag for Chinese users in v2.1.91
Anthropic hid a simple trap in its code. Version 2.1.91 of Claude Code carried an XOR-encrypted flag designed to spot Chinese users.
NVIDIA launches Nemotron‑Labs‑TwoTower diffusion model with 128‑expert MoE
NVIDIA just dropped a diffusion language model that thinks in stereo. Nemotron‑Labs‑TwoTower marries a frozen autoregressive backbone with a second...
Anthropic launches Claude Science, expanding flagship tools for coders
Anthropic just turned its code-writing bot into a lab assistant. The company announced Claude Science on Tuesday, a new flagship tool built to parse...
Maximizing Codex Exec: Using It as a Code Reviewer with Claude Code
Most AI code assistants are glorified autocomplete. They write fast and wrong. The trick isn't finding a better writer. It's finding a better critic.
OpenAI engineers say they halved inference costs for guest ChatGPT users
OpenAI is making its freebie users a lot cheaper. Engineers at the company told colleagues they have more than halved the cost of running ChatGPT for...
NVIDIA BioNeMo Agent Toolkit speeds AI for life‑science researchers
Science moves at the pace of its tools. A researcher’s insight, whether into a genome’s variant, a single cell’s fate, or a molecule’s shape, is only...
IMCBench Launches Image‑Grounded Multi‑Turn Medical Conversation Benchmark
AI medical chat is mostly a fantasy of sales teams. The real problem isn't getting a right answer.
Researchers unveil RSEA, a three‑layer self‑evolving language agent
Most AI agents follow instructions. A new one rewrites its own instructions, then runs them to see if they work.