AI Regulation News - Page 11 of 19
370 articles • Page 11 of 19
NeuronFuzz Uses Safety Oracle to Guide Fuzzing Without Generating LLM Responses
A team behind a new paper on LLM safety testing is targeting a specific bottleneck in how researchers check whether jailbreak defenses actually hold...
Counter-Strike Sets New Benchmark for Vibe Coding, Says Ex-Mixpanel CEO
Counter-Strike is not just a game anymore, it’s a proving ground for artificial intelligence.
NVIDIA releases Cosmos 3 with Super‑Text2Image and Nano‑Policy‑DROID
NVIDIA’s Cosmos 3 isn’t just another model drop. It’s a two-tower mixture-of-transformers foundation model that fuses physical reasoning, world...
Pentagon’s Claude platform cuts costs 70% with secure single‑tenant AI
They wanted cost savings. They got a 70% reduction. The Pentagon’s deployment of Anthropic’s Claude platform isn’t just another efficiency play, it’s...
Key Topics for LLM Engineers: Using Instruction Data to Align Models
A pretrained language model can generate coherent text. It doesn't, however, know how to answer a question directly or avoid harmful content without...
AI backlash surges as politicians finally grasp public sentiment
Pollsters are stunned. Public sentiment on artificial intelligence is shifting faster than any issue they can remember.
Nonprofits lobbying OpenAI regulation later received company subpoenas
OpenAI is using subpoenas as a weapon. At least seven nonprofits that lobbied for stricter oversight of the company’s controversial shift to a...
UK tests Mythos AI, noting its ability to chain multistep attacks
Last year, OpenAI's GPT-3.5 Turbo failed every one of the UK AI Safety Institute's basic hacking puzzles. That was the baseline.
The Vergecast on Claude's vibe coding, email safety, and phone upgrade timing
The latest Vergecast lands with a trio of questions that cut straight to the heart of modern tech anxiety: Can an AI actually write its own code, and...
AgentWall adds runtime safety layer for local AI agents' actions
The real danger isn't a rogue thought; it's a rogue command. Take the developer running an AI agent locally, pointed at a filesystem littered with...
Claude Mythos USD 36,428 for 122 exploit episodes; GPT‑5.5 USD 3,075 for 123
Anthropic's latest AI model costs twelve times more than OpenAI's to do roughly the same job.
Researchers Propose Method to Block Illegal AI Content for Kids
The National Center for Missing and Exploited Children logged more than 1.5 million reports of AI-generated child sexual abuse material in 2025, up...
Meta plans facial recognition for AI smart glasses, amid privacy concerns
Meta is quietly reopening the door to facial recognition, this time through the lens of its AI-powered smart glasses.
Anthropic files lawsuit against U.S. government in federal court
Anthropic just filed a lawsuit against the U.S. government in federal court. The AI company is taking its fight with regulators out of the boardroom...
Supabase Launches Evals to Benchmark Claude, Codex, and OpenCode on Real Tasks
Supabase pushed its evaluation framework to GitHub this week under an Apache-2.0 license, giving developers a way to check whether AI coding agents...
Google AI's EnvHarness Makes Static AI Training Worlds Adaptable
Most agent training environments don't change. A robot arm simulator behaves the same on day one as it does after ten thousand episodes of the policy...
LWiAI Podcast #228: OpenAI unveils GPT-5.2, Runway rolls out first world model
The artificial intelligence arms race just got a little more personal. OpenAI drops GPT-5.2 into the agentic arena, and Runway finally unveils its...
Trump signs first major AI regulation order of second term, cites Anthropic
Donald Trump put his signature on the first major AI rule of his new term. The document might as well have come stamped with an Anthropic watermark.
Tim Cook and Sundar Pichai called cowards over Tumblr, Musk, and shareholder value debate
These are the men who sell safety. Tim Cook and Sundar Pichai have built the most valuable companies on earth by promising control.
Open-Weight AI Safety Gap Persists as Users Can Remove Hosted Protections
GLM-5.2, the open-weight model released by China's Z.ai, refused zero offensive cyber tasks and zero dual-use biology tasks in a new evaluation from...