📂 Category
Research & Benchmarks News Archive - Page 4 of 6
547 articles in this category • Page 4 of 6
- 301. AI spots trends but misses significance, keeping humans essential
- 302. Large CUDA Tiles Reduce Flash Attention TFLOPS by 18‑43% Across Sequences
- 303. KV cache compaction cuts LLM memory 50×, chunked processing long contexts
- 304. AI system flags probable matches, narrows anonymous accounts to shortlist
- 305. Seven tech giants sign Trump pledge to curb data‑center power cost spikes
- 306. Microsoft's Phi-4 Reasoning Vision 15B offers low‑latency, compact AI
- 307. LangSmith CLI adds three portable skills for coding agents in the repo
- 308. Secret meeting sees 94% approve even least‑popular AI resistance stance
- 309. AI data centers move to Arctic edge, boosting Nordic rural economies
- 310. Microsoft's OPCD cuts system prompts while preserving AI performance
- 311. Wall Street shows persistent AI anxiety, sparking frequent mini‑panics
- 312. Riley Walz, the ‘Jester of Silicon Valley,’ joins OpenAI’s OAI Labs team
- 313. AI enables scientists to integrate multiple cell measurements
- 314. Researchers argue building conscious AI could foster empathy, despite doubts
- 315. AI Researchers Resign, Bots Hire Humans, Anthropic Targeted, Evie Party
- 316. Researchers embed mask token in LLM weights to achieve 3× faster inference
- 317. Run:ai on 64 GPUs serves 10,200 users, matching native scheduler
- 318. Google unveils Gemini 3.1 Pro, hits 94.3% GPQA Diamond and coding Elo 2
- 319. Google launches AI Professional Certificate to boost fluency for workers
- 320. Google.org launches USD 30 M AI for Government Innovation Impact Challenge
- 321. Google urges full‑stack, collaborative security to fight bad actors at MSC 2026
- 322. SurrealDB 3.0 stores agent memory, business logic, and multimodal data in one DB
- 323. Anthropic-Pentagon AI feud escalates as You.com co-founders Socher, McCann cited
- 324. AI's new physics discovery; Spotify devs wrote no code this year, CEO says
- 325. Study Finds Stigma Causes Shame for Some in AI Relationships
- 326. Google's upgrade teaches zero-shot selection, embeddings, QA workflows
- 327. MLOps Workflow Normalizes and Enriches Occupational Wage Data from Excel
- 328. Full‑stack resilience: protecting democracies from digital threats to subsea cables
- 329. Anthropic aims to curb costs as it launches USD 50B of data centers in NY, Texas
- 330. Qwen-Image-2.0 renders calligraphy with near‑perfect text, ranks behind Nano Banana Pro
- 331. EU AI Lacks Models and Compute; Germany Urged to Lead Coalition
- 332. New benchmark finds AI still hallucinates despite citing legitimate sources
- 333. No firm admits AI replacing New York workers; Amazon cites AI for 30,000 layoffs
- 334. AI Proposed to Supplant Nuclear Treaties, Raising Cheating Concerns
- 335. Study finds GPT‑4o updates trigger real mourning as users personify model
- 336. Deepseek‑R1 and QwQ‑3 exhibit competing personalities that improve reasoning
- 337. Google's PaperBanana uses five AI agents to auto-generate diagrams, missing icons
- 338. Team embeds compressed docs index in AGENTS.md to guide AI coding agents
- 339. Waymo launches Waymo World Model using DeepMind's Genie 3 for unseen scenarios
- 340. AI Social Network Moltbook Leaks Real Human Data, Raising Security Concerns
- 341. Recommendation engine lifts click-through 10%; efficiency needed for deployment
- 342. TTT-Discover uses inference-time RL to double GPU kernel speed vs experts
- 343. OpenClaw AI skill extensions flagged as security nightmare by OpenSourceMalware
- 344. Anthropic teams with Allen Institute and HHMI to boost transparent scientific AI
- 345. Infiltrator reports AI agents on Moltbook ignore pleas, share odd links
- 346. Musk merges SpaceX with xAI and X, cites new AI‑compute satellite plan
- 347. Game Arena launches chess benchmark to test AI strategic reasoning
- 348. Testing Google’s Auto Browse AI in Chrome: the results fell short
- 349. Tely AI auto‑creates and publishes website answers, delivering high‑quality leads
- 350. AI models using internal debate spot errors and boost accuracy on complex tasks
- 351. Vibe Coding’s 7 Plans Start at USD 3/Month, Provide Prompt Capacity
- 352. 7 Scikit-learn Tricks: Embed Preprocessing Pipelines in Hyperparameter Tuning
- 353. AI Toy Leaks 50,000 Kids' Chat Logs to Any Gmail User, Privacy Breach
- 354. Airtable Superagent provides full execution visibility, cites data semantics over model
- 355. AI scans 100 M Hubble cutouts in 2.5 days, flags 1,400 odd objects
- 356. Google DeepMind staff request physical safety from ICE agents in offices
- 357. Animators and AI Researchers Build ‘Dear Upstairs Neighbors’ Despite Unique Style
- 358. Microsoft's Maia 200 AI chip, with 100B+ transistors, rivals Amazon, Google
- 359. Researchers breach all AI defenses; Walmart CISO warns of agentic AI risks
- 360. Rust meets Python: Enhancing the NumPy‑pandas‑scikit‑learn‑PyTorch workflow
- 361. AI Foundry by Tredence to Host Builders Forum Feb 7, 2026 in Chennai
- 362. AI video hits high bar; new tools for consistency and customization at Davos
- 363. Trust Drives C‑suite Adoption and Scaling of Agentic AI, Research Finds
- 364. Chinese-born AI scholars in US forge ties, deepening US-China collaboration
- 365. Adobe adds Firefly‑Premiere AI video tools, announces USD 10M Sundance grants
- 366. Anthropic, DeepMind, Node.js Leaders Say AI Will Replace Most Coding in a Year
- 367. Hyperparameter Tuning Reaches 0.9617 Accuracy in 64.59 Seconds
- 368. RLVR lifts sampling efficiency, not reasoning; base models hold trajectories
- 369. OpenAI Safety Lead Moves to Anthropic's AI Risk Research Team
- 370. AI Researchers Reveal Token Warehousing Strategy to Cut GPU Computational Waste
- 371. AI Tool Detects Dangerous Blood Cells Doctors Might Overlook
- 372. IndiaAI Mission Launches 62 AI and Data Labs Across Uttar Pradesh
- 373. Stanford AI Detects Hidden Disease Signals in Large-Scale Sleep Data
- 374. X limits Grok image tool to paid users; 1 obscene request/min, 102 in 5 mins
- 375. Replit CEO says using more tokens yields higher-quality inputs, then tests apps
- 376. Dell says AI-focused PCs confuse consumers, who show little interest
- 377. Vibe Coding Remains Early Stage, Real-World Reliability Still Distant
- 378. New Magnetic Nanoparticle Approach Merges Heating and Healing for Bone Cancer
- 379. Tredence hosts AI Foundry workshop in Chennai for AI system designers
- 380. Test-Time Training adds dual-memory to Transformers, keeping inference cheap
- 381. Analysis overhauls AI Index; GPT-5.2 beats professionals on 70.9% of tasks
- 382. MIT study probes memorization risk of clinical AI with de-identified data
- 383. AMD announces Ryzen AI 400 at CES, resembles AI 300 in laptops
- 384. Docker Trick: Deterministic OS Packages in One Layer to Prevent ML Failures
- 385. Notion’s simplified AI agent feature feels indispensable, says engineer
- 386. DeepSeek's architectural fix improves large-scale reasoning, follows GRPO work
- 387. Nested Learning's Continuum Memory System Redefines AI Memory for 2026
- 388. New framework lets agentic AI tools adapt to fill main agent knowledge gaps
- 389. Opera Neon: AI-native browser that researches, compares prices, codes
- 390. Fusion reactors could produce dark-sector particles via neutron emissions
- 391. CIOs drive AI experiments by embedding ready-to-use features into everyday tools
- 392. Dell and NVIDIA Host AI Developer Meetup in Hyderabad to Discuss Solutions
- 393. Google highlights AI-driven chip, infrastructure and robotics advances in 2025
- 394. LeCun and Hassabis dispute meaning of ‘general intelligence’
- 395. Qwen3-4B-Instruct-2507: 4B-parameter model boosts Raspberry Pi AI
- 396. Dell and NVIDIA host AI developer meetup in Bengaluru on deployment trade-offs
- 397. GPT-5.2 leads FrontierScience test, but falters on real research tasks
- 398. OpenUSD and NVIDIA Halos Enhance Robotaxi Safety with Synthetic Data, SimReady
- 399. Audio Dataset Valuable for Listening Models, Tackles Noise, Accents, Timing
- 400. Fastweb and Vodafone use LangGraph LLM Compiler to automate customer requests