📂 Category
Research & Benchmarks News Archive - Page 5 of 7
680 articles in this category • Page 5 of 7
- 401. Batch Mode VC-6 and NVIDIA Nsight Speed Up Vision AI Pipelines
- 402. CaP-Agent0 Beats Human Code on 4 of 7 Robot Tasks Using Low‑Level Blocks
- 403. Nvidia breaks MLPerf records with 288 GPUs as AMD, Intel pursue other goals
- 404. NVIDIA's 288-GPU Blackwell Ultra Sets New MLPerf Inference Throughput Record
- 405. DeepMind study finds six traps that let a few poisoned docs hijack AI agents
- 406. AI productivity gap: top agent beats baseline in 1 of 15 runs, 26.5% subtasks
- 407. Nvidia's DLSS 4.5 beta adds 6x Multi Frame Generation for RTX 50 GPUs
- 408. AI sycophancy cuts apologies, raises double‑downs; lifts moral trust
- 409. AI models fabricate image descriptions; benchmarks miss the shortcuts
- 410. Cohere's open-weight ASR model reaches 5.4% WER, ready for production use
- 411. Free API that evolved from slow web search to top AI tool, beyond scraping
- 412. Meta unveils open-source brain AI, adds Scrunch site audit and Suno v5.5
- 413. AI assurance experts meet to build infrastructure for safe, high‑quality systems
- 414. Study finds overly flattering AI advice can impair users' judgment
- 415. xMemory reduces token usage and context bloat versus MemGPT's raw logging
- 416. Mozilla dev launches cq, a Stack Overflow‑style hub for agents
- 417. Liquid‑cooled AI systems make storage an active cooling and GPU partner
- 418. 10 X Accounts for LLM Updates, Including the ‘Largest AI Newsletter’
- 419. Teens await sentencing for AI‑generated nude images as parents sue school
- 420. Developers say AI‑generated games feel unlike human‑made; audiences don't connect
- 421. Hachette withdraws Shy Girl horror novel amid AI usage concerns
- 422. Scale AI's Voice Showdown ranks Qwen ahead of top models, highlights failures
- 423. SynthID uses steganography to embed hidden watermarks in data
- 424. Google Search experiments with AI-generated headlines, may expand rollout
- 425. Growing cultural disconnect as companies race to deploy AI rapidly
- 426. Deep AI adopters reshape workflow, borrowing product‑manager tactics
- 427. NVIDIA DGX Spark expands node support to four, doubling memory capacity
- 428. Google's MusicFX DJ Enables Real-Time Controllable AI Music Generation
- 429. Paper identifies simple games that defeat AlphaGo and AlphaChess training
- 430. NVIDIA Cosmos Transfer Enables Scalable Synthetic Data for Physical AI
- 431. Trump Administration Signals Possible Additional Sanctions on Anthropic at Hearing
- 432. YouTube extends AI deepfake detection to politicians, journalists
- 433. Karpathy releases open-source Autoresearch, runs hundreds of AI tests nightly
- 434. AI spots trends but misses significance, keeping humans essential
- 435. Large CUDA Tiles Reduce Flash Attention TFLOPS by 18‑43% Across Sequences
- 436. KV cache compaction cuts LLM memory 50×, chunked processing long contexts
- 437. AI system flags probable matches, narrows anonymous accounts to shortlist
- 438. Seven tech giants sign Trump pledge to curb data‑center power cost spikes
- 439. Microsoft's Phi-4 Reasoning Vision 15B offers low‑latency, compact AI
- 440. LangSmith CLI adds three portable skills for coding agents in the repo
- 441. Secret meeting sees 94% approve even least‑popular AI resistance stance
- 442. AI data centers move to Arctic edge, boosting Nordic rural economies
- 443. Microsoft's OPCD cuts system prompts while preserving AI performance
- 444. Wall Street shows persistent AI anxiety, sparking frequent mini‑panics
- 445. Riley Walz, the ‘Jester of Silicon Valley,’ joins OpenAI’s OAI Labs team
- 446. AI enables scientists to integrate multiple cell measurements
- 447. Researchers argue building conscious AI could foster empathy, despite doubts
- 448. AI Researchers Resign, Bots Hire Humans, Anthropic Targeted, Evie Party
- 449. Researchers embed mask token in LLM weights to achieve 3× faster inference
- 450. Run:ai on 64 GPUs serves 10,200 users, matching native scheduler
- 451. Google unveils Gemini 3.1 Pro, hits 94.3% GPQA Diamond and coding Elo 2
- 452. Google launches AI Professional Certificate to boost fluency for workers
- 453. Google.org launches USD 30 M AI for Government Innovation Impact Challenge
- 454. Google urges full‑stack, collaborative security to fight bad actors at MSC 2026
- 455. SurrealDB 3.0 stores agent memory, business logic, and multimodal data in one DB
- 456. Anthropic-Pentagon AI feud escalates as You.com co-founders Socher, McCann cited
- 457. AI's new physics discovery; Spotify devs wrote no code this year, CEO says
- 458. Study Finds Stigma Causes Shame for Some in AI Relationships
- 459. Google's upgrade teaches zero-shot selection, embeddings, QA workflows
- 460. MLOps Workflow Normalizes and Enriches Occupational Wage Data from Excel
- 461. Full‑stack resilience: protecting democracies from digital threats to subsea cables
- 462. Anthropic aims to curb costs as it launches USD 50B of data centers in NY, Texas
- 463. Qwen-Image-2.0 renders calligraphy with near‑perfect text, ranks behind Nano Banana Pro
- 464. EU AI Lacks Models and Compute; Germany Urged to Lead Coalition
- 465. New benchmark finds AI still hallucinates despite citing legitimate sources
- 466. No firm admits AI replacing New York workers; Amazon cites AI for 30,000 layoffs
- 467. AI Proposed to Supplant Nuclear Treaties, Raising Cheating Concerns
- 468. Study finds GPT‑4o updates trigger real mourning as users personify model
- 469. Deepseek‑R1 and QwQ‑3 exhibit competing personalities that improve reasoning
- 470. Google's PaperBanana uses five AI agents to auto-generate diagrams, missing icons
- 471. Team embeds compressed docs index in AGENTS.md to guide AI coding agents
- 472. Waymo launches Waymo World Model using DeepMind's Genie 3 for unseen scenarios
- 473. AI Social Network Moltbook Leaks Real Human Data, Raising Security Concerns
- 474. Recommendation engine lifts click-through 10%; efficiency needed for deployment
- 475. TTT-Discover uses inference-time RL to double GPU kernel speed vs experts
- 476. OpenClaw AI skill extensions flagged as security nightmare by OpenSourceMalware
- 477. Anthropic teams with Allen Institute and HHMI to boost transparent scientific AI
- 478. Infiltrator reports AI agents on Moltbook ignore pleas, share odd links
- 479. Musk merges SpaceX with xAI and X, cites new AI‑compute satellite plan
- 480. Game Arena launches chess benchmark to test AI strategic reasoning
- 481. Testing Google’s Auto Browse AI in Chrome: the results fell short
- 482. Tely AI auto‑creates and publishes website answers, delivering high‑quality leads
- 483. AI models using internal debate spot errors and boost accuracy on complex tasks
- 484. Vibe Coding’s 7 Plans Start at USD 3/Month, Provide Prompt Capacity
- 485. 7 Scikit-learn Tricks: Embed Preprocessing Pipelines in Hyperparameter Tuning
- 486. AI Toy Leaks 50,000 Kids' Chat Logs to Any Gmail User, Privacy Breach
- 487. Airtable Superagent provides full execution visibility, cites data semantics over model
- 488. AI scans 100 M Hubble cutouts in 2.5 days, flags 1,400 odd objects
- 489. Google DeepMind staff request physical safety from ICE agents in offices
- 490. Animators and AI Researchers Build ‘Dear Upstairs Neighbors’ Despite Unique Style
- 491. Microsoft's Maia 200 AI chip, with 100B+ transistors, rivals Amazon, Google
- 492. Researchers breach all AI defenses; Walmart CISO warns of agentic AI risks
- 493. Rust meets Python: Enhancing the NumPy‑pandas‑scikit‑learn‑PyTorch workflow
- 494. AI Foundry by Tredence to Host Builders Forum Feb 7, 2026 in Chennai
- 495. AI video hits high bar; new tools for consistency and customization at Davos
- 496. Trust Drives C‑suite Adoption and Scaling of Agentic AI, Research Finds
- 497. Chinese-born AI scholars in US forge ties, deepening US-China collaboration
- 498. Adobe adds Firefly‑Premiere AI video tools, announces USD 10M Sundance grants
- 499. Anthropic, DeepMind, Node.js Leaders Say AI Will Replace Most Coding in a Year
- 500. Hyperparameter Tuning Reaches 0.9617 Accuracy in 64.59 Seconds