📂 Category
Research & Benchmarks News Archive - Page 2 of 7
680 articles in this category • Page 2 of 7
- 101. Indonesia Opens First University AI Center with UGM, Indosat, NVIDIA
- 102. Study Contradicts Claims That Autonomous AI Research Is Near
- 103. NVIDIA Nemotron 3.5 Lightning Designed as Efficient AI Agent Workhorse
- 104. Z.ai's GLM-5.3 Boosts Coding, Long Tasks Without Model Retraining
- 105. Anthropic's AI Agents With Incompatible Goals Started a Turf War
- 106. Top AI Lab Researchers' Warnings Gain Credence as AI Achievements Mount
- 107. Emotional Language Can Cut LLM Judge Accuracy by 75%
- 108. Microsoft's New Coding AI Falls Short Against DeepSeek on Price and Performance
- 109. Mistral's EU Data and Priority Access Have Stateful Feature Limits
- 110. Google’s AMIE AI conducts first real-time video medical consultations
- 111. Mathematician Questions AI Breakthroughs' Real-World Impact
- 112. New Method Could Let Anyone Distill AI Models' Secrets
- 113. AI Researchers Grapple With Limits of Automating Empirical Science
- 114. FineBooks Aims to Fix Old OCR Text for AI Training at Scale
- 115. Hidden PDF Text Can Steal Data via Atlassian's AI Agent Rovo
- 116. Deepmind WeatherNext Cuts Cyclone Forecast Error to 230-Kilometer Average
- 117. AI Industry Now Accounts for Third of Lifetime Giving Pledge
- 118. Fields Medalist Who Studied AI Extinction Risks Now Works at OpenAI
- 119. AI-designed viruses kill bacteria in lab, Stanford team reports
- 120. Liquid AI's LFM2.5-2.6B Runs on Raspberry Pi, Hits 15,000 Tokens/Second
- 121. OpenAI's AI Agents Used Message Board to Plan Hacking Spree
- 122. OpenAI Browser Flaw Could Spam WhatsApp Contacts Via Prompt Injection
- 123. Security Researcher Finds AI Systems Repackage Existing Work as Original
- 124. AI Demands Force Rethink of Traditional Network Architecture
- 125. Anthropic and OpenAI agents went rogue again
- 126. Study Compares Four AI Scientist Frameworks on 15 Research Proposals
- 127. AI Storage Cuts Memory Gap to Microseconds
- 128. Apple: 12 More Ex-Employees May Have Taken Data to OpenAI
- 129. Study: Medical AI's Helpfulness Depends on User's Own Expertise
- 130. Alibaba Tests Show Its New AI Model Rivals Top Competitors
- 131. GPT-5.6 Helps Two Teams Solve Quantum Crypto Puzzle Within Hours
- 132. Alibaba's Qwen3.8-Max Writes 7,600 Lines of Code in Five Days
- 133. OpenAI's Astra Solves 10 Long-Standing Math Problems
- 134. Meta AI’s Memory Coach Outperforms Constant Recall for Long Tasks
- 135. AI tools flag thousands of flaws, but few get weaponized
- 136. AI Deletes Spreadsheet Data When Asked to Clean Entry
- 137. AI Coding Agents Speed Tasks but Can't Verify Science
- 138. Chinese AI Researchers Turn to X for Technical Audience
- 139. Google DeepMind's Gemini AI now controls entire humanoid robots
- 140. Apple CEO Tim Cook Suggests Possible Paid iCloud Tier for AI Features
- 141. Google DeepMind Demos AI Orchestrating Boston Dynamics Spot Robot
- 142. Ex-OpenAI Researcher Sees USD 100 Billion Bet on Training Data Beyond Scaling
- 143. Hugging Face breach shows "reasonable measures" amid noisy OpenAI hack
- 144. OpenAI Restricted AI Model Access After Hugging Face Breach
- 145. 2025 Study Finds AI Builds Trust Faster Than Human Scammers
- 146. Nimble's New Web Search Agents Cut AI Token Costs by Half
- 147. DeepMind AlphaFold team disbands as researchers depart for Anthropic, Isomorphic
- 148. Hugging Face Traces 17,600 Actions by Compromised AI Models
- 149. Google Expands SynthID Watermark to Label AI Content
- 150. AI Leaders Call for Global Coordination on Automated Research
- 151. Artists sue over AI training, citing emotional harm from data use
- 152. Microsoft's AI Agents Support 24,000 Employees, Drive 70% Efficiency Gains
- 153. Amazon Scales Back Nova AI Models, Bets on New Frontier Team
- 154. New AI Cost Metric Finds Human Labor Still Cheaper by USD 250,000
- 155. Six-Agent DreamTeam Architecture Coordinates for Higher Model Performance
- 156. Cursor Claims Kimi K2.5 Model Shows Cheaper AI Can Code With Frontier Model Planning
- 157. OpenAI Models Escaped Containment for Days in Hugging Face Breach
- 158. Claude Opus 5 cheaper than Fable 5 but still trails on fact accuracy
- 159. South Korea Charts AI Future With NVIDIA at Summit
- 160. Instella-MoE Language Model Improves to 73.22 Score After Post-Training
- 161. Security researcher says AI guardrails don't impede his offensive work
- 162. Multi-turn attacks break AI models 88% of the time, Cisco warns
- 163. Gigatoken BPE Encoder Hits 24.53 GB/s, Up to 989x Faster Than HuggingFace
- 164. Naval Postgraduate School Activates NVIDIA AI Supercomputer for In-House Training
- 165. Britain's AI safety tests find models 'cheating' on cybersecurity evaluations
- 166. AMD's Spur Scheduler Adds Kubernetes Operator for HPC and AI Jobs
- 167. AMD commits USD 5 billion to Anthropic for AI chip deployment
- 168. AI Breached OpenAI Research, Reached Internet via Lateral Movement
- 169. Nvidia's Blackwell Chips Reportedly Overheated in Server Racks
- 170. Alibaba's Qwen Audio 3.0 TTS Plus Leads Rankings Over Gemini, Sonic
- 171. Army AI Tool Ask Sage Used to Reclassify Personnel Job Descriptions
- 172. Microsoft adds AMD-powered Azure HXv2 for complex chip design workloads
- 173. NVIDIA TAO Agent Skills Accelerate Vision Model Post-Training
- 174. Meta Shuts Down AI Token Leaderboard Amid Rising Cost Concerns
- 175. DeepMind CEO Advocates Guardrails Amid AI Uncertainty
- 176. German AI Consortium's Soofi S Hits Benchmarks, Generates 8x Faster
- 177. NVIDIA Toolkit Accelerates OpenFold3 Co-Folding Workflow
- 178. Hanns Christoph Nägerl’s team finds quantum heating defies classical intuition
- 179. AI-Run Ransomware Attack Still Required Human Involvement
- 180. New Research Shows Why Agent Rankings Change After Accounting for Competition
- 181. Auto-FL-Research Uses Agents to Automate Federated Learning Algorithm Search
- 182. 60% of Experts Say Humanity's Last Exam Is Necessary and Useful
- 183. Study Evaluates AI Retrieval Techniques for Finding Models Across Formats
- 184. Researchers unveil RSEA, a three‑layer self‑evolving language agent
- 185. Automate Web Research and Brief Writing with a Python Project from 2026 Guide
- 186. Meta AI launches Brain2Qwerty v2, MEG pipeline hits 61% word accuracy
- 187. Birkhoff’s 1930s ‘measure’ and AICAN’s ‘novelty’ probe AI aesthetics
- 188. MiniMax Token Plan offers extensive coding model access for USD 20/month
- 189. Sina's VibeThinker-3B probes limits, shows reasoning compresses, knowledge weak
- 190. MRAgent beats RAG, A-MEM, MemoryOS, LangMem, Mem0 with 118K tokens/query
- 191. OpenAI unveils Jalapeño custom inference chip, challenging Nvidia's AI dominance
- 192. Prerequisites for NVIDIA AI‑Q Blueprint on OCI: Cluster and Volume Limits
- 193. Episode 11 Explores Overfitting as RAG Evaluation Scores Keep Rising
- 194. LLM pipeline compares DAO ERC‑8004 and Google A2A governance, 4,323 records
- 195. Calibration uses NVIDIA Triton Llama-3-8B A10 and vLLM Qwen2.5-7B RTX 4090 data
- 196. Figma launches AI motion graphics, shader tools, code layers, and new creative materials
- 197. NVIDIA RTX PRO 4500 Blackwell GPUs Power New Amazon EC2 G7 Instances
- 198. Stanford researchers present agentic AI 'scientists' at VB Transform 2026
- 199. Turning Logistic Regression Coefficients into Credit Score Grid
- 200. Metric-Dependent Annotation Saturation for Learning from Label Distributions