📂 Category
LLMs & Generative AI News Archive - Page 3 of 13
1,253 articles in this category • Page 3 of 13
- 201. Experts: Kimi K3's Gains Not From Costly, Slow Frontier Model API
- 202. Cisco’s Small AI Models Outperform Larger Rivals on Cost for Vulnerability Detection
- 203. AMD Opens ROCm KV Cache Tech to Upstream Projects
- 204. AMD Releases Preview of Gemma 3B LLM Container for Radeon GPUs
- 205. Gemini 3.6 Flash Release Tests Expose Model's Reliance on Retrieval
- 206. Gemini 3.6 Flash Boosts Coding and Token Efficiency
- 207. LWiAI Podcast #252: GPT 5.6, Grok 4.5, and AI 2040 Discussed
- 208. Report: US Weighs Ban on Chinese AI Models Amid IP Theft Concerns
- 209. NVIDIA GB300 NVL72 Achieves Record MoE Pre-Training Performance
- 210. Xiaomi's New Robot AI Model Favors More Data Over Bigger Architectures
- 211. MCP Protocol Update Simplifies AI Server Session Management
- 212. Ex-Trump AI Advisor Criticizes China's AI Model Rules
- 213. Adobe’s Indigo app adds AI Playground with generative photo editing
- 214. OpenAI Targets 2027 for First Major Hardware: A ChatGPT Speaker
- 215. Anthropic's Claude for Teachers Vows Not to Train on Student Data
- 216. Superhuman's AI Feature Drafts Email Replies in Your Own Tone
- 217. ChatGPT Returns to WhatsApp in Europe Using GPT-5.5
- 218. New York enacts first US state data center moratorium
- 219. Anthropic Study Shows Claude's Values Shift by Language
- 220. ELIZA's Historical Algorithms Foreshadow Why People Confide in AI
- 221. Google's SensorFM outperforms models on 34 of 35 health data tasks
- 222. OpenAI Launches ChatGPT Work, a Unified AI Agent Powered by GPT-5.6
- 223. Terrorist Groups Use Major AI Chatbots for Attack Planning, Weapons Development
- 224. OpenAI Targets Families as ChatGPT Adapts for Household Use
- 225. Meta's Muse Spark 1.1 coding score hits 71.3, edges past GLM-5.2
- 226. Anthropic's 'Logit Lens' Reveals How Claude Puzzles Over Words
- 227. OpenAI Releases GPT-Live Voice Models That Delegate Reasoning to GPT-5.5
- 228. Bezos-backed startup raises USD 320 million to build AGI from gaming data
- 229. Anthropic Launches Claude Cowork AI Agent for Mobile and Web
- 230. Anthropic Moves Its Claude Agent to Phones as AI Rivals Target Mobile
- 231. DeepSeek Plans Own Chips Amid US Export Controls, Following Huawei, Alibaba
- 232. Trump’s Davos drama, AI‑fuelled midterms draw tens of millions early
- 233. Dynamic Power Boost Compensates for GPU Failures in LLM Training
- 234. 16B-Parameter Diffusion Models Generate Multi-Speaker, Multilingual Speech
- 235. PathMoE Study Shows More Concentrated, Robust Routing Paths
- 236. Five Small, Open-Weight Models Built for Agentic Tool Calling
- 237. Meta Tests GPT-5.5 With 'Watermelon' Model
- 238. GPT-4o, Gemini, and Claude Vision Can Reason Over Visual Details
- 239. AI Search Agents Struggle With Ambiguous Queries, Study Finds
- 240. pxpipe hides text in PNGs to cut Claude token costs by up to 70%
- 241. Midjourney Challenges Studios' AI Document Secrecy in Court Filing
- 242. Run Local AI on 8GB Macs With Smaller Models, Avoid Complex Setup
- 243. Deep Learning AI Models Identify Data Features Without Human Input
- 244. Developer Replaces LLM Wiki With Pure Python Compiler, Citing Over-Engineering
- 245. Wiola Architecture Introduces Five Novel Components for Efficient Small Language Models
- 246. t0-alpha Shows Tight 0.015 CRPS Spread in Time-Series LLM Cluster
- 247. VideoFlexTok's Flow Decoder Enables Variable-Length Video Tokenization
- 248. AI Engineers Face Rising Costs, Need New Strategies for Efficiency
- 249. Square's ChatGPT integration charges restaurants 6% fee for pickup orders
- 250. Anthropic adds security measure; Commerce Dept clears Fable 5 for release
- 251. Claude Sonnet 5 lists unchanged token rates despite double real cost
- 252. Anthropic's Claude Code hides XOR‑encrypted flag for Chinese users in v2.1.91
- 253. Maximizing Codex Exec: Using It as a Code Reviewer with Claude Code
- 254. OpenAI engineers say they halved inference costs for guest ChatGPT users
- 255. NVIDIA BioNeMo Agent Toolkit speeds AI for life‑science researchers
- 256. IMCBench Launches Image‑Grounded Multi‑Turn Medical Conversation Benchmark
- 257. GPTNT Benchmarks Real-Time Collaboration of Multimodal Agents on KTaNE
- 258. Omniverse Workflows Boost Vision AI Accuracy Using Synthetic Data, Fine‑Tuning
- 259. Dynamic Representation Editing Framework Aims to Steer LLM Reasoning Paths
- 260. Hybrid LLM Guide: Local Model Sanitizes Household Data Before Cloud Scheduling
- 261. New Benchmark Assesses AI Text-to-Image and Multimodal Models for Scientific Figures
- 262. Meta hired teen‑posing contractors to test rival chatbots on suicide, sex, drugs
- 263. Google's Gemini offers free Nano Banana AI image generation for US users
- 264. Small models lag in multi‑step reasoning, >128K context, and large‑scale coding
- 265. Researchers Spot Format‑Capability Gap in Post‑Training Look‑Ahead Fine‑Tuning
- 266. DysLexLens: Low‑Resource LLM Turns Forum Posts into Traceable KG Insights
- 267. OpenAI's GPT-5.6 Sol cheats on software tests more than any model, METR says
- 268. Anthropic receives US approval to relaunch Claude Mythos 5 model
- 269. Routing Layer Cut AI Costs but Dropped Customer Satisfaction Scores
- 270. New Methods Let LLMs Auto‑Search Knowledge Bases, Replacing Manual Checks
- 271. GPT‑5.6 Sol outperforms GPT‑5.5 on GeneBench v1 genomics benchmarks
- 272. ByteDance's iLLaDA Diffusion Model Generates Text 4× Faster, Scores Lower on MMLU
- 273. AlgoEvolve uses LLMs to evolve and evaluate Python trading strategies
- 274. Company retains account, IP, session data despite “temporary” AI chats
- 275. KRAFTON’s PUBG Ally uses NVIDIA ACE TTS and behavior trees for real‑time play
- 276. Physics‑Guided CNN Predicts Phase‑Separation Evolution in Binary Mixtures
- 277. OpenAI postpones GPT‑5.6 rollout after Trump administration request
- 278. Meta says AI moderators make 13% fewer errors than humans, defends rollout speed
- 279. NVIDIA TensorRT Enables Context Parallelism for Multi‑GPU AI Inference
- 280. TokenSpeed-Kernel Delivers Top Performance on AMD GPT-OSS 120B via Gluon Kernels
- 281. OpenAI and Deepseek chatbots remain left‑leaning despite anti‑woke push
- 282. MiniCPM‑o 4.5 powers image understanding, captioning and text‑to‑image generation
- 283. Google adds screen-control to Gemini 3.5 Flash for cross‑platform agents
- 284. LLM embeddings and HDBSCAN cluster text; visualized with pairwise scatterplots
- 285. AI Agents Risk Fatal Traps When Treating Context Windows as Memory
- 286. Two-Stage RAG Pipeline Uses Initial LLM Call to Match TOC Sections
- 287. Harness-1 20B Model Beats GPT-5.4, Curates Top 8 Fairness‑Rated Results
- 288. DFlash drafts whole token blocks, achieving 15× throughput on NVIDIA Blackwell
- 289. RIFT-Bench Introduces Graph-Driven Dynamic Red-Teaming for Agentic AI
- 290. Survey of AI Agents: Descartes, Sci‑Fi Roots, and Current Architectures
- 291. Correlated errors cut panel accuracy 8‑22 points; top judge matches panel
- 292. OpenAI's GPT-5.5-Cyber Beats Anthropic Mythos, Starts Patching Initiative
- 293. Pull Gemma4:e4b with Ollama to Build a Local AI Coding Agent (v9.6)
- 294. Anthropic, Micron to Design AI Memory Architecture for Performance, Efficiency
- 295. Sakana's Fugu multi-model hits frontier performance, cites geopolitical edge
- 296. Guide to Using Claude Code for Browser Navigation and Its Simple Mechanics
- 297. Combining hidden neurons still yields a line, highlighting activation's role
- 298. Three NLTK tricks, including MWETokenizer, preserve domain terms in NLP
- 299. Sakana AI Fugu Ultra aims to match models; base Fugu low‑latency coding and chat
- 300. Samsung Deploys ChatGPT and Codex in Software, Marketing, Product, Manufacturing