📂 Category
LLMs & Generative AI News Archive - Page 2 of 11
1,086 articles in this category • Page 2 of 11
- 101. Anthropic receives US approval to relaunch Claude Mythos 5 model
- 102. Routing Layer Cut AI Costs but Dropped Customer Satisfaction Scores
- 103. New Methods Let LLMs Auto‑Search Knowledge Bases, Replacing Manual Checks
- 104. GPT‑5.6 Sol outperforms GPT‑5.5 on GeneBench v1 genomics benchmarks
- 105. ByteDance's iLLaDA Diffusion Model Generates Text 4× Faster, Scores Lower on MMLU
- 106. AlgoEvolve uses LLMs to evolve and evaluate Python trading strategies
- 107. Company retains account, IP, session data despite “temporary” AI chats
- 108. KRAFTON’s PUBG Ally uses NVIDIA ACE TTS and behavior trees for real‑time play
- 109. Physics‑Guided CNN Predicts Phase‑Separation Evolution in Binary Mixtures
- 110. OpenAI postpones GPT‑5.6 rollout after Trump administration request
- 111. Meta says AI moderators make 13% fewer errors than humans, defends rollout speed
- 112. NVIDIA TensorRT Enables Context Parallelism for Multi‑GPU AI Inference
- 113. TokenSpeed-Kernel Delivers Top Performance on AMD GPT-OSS 120B via Gluon Kernels
- 114. OpenAI and Deepseek chatbots remain left‑leaning despite anti‑woke push
- 115. MiniCPM‑o 4.5 powers image understanding, captioning and text‑to‑image generation
- 116. Google adds screen-control to Gemini 3.5 Flash for cross‑platform agents
- 117. LLM embeddings and HDBSCAN cluster text; visualized with pairwise scatterplots
- 118. AI Agents Risk Fatal Traps When Treating Context Windows as Memory
- 119. Two-Stage RAG Pipeline Uses Initial LLM Call to Match TOC Sections
- 120. Harness-1 20B Model Beats GPT-5.4, Curates Top 8 Fairness‑Rated Results
- 121. DFlash drafts whole token blocks, achieving 15× throughput on NVIDIA Blackwell
- 122. RIFT-Bench Introduces Graph-Driven Dynamic Red-Teaming for Agentic AI
- 123. Survey of AI Agents: Descartes, Sci‑Fi Roots, and Current Architectures
- 124. Correlated errors cut panel accuracy 8‑22 points; top judge matches panel
- 125. OpenAI's GPT-5.5-Cyber Beats Anthropic Mythos, Starts Patching Initiative
- 126. Pull Gemma4:e4b with Ollama to Build a Local AI Coding Agent (v9.6)
- 127. Anthropic, Micron to Design AI Memory Architecture for Performance, Efficiency
- 128. Sakana's Fugu multi-model hits frontier performance, cites geopolitical edge
- 129. Guide to Using Claude Code for Browser Navigation and Its Simple Mechanics
- 130. Combining hidden neurons still yields a line, highlighting activation's role
- 131. Three NLTK tricks, including MWETokenizer, preserve domain terms in NLP
- 132. Sakana AI Fugu Ultra aims to match models; base Fugu low‑latency coding and chat
- 133. Samsung Deploys ChatGPT and Codex in Software, Marketing, Product, Manufacturing
- 134. Retrieval quality quickly becomes bottleneck for parametric memory’s long‑term weights
- 135. AI agents pick tools using function and parameter descriptions, study shows
- 136. Tip: Ask Clarifying Questions First to Refine ChatGPT Prompts
- 137. Altman says researchers underestimated scaling, calls LeCun's LLM view a dead end
- 138. Convert FP16 LLM to 4‑bit Q4_K_M on Windows AMD Radeon GPUs via llama.cpp
- 139. IEEE launches five‑course online program on large language models
- 140. CUDA Kernel Keeps Corpus on GPU, Cutting Retrieval Latency in RAG
- 141. Anthropic adds live dashboards, Cloudflare‑compatible code to Claude
- 142. Gemma-2-2B-Instruct with Llama-3.1-8B-Instruct cuts 99.9 tokens on 248‑prompt test
- 143. Modeling multi-agent deliberation as closed-loop system with hidden anchors
- 144. Lightweight model cuts RMSE in meteorology, carbon flux, soil moisture, grids
- 145. Understanding JSON Mode, Function Calling, and Structured Output in LLMs
- 146. Audit AI tools: inventory, VPN/zero-trust, continuous fingerprinting
- 147. Adobe adds AI Assistant to Photoshop, Premiere, Illustrator, InDesign, Frame.io
- 148. Claude Fable (Mythos) 5 shows limited bug‑finding and refactoring aid
- 149. Helion adopts LFBO with on‑the‑fly Random Forest for autotuning
- 150. NAVI‑Orbital performs first in‑orbit autonomous vision‑language inference
- 151. TurboQuant and OSCAR vie in KV cache compression race at ICLR 2026
- 152. Study probes if language models can hypothesize new math structures
- 153. NVIDIA XR AI Enables Real‑Time Multimodal Agents for AR Glasses
- 154. PrologMCP Launches as Task-Agnostic Open-Source Server for LLM Agents
- 155. Reconfigure OpenClaw on Mac Mini to Deploy a Local LLM Model
- 156. Roadmap to LLM Engineer in 2026: Foundations, Prompting, Fine‑Tuning, Alignment
- 157. Attention output GEMM reduces blended Fprop speedup to 1.47× in NVFP4 training
- 158. Estonian institute benchmarks AI models' vulnerability to Russian propaganda
- 159. Study quantifies AI agent trust formation, breakage, recovery in survival game
- 160. UP‑NRPA Allows Dynamic Customization of Dialogue Strategies Without Offline RL
- 161. Mobile NPU powers on‑device diffusion LLM with Multi‑Block Speculative Decoding
- 162. Orchestra‑o1 Enables Efficient Omnimodal Agent Collaboration
- 163. Vision LLMs Expand PDF Parsing to Charts, Diagrams, and Tables
- 164. Claude Fable 5 beats GPT‑5.5 by 13 points on FrontierMath tier‑4 tests
- 165. German Court Holds Google Liable for False AI-Generated Overviews
- 166. Google's DiffusionGemma: open diffusion model for faster text generation
- 167. Google sues Chinese Outsider Enterprise for Gemini-driven phishing on Telegram
- 168. PersonaDrive conditions VLA agents on human driving demos for simulation
- 169. ToolSense Framework Audits LLM Tool Knowledge Beyond Constrained Decoding
- 170. Gemini Omni adds AI video generation, using compute limits based on complexity and size
- 171. Xiaomi's MiMo Code beats Claude Code on 200+ step tasks, free MiMo Auto to V2.5
- 172. OpenAI hires Sottiaux in 2024, shifts from internal tools to ChatGPT overhaul
- 173. Low Kruskal-Rank Adaptation Shows Matrix Rank Stays r, Kruskal Rank Falls to 1
- 174. Anthropic apologizes for invisible guardrails on Claude Fable, first Mythos model
- 175. AI pre‑mediation matched professional mediators in multi‑issue negotiation test
- 176. AVLLMs Mirror VLM and VideoLLM Sequential Flow in Audio‑Visual Tasks
- 177. vLLM uses custom GPU kernels, TorchInductor and CUTLASS for portable inference
- 178. Claude Fable declines basic biology queries; Opus 4.8 responds
- 179. Run DiffusionGemma on NVIDIA GPUs for high‑throughput text generation
- 180. SynIB Introduces Information Bottleneck to Boost Multimodal Synergy
- 181. Understanding AgentOps: Discipline and the agentops.ai Platform Explained
- 182. Grab, CJ ENM, LiveKit praise Gemini 3.5 Live Translate for quality and accuracy
- 183. Apple's top AI concept mirrors vibe coding, using Shortcuts as a model
- 184. CoCoNuT paradigm expands residual stream for latent‑space, multi‑path reasoning
- 185. OmniMem adds modality-aware memory allocation for audio‑visual LLMs
- 186. PathoSage Introduces Three‑Stage Framework for Patch‑Level Pathology Reasoning
- 187. Apple unveils third‑gen foundation model, AFM 3 Cloud shows 36% boost
- 188. NVFP4 recipe speeds JAX/MaxText training on NVIDIA Blackwell and Rubin
- 189. Weaker LLMs Accidentally Delete Content, Shrinking Documents Over Time
- 190. Four New Specific Techniques to Boost Productivity with Claude Code
- 191. Jensen Huang sees token market segmenting into distinct value tiers
- 192. OpenAI to revamp ChatGPT, shift to business customers, rival Anth
- 193. MLP Networks Fit High-Frequency Functions One Oscillation at a Time
- 194. SafeGene Introduces Reusable Safety-Adapter for Cross-Task Model Families
- 195. FAIR-Calib Introduces Two-Stage PTQ Framework for Diffusion LLM Quantization
- 196. Elmes* Automates Fine-Grained Rubric Building for LLMs in Niche Education
- 197. Lean4Agent launches FormalAgentLib to model and verify workflow consistency
- 198. Study Finds No One-Size-Fits-All Strategy for Multi-Agent Communication
- 199. xAI used Anthropic’s Claude via personal accounts after access revoked for months
- 200. Study examines temporal preference concepts in large language models