📂 Category
LLMs & Generative AI News Archive - Page 4 of 11
1,086 articles in this category • Page 4 of 11
- 301. Vercel Labs launches Zero, a systems language for AI agents to read and ship
- 302. Enterprise‑grade AI platform merges chatbot, voice, video; customizable via APIs
- 303. NightCafe Remains Long‑Running, Community‑Focused AI Art Platform
- 304. Claude Mythos USD 36,428 for 122 exploit episodes; GPT‑5.5 USD 3,075 for 123
- 305. Use Automated Dashboards and Weekly Review Cadence for GenAI Interviews
- 306. OpenAI partners with Malta to offer ChatGPT Plus to every citizen
- 307. Tool Highlights Time‑Consuming Stalls and Faulty Calls in Claude Code
- 308. Zyphra launches ZAYA1-8B Diffusion Preview, a MoE model with 7.7× speedup
- 309. Claude targets agent control plane while Microsoft stays enterprise default
- 310. Invisible orchestration raises collective dissociation (g = 0.975, p = .001)
- 311. Automatic alerts trigger when LLM accuracy falls or latency spikes
- 312. Poetiq’s Meta‑System Improves LLMs on LiveCodeBench for Reasoning, Retrieval
- 313. ChatGPT traffic falls to 54% as Gemini climbs to 26.7% in a year
- 314. Inference Systems, Not Models, Emerge as the Next AI Bottleneck
- 315. Alibaba's Qwen-Image-2.0 doubles compression, slashes steps to 4 with Qwen3.5-9B
- 316. B2B Document Extractor Rebuilt: Rule-Based vs. LLM Using pytesseract OCR
- 317. Anthropic adds Claude plugins for CoCounsel, DocuSign, Everlaw, Box, Harvey
- 318. Rubrics-as-Reward seeks explicit criteria; scalable rubrics remain elusive
- 319. Avoid TensorRT Slowdowns or Build Failures by Adding Plugin Extensions
- 320. Audit matrix flags token rotation via npm postinstall hook in Claude Code
- 321. BalCapRL adds length-based reward masking, boosting LLaVA-1.5-7B and Qwen2.5-VL
- 322. SFT and RL Reweight Pretrained Distributions via Demonstration and Reward Signals
- 323. Spatial priming beats semantic prompting in chart data extraction study
- 324. GraphDC Uses Divide‑and‑Conquer Agents to Scale Graph Reasoning
- 325. RateQuant reveals mixed-precision KV cache pitfall: β decay rates span 3.6‑5.3
- 326. Top 10 2026 LLM Papers Highlight Pass@k Efficiency for Reasoning Models
- 327. Generative AI fuels industrial-scale record 2025 data breaches, ITRC reports
- 328. Strain drives exponential error growth; vorticity only linear impact
- 329. LKV learns head-wise budgets and token selection for LLM KV cache eviction
- 330. LLM Summarizers Omit Identification, Distinguish Observed vs Inferred Claims
- 331. NVIDIA's Star Elastic bundles 30B, 23B, 12B models; 23B hits 85.63 on AIME-2025
- 332. Understanding 'Compute': The Core Power Driving Modern AI Models
- 333. Fields Medalist: ChatGPT 5.5 Pro produced PhD-level math proof in under an hour
- 334. Key Topics for LLM Engineers: Using Instruction Data to Align Models
- 335. Semantic memory query retrieves Friday deployment approval for user-123
- 336. OpenAI launches Realtime‑Translate for 70+ languages and Realtime‑Whisper transcription
- 337. Google's Chrome 4GB on-device AI model unchanged, but explanation lacking
- 338. RVPO boosts HealthBench score to 0.261, beating GDPO’s 0.215 at 14B (p < 0.001)
- 339. Adding Tools and Memory Expands AI Agent Threat Surface, Study Finds
- 340. Grammar-Constrained Decoding Adds 495 Bash Tasks, Some Regressions Seen
- 341. SAT ensures improvement and plug‑and‑play upgrades in multi‑LLM training
- 342. Transfer Learning Boosts Efficiency of Physics-Informed Neural Networks
- 343. GPT‑5.5 and GPT‑5.5‑Cyber Scale Trusted Access for Cyber Defense
- 344. Google's Gemini Nano auto-uninstalls on low-resource Chrome devices
- 345. ChatGPT in Excel and Google Sheets updates formulas and summarizes sales
- 346. WPP refines AI model components, gains 10% accuracy with AlphaEvolve
- 347. Major reasoning models converge on a shared “brain” as they better model reality
- 348. AI models follow values when taught reasons, avoid harmful rationales
- 349. Lean‑4 Lyapunov proof confirms controllability, ISS robustness for cyber defense
- 350. arXiv paper adds programmatic context to LLM-based symbolic regression
- 351. Anthropic adds 220,000 GPUs at SpaceX's Colossus-1, doubles Claude Code limits
- 352. Gemini API File Search Enables Text‑Only and Multimodal RAG with Embedding‑2
- 353. Timer-XL Uses TimeAttention to Blend Encoder Scatter and Decoder Zoom
- 354. OpenCode adds plugins for citations, source lists and multi‑search support
- 355. Google AI's MTP Drafters for Gemma 4 cut inference time up to threefold
- 356. ClinicBot Introduces Prioritized Evidence RAG with Verifiable Citations
- 357. Study links emergent misalignment to overlapping feature superposition geometry
- 358. Study Finds Systematic Verification Errors Can Stall or Undermine RLVR Training
- 359. OpsLLM: Domain‑Specific LLM Enables QA and Root‑Cause Analysis for Software Ops
- 360. Apple's iOS 27 to add “Extensions” for on‑demand AI via Siri and Writing Tools
- 361. Stochastic KV Routing Enables Depth‑Wise Cache Sharing in Training
- 362. Amazon adds agentic fine‑tuning to SageMaker for Llama, Qwen, Deepseek, Nova
- 363. LLaMA-Factory Enables UI-Based Fine-Tuning and Multi-Model Support Locally
- 364. Self-Healing Layer Scores Answers in Real Time to Counter RAG Hallucinations
- 365. Study Finds Local Causal Directions Encode Harmfulness in LLM Jailbreaks
- 366. New arXiv paper introduces TADI, an agentic AI system for drilling data
- 367. Agentopic uses multiple agents for identification, validation, and explanations
- 368. Google adds event-driven webhooks to Gemini API, ending polling for long AI jobs
- 369. Cheaper tokens, bigger bills: Agentic workloads test AI infrastructure
- 370. Zuckerberg commits USD 500 million to AI-driven biology research at Meta
- 371. Claude Code, Copilot, Codex hacked; attackers stole credentials, not models
- 372. Artificial Intelligence Shows Skill in Emergency Room Triage Process
- 373. OpenAI activates default marketing cookies for free ChatGPT users
- 374. LlamaIndex CEO: AI scaffolding collapses as models surpass humans on massive data
- 375. ChatGPT's 'Nerdy' tweak rewards goblin metaphors in answers, study finds
- 376. SMG releases smg-grpc-proto on PyPI; vLLM integrates via PR #36169
- 377. Google TV adds Gemini tools, including Nano Banana voice‑prompt image editor
- 378. Google launches Gemini memory in Europe, default on, pulls personal data
- 379. Nvidia's Nemotron 3 Nano Omni: 30B model processes text, images, video, audio
- 380. LLM using pre‑1930 sources draws on etiquette manuals, cookbooks for post‑training
- 381. Alibaba's Metis agent cuts redundant AI tool calls to 2% and boosts accuracy
- 382. Qiushi Discovery Engine Enables Autonomous Science on Optical Platform
- 383. RL Agent Retrieves Relevant Memories to Boost LLM Question Answering
- 384. New Session Details Hardware and Software Methods to Speed Multimodal Models
- 385. ChatGPT Images 2.0 and Nano Banana 2 Produce Professional Results
- 386. vLLM Enables Fast, Memory‑Efficient, High‑Throughput Serving of Open‑Source LLMs
- 387. xAI's grok-voice-think-fast-1.0 leads τ-voice Bench with 67.3%
- 388. Synthetic pipelines speed edge‑case curation for LLM behavior monitoring
- 389. GitNexus indexes repositories into a knowledge graph for code intelligence
- 390. Google Cloud Next ’26 launches Agent Studio and Gemini Enterprise AI app
- 391. OpenAI's 'Spud' Beats Claude; April 30 Webinar on Agentspan 4‑Layer Production
- 392. Why ChatGPT and Other Bots May Mislead You on Financial Advice
- 393. Claude adds direct connectors for Spotify, Uber Eats, TurboTax; mobile beta
- 394. OpenAI launches GPT-5.5, hits 82.7% on Terminal-Bench 2.0, 84.9% on GDPval
- 395. Google launches TPU 8t for high‑throughput training, TPU 8i for memory bandwidth
- 396. Google unveils dual high‑powered TPUs, sidestepping Nvidia tax for enterprises
- 397. X to let Grok personalize timelines based on selected topics for each user
- 398. OpenAI introduces workspace agents that autonomously report product feedback
- 399. Google Meet adds AI notes, summaries and transcripts to in‑person meetings
- 400. OpenAI regains image lead as Algolia releases AI agent guide