📂 Category
LLMs & Generative AI News Archive - Page 6 of 13
1,263 articles in this category • Page 6 of 13
- 501. GraphDC Uses Divide‑and‑Conquer Agents to Scale Graph Reasoning
- 502. RateQuant reveals mixed-precision KV cache pitfall: β decay rates span 3.6‑5.3
- 503. Top 10 2026 LLM Papers Highlight Pass@k Efficiency for Reasoning Models
- 504. Generative AI fuels industrial-scale record 2025 data breaches, ITRC reports
- 505. Strain drives exponential error growth; vorticity only linear impact
- 506. LKV learns head-wise budgets and token selection for LLM KV cache eviction
- 507. LLM Summarizers Omit Identification, Distinguish Observed vs Inferred Claims
- 508. NVIDIA's Star Elastic bundles 30B, 23B, 12B models; 23B hits 85.63 on AIME-2025
- 509. Understanding 'Compute': The Core Power Driving Modern AI Models
- 510. Fields Medalist: ChatGPT 5.5 Pro produced PhD-level math proof in under an hour
- 511. Key Topics for LLM Engineers: Using Instruction Data to Align Models
- 512. Semantic memory query retrieves Friday deployment approval for user-123
- 513. OpenAI launches Realtime‑Translate for 70+ languages and Realtime‑Whisper transcription
- 514. Google's Chrome 4GB on-device AI model unchanged, but explanation lacking
- 515. RVPO boosts HealthBench score to 0.261, beating GDPO’s 0.215 at 14B (p < 0.001)
- 516. Adding Tools and Memory Expands AI Agent Threat Surface, Study Finds
- 517. Grammar-Constrained Decoding Adds 495 Bash Tasks, Some Regressions Seen
- 518. SAT ensures improvement and plug‑and‑play upgrades in multi‑LLM training
- 519. Transfer Learning Boosts Efficiency of Physics-Informed Neural Networks
- 520. GPT‑5.5 and GPT‑5.5‑Cyber Scale Trusted Access for Cyber Defense
- 521. Google's Gemini Nano auto-uninstalls on low-resource Chrome devices
- 522. ChatGPT in Excel and Google Sheets updates formulas and summarizes sales
- 523. WPP refines AI model components, gains 10% accuracy with AlphaEvolve
- 524. Major reasoning models converge on a shared “brain” as they better model reality
- 525. AI models follow values when taught reasons, avoid harmful rationales
- 526. Lean‑4 Lyapunov proof confirms controllability, ISS robustness for cyber defense
- 527. arXiv paper adds programmatic context to LLM-based symbolic regression
- 528. Anthropic adds 220,000 GPUs at SpaceX's Colossus-1, doubles Claude Code limits
- 529. Gemini API File Search Enables Text‑Only and Multimodal RAG with Embedding‑2
- 530. Timer-XL Uses TimeAttention to Blend Encoder Scatter and Decoder Zoom
- 531. OpenCode adds plugins for citations, source lists and multi‑search support
- 532. Google AI's MTP Drafters for Gemma 4 cut inference time up to threefold
- 533. ClinicBot Introduces Prioritized Evidence RAG with Verifiable Citations
- 534. Study links emergent misalignment to overlapping feature superposition geometry
- 535. Study Finds Systematic Verification Errors Can Stall or Undermine RLVR Training
- 536. OpsLLM: Domain‑Specific LLM Enables QA and Root‑Cause Analysis for Software Ops
- 537. Apple's iOS 27 to add “Extensions” for on‑demand AI via Siri and Writing Tools
- 538. Stochastic KV Routing Enables Depth‑Wise Cache Sharing in Training
- 539. Amazon adds agentic fine‑tuning to SageMaker for Llama, Qwen, Deepseek, Nova
- 540. LLaMA-Factory Enables UI-Based Fine-Tuning and Multi-Model Support Locally
- 541. Self-Healing Layer Scores Answers in Real Time to Counter RAG Hallucinations
- 542. Study Finds Local Causal Directions Encode Harmfulness in LLM Jailbreaks
- 543. New arXiv paper introduces TADI, an agentic AI system for drilling data
- 544. Agentopic uses multiple agents for identification, validation, and explanations
- 545. Google adds event-driven webhooks to Gemini API, ending polling for long AI jobs
- 546. Cheaper tokens, bigger bills: Agentic workloads test AI infrastructure
- 547. Zuckerberg commits USD 500 million to AI-driven biology research at Meta
- 548. Claude Code, Copilot, Codex hacked; attackers stole credentials, not models
- 549. Artificial Intelligence Shows Skill in Emergency Room Triage Process
- 550. OpenAI activates default marketing cookies for free ChatGPT users
- 551. LlamaIndex CEO: AI scaffolding collapses as models surpass humans on massive data
- 552. ChatGPT's 'Nerdy' tweak rewards goblin metaphors in answers, study finds
- 553. SMG releases smg-grpc-proto on PyPI; vLLM integrates via PR #36169
- 554. Google TV adds Gemini tools, including Nano Banana voice‑prompt image editor
- 555. Google launches Gemini memory in Europe, default on, pulls personal data
- 556. Nvidia's Nemotron 3 Nano Omni: 30B model processes text, images, video, audio
- 557. LLM using pre‑1930 sources draws on etiquette manuals, cookbooks for post‑training
- 558. Alibaba's Metis agent cuts redundant AI tool calls to 2% and boosts accuracy
- 559. Qiushi Discovery Engine Enables Autonomous Science on Optical Platform
- 560. RL Agent Retrieves Relevant Memories to Boost LLM Question Answering
- 561. New Session Details Hardware and Software Methods to Speed Multimodal Models
- 562. ChatGPT Images 2.0 and Nano Banana 2 Produce Professional Results
- 563. vLLM Enables Fast, Memory‑Efficient, High‑Throughput Serving of Open‑Source LLMs
- 564. xAI's grok-voice-think-fast-1.0 leads τ-voice Bench with 67.3%
- 565. Synthetic pipelines speed edge‑case curation for LLM behavior monitoring
- 566. GitNexus indexes repositories into a knowledge graph for code intelligence
- 567. Google Cloud Next ’26 launches Agent Studio and Gemini Enterprise AI app
- 568. OpenAI's 'Spud' Beats Claude; April 30 Webinar on Agentspan 4‑Layer Production
- 569. Why ChatGPT and Other Bots May Mislead You on Financial Advice
- 570. Claude adds direct connectors for Spotify, Uber Eats, TurboTax; mobile beta
- 571. OpenAI launches GPT-5.5, hits 82.7% on Terminal-Bench 2.0, 84.9% on GDPval
- 572. Google launches TPU 8t for high‑throughput training, TPU 8i for memory bandwidth
- 573. Google unveils dual high‑powered TPUs, sidestepping Nvidia tax for enterprises
- 574. X to let Grok personalize timelines based on selected topics for each user
- 575. OpenAI introduces workspace agents that autonomously report product feedback
- 576. Google Meet adds AI notes, summaries and transcripts to in‑person meetings
- 577. OpenAI regains image lead as Algolia releases AI agent guide
- 578. AI backlash surges as politicians finally grasp public sentiment
- 579. OpenAI upgrades ChatGPT image model, improves English text rendering
- 580. Starbucks ChatGPT app forces users to pick from suggested iced‑coffee options
- 581. OpenCode now supports Qwen3-Coder via config.json on Linux, macOS, Windows
- 582. Yelp expands AI chatbot to new Assistant tab across all categories
- 583. Qwen 3.6-35B-A3B Demo Implements Multimodal Inference, Thinking Control and RAG
- 584. Microsoft’s Phi-4-Mini 3.8B-Parameter Used in RAG Pipeline with LoRA Fine‑Tuning
- 585. OpenAI expands Trusted Access for Cyber Defense with GPT-5.4‑Cyber model
- 586. Moonshot AI, Tsinghua unveil PrfaaS KVCache that auto‑balances LLM nodes for throughput
- 587. Anthropic's Claude Opus 4.7 lifts coding benchmark 13% and solves four new tasks
- 588. OpenAI API guide demonstrates gpt-4o call, returning 'Late 2024-early 2025
- 589. Microsoft’s MarkItDown library converts zip files, unifying supported content
- 590. NVIDIA KVPress Enables Long‑Context LLM Inference with KV Cache Compression
- 591. Complete Real-World Example Shows Crawl4AI CSS Extraction and Filtering
- 592. Implementing Context-Aware Long-Term Memory for AI Agents via Mem0 and OpenAI
- 593. Transformer-Based Neural Quantum States for Frustrated Spins Using NetKet
- 594. Schematik ‘Cursor for Hardware’ secures USD 4.6M Lightspeed; Anthropic wants in
- 595. Google AI launches Auto-Diagnose, LLM tool flags 84.3% of reports as ‘Please fix’
- 596. Penligent and Giskard among top AI red‑team tools for model security
- 597. AI protein-design tools offer flexible workflows for any protein class
- 598. Google's AI mode will open linked pages beside search on Chrome desktop
- 599. Physical Intelligence robot model shows LLM-like skill composition, flaws noted
- 600. OpenAI's Codex powers Lovable AI, letting millions create apps from text