News about foundation models, LLMs, multimodal models, benchmarks, and model releases from OpenAI, Anthropic, Google, Meta, Mistral, and others.
987 articles
Importance:ResearchLLM memory architecture
MeMo's memory model lets teams upgrade their LLM without retraining it — and performance jumps 26%
Researchers' MeMo keeps AI memory separate from reasoning, so teams can upgrade their LLM without retraining it and see a 26% performance gain,... Source: venturebeat.com
Importance:NewsLLM inference economics
Databricks’ Model Units Redefine LLM Inference Economics, But Can Reliability Scale?
Databricks' Model Units slash GPU costs by 80% while scaling LLM inference reliability. Discover how this new abstraction reshapes enterprise AI. Source: futurumgroup.com
Importance:NewsLLM security/cybersecurity
AI vs AI Cybersecurity: Sysdig Documents First LLM-Agent Intrusion in the Wild
Security professionals have spent two decades defending against human attackers who use automation as a force multiplier. That model is obsolete. Source: techtimes.com
Importance:ResearchLLM applications in pharmacology
Predicting Drug Side Effects via LLM Pharmacology
In an era when artificial intelligence continues reshaping the landscape of biomedical research, a new study promises to transform drug safety evaluation by... Source: bioengineer.org
Importance:LaunchMeta open-source protein structure model
Biology's Major Change: Zuckerberg's New Open-Source Model Dethrones Google's AlphaFold
AlphaFold has long dominated protein AI, but it has been directly defeated. Zuckerberg's Biohub released ESMFold2, which can predict 1.1 billion protein... Source: eu.36kr.com
Importance:Newsembodied AI / robotics
Alibaba, Tencent lead pivot from chatbots to embodied AI for robotics
Embodied AI and autonomous agents are increasingly viewed by investors as one of the next major growth areas for the sector, UBS says. Source: amp.scmp.com
Importance:NewsLLM security/adversarial use
Attackers Use LLM Agent for Post-Exploitation After Marimo CVE-2026-39987 Exploit
An unknown threat actor has been observed using a large language model (LLM) agent to conduct post-compromise actions after obtaining initial access... Source: thehackernews.com
Importance:ResearchLLM performance optimization
MIT’s MeMo framework boosts LLM performance by 26% without retraining
MIT's MeMo framework trains a compact memory model that boosts LLM performance by up to 26.73% without retraining, with major implications for crypto AI... Source: cryptobriefing.com
Importance:NewsLLM infrastructure
Virtual AI testbed lets developers verify massive LLM servers before construction
Operating large language model (LLM) services like ChatGPT requires a server infrastructure on the scale of tens of thousands of units. Source: techxplore.com
Importance:LaunchAnthropic Claude model release & funding
As Anthropic launches Claude Opus 4.8, it raises $65B in new funding
Anthropic PBC today introduced a new large language model, Claude Opus 4.8, that's significantly better than its predecessor at complex coding tasks. Source: siliconangle.com
Importance:OpinionLLM understanding/evaluation
AI may know the answers but it, as yet, does not understand the questions
Large-scale models can be evaluated across diverse behavioural tasks - but what do they actually understand? Source: digitaljournal.com
Importance:NewsMiniMax M3 model launch
MiniMax Prepares to Launch Next-Generation M3 Large Language Model
Chinese AI unicorn MiniMax is preparing to launch its M3 large language model featuring a custom sparse attention mechanism, claiming 9.7x prefilling speed... Source: pandaily.com
Importance:ResearchLLM agent memory architectures
FluxMem: Dynamic Memory for LLM Agents
FluxMem revolutionizes LLM agent memory, treating it as a dynamic, evolving graph to achieve state-of-the-art performance in complex environments. Source: startuphub.ai
Importance:ResearchLLM reliability/medical AI
The ‘Suggestible’ Orthopaedic Large Language Model
It was right 78% of the time. Then the user confidently handed it the wrong hint, and the large language model (LLM) followed the misdirection. Source: ryortho.com
Importance:NewsLLM API pricing trends
Major LLM providers cut API prices as CoreWeave expands
Leading large language model providers, including OpenAI, Google, Anthropic, xAI, and DeepSeek, have sharply reduced API pricing amid intensifying... Source: msn.com
Importance:ResearchMultimodal visual segmentation
Let the Large Model "View and Modify Simultaneously": Visual Segmentation Accuracy Soars by 9%
In the era of agents, how to make visual segmentation more accurate? Fudan University and Chuangzhi College jointly launched RSAgent, providing the latest... Source: eu.36kr.com
Importance:ResearchLLM knowledge update methods
MEMO: A Modular Framework for Training a Dedicated Memory Model on New Knowledge Without Modifying LLM Parameters
Large language models become static after pretraining. Their knowledge does not update as the world changes. Retraining a full LLM is too expensive at... Source: marktechpost.com
Importance:NewsLLM pricing competition
China's LLM price war puts DeepSeek, Xiaomi on a collision course with OpenAI
China's large language model (LLM) market is coming under mounting pricing pressure, as domestic AI model developers cut fees closer to cost, while the gap... Source: digitimes.com
Importance:NewsAnthropic Claude — security/vulnerability detection
Anthropic: Claude Mythos identified 10,000+ software flaws
Claude Mythos identified thousands of high- and critical-severity vulnerabilities, Anthropic revealed in a Project Glasswing update. Source: helpnetsecurity.com
Importance:NewsLLM pricing wars — DeepSeek/Doubao
Collective Price Hikes: Large Language Models Begin Demanding "Money"
DeepSeek Achieves Legendary Status through Price - Cutting, while Doubao Faces Criticism for Charging Fees! Large Models Caught in Price War. Source: eu.36kr.com