AIskimIQ

Daily AI & tech news brief

Archive/large language models

🧠 Large Language Models

News about foundation models, LLMs, multimodal models, benchmarks, and model releases from OpenAI, Anthropic, Google, Meta, Mistral, and others.

987 articles

Importance:ResearchLLM memory architecture

MeMo's memory model lets teams upgrade their LLM without retraining it — and performance jumps 26%

Researchers' MeMo keeps AI memory separate from reasoning, so teams can upgrade their LLM without retraining it and see a 26% performance gain,... Source: venturebeat.com

Importance:NewsLLM inference economics

Databricks’ Model Units Redefine LLM Inference Economics, But Can Reliability Scale?

Databricks' Model Units slash GPU costs by 80% while scaling LLM inference reliability. Discover how this new abstraction reshapes enterprise AI. Source: futurumgroup.com

Importance:NewsLLM security/cybersecurity

AI vs AI Cybersecurity: Sysdig Documents First LLM-Agent Intrusion in the Wild

Security professionals have spent two decades defending against human attackers who use automation as a force multiplier. That model is obsolete. Source: techtimes.com

Importance:ResearchLLM applications in pharmacology

Predicting Drug Side Effects via LLM Pharmacology

In an era when artificial intelligence continues reshaping the landscape of biomedical research, a new study promises to transform drug safety evaluation by... Source: bioengineer.org

Importance:LaunchMeta open-source protein structure model

Biology's Major Change: Zuckerberg's New Open-Source Model Dethrones Google's AlphaFold

AlphaFold has long dominated protein AI, but it has been directly defeated. Zuckerberg's Biohub released ESMFold2, which can predict 1.1 billion protein... Source: eu.36kr.com

Importance:Newsembodied AI / robotics

Alibaba, Tencent lead pivot from chatbots to embodied AI for robotics

Embodied AI and autonomous agents are increasingly viewed by investors as one of the next major growth areas for the sector, UBS says. Source: amp.scmp.com

Importance:NewsLLM security/adversarial use

Attackers Use LLM Agent for Post-Exploitation After Marimo CVE-2026-39987 Exploit

An unknown threat actor has been observed using a large language model (LLM) agent to conduct post-compromise actions after obtaining initial access... Source: thehackernews.com

Importance:ResearchLLM performance optimization

MIT’s MeMo framework boosts LLM performance by 26% without retraining

MIT's MeMo framework trains a compact memory model that boosts LLM performance by up to 26.73% without retraining, with major implications for crypto AI... Source: cryptobriefing.com

Importance:NewsLLM infrastructure

Virtual AI testbed lets developers verify massive LLM servers before construction

Operating large language model (LLM) services like ChatGPT requires a server infrastructure on the scale of tens of thousands of units. Source: techxplore.com

Importance:LaunchAnthropic Claude model release & funding

As Anthropic launches Claude Opus 4.8, it raises $65B in new funding

Anthropic PBC today introduced a new large language model, Claude Opus 4.8, that's significantly better than its predecessor at complex coding tasks. Source: siliconangle.com

Importance:OpinionLLM understanding/evaluation

AI may know the answers but it, as yet, does not understand the questions

Large-scale models can be evaluated across diverse behavioural tasks - but what do they actually understand? Source: digitaljournal.com

Importance:NewsMiniMax M3 model launch

MiniMax Prepares to Launch Next-Generation M3 Large Language Model

Chinese AI unicorn MiniMax is preparing to launch its M3 large language model featuring a custom sparse attention mechanism, claiming 9.7x prefilling speed... Source: pandaily.com

Importance:ResearchLLM agent memory architectures

FluxMem: Dynamic Memory for LLM Agents

FluxMem revolutionizes LLM agent memory, treating it as a dynamic, evolving graph to achieve state-of-the-art performance in complex environments. Source: startuphub.ai

Importance:ResearchLLM reliability/medical AI

The ‘Suggestible’ Orthopaedic Large Language Model

It was right 78% of the time. Then the user confidently handed it the wrong hint, and the large language model (LLM) followed the misdirection. Source: ryortho.com

Importance:NewsLLM API pricing trends

Major LLM providers cut API prices as CoreWeave expands

Leading large language model providers, including OpenAI, Google, Anthropic, xAI, and DeepSeek, have sharply reduced API pricing amid intensifying... Source: msn.com

Importance:ResearchMultimodal visual segmentation

Let the Large Model "View and Modify Simultaneously": Visual Segmentation Accuracy Soars by 9%

In the era of agents, how to make visual segmentation more accurate? Fudan University and Chuangzhi College jointly launched RSAgent, providing the latest... Source: eu.36kr.com

Importance:ResearchLLM knowledge update methods

MEMO: A Modular Framework for Training a Dedicated Memory Model on New Knowledge Without Modifying LLM Parameters

Large language models become static after pretraining. Their knowledge does not update as the world changes. Retraining a full LLM is too expensive at... Source: marktechpost.com

Importance:NewsLLM pricing competition

China's LLM price war puts DeepSeek, Xiaomi on a collision course with OpenAI

China's large language model (LLM) market is coming under mounting pricing pressure, as domestic AI model developers cut fees closer to cost, while the gap... Source: digitimes.com

Importance:NewsAnthropic Claude — security/vulnerability detection

Anthropic: Claude Mythos identified 10,000+ software flaws

Claude Mythos identified thousands of high- and critical-severity vulnerabilities, Anthropic revealed in a Project Glasswing update. Source: helpnetsecurity.com

Importance:NewsLLM pricing wars — DeepSeek/Doubao

Collective Price Hikes: Large Language Models Begin Demanding "Money"

DeepSeek Achieves Legendary Status through Price - Cutting, while Doubao Faces Criticism for Charging Fees! Large Models Caught in Price War. Source: eu.36kr.com