News about foundation models, LLMs, multimodal models, benchmarks, and model releases from OpenAI, Anthropic, Google, Meta, Mistral, and others.
987 articles
Importance:NewsLLM industry — China vs US
Former Tencent AI lead says Chinese firms are trailing US on core LLM innovation
A former Tencent large-model technical chief warned that Chinese artificial intelligence companies lack independent paradigm-level breakthroughs and are... Source: digitimes.com
Importance:Newsquantum AI / LLM hardware
A quantum AI test on IBM hardware points to a new compute race
A quantum-enhanced language model has done something small but important: it answered questions its base model missed. That does not make quantum AI... Source: startupfortune.com
Importance:ResearchLLM usage analysis
Grok in the Wild: Characterizing the Roles and Uses of Large Language Models on Social Media
Authors. Katelyn Xiaoying Mei University of Washington; Robert Wolfe Rutgers University University of Washington; Nic Weber University of Washington... Source: ojs.aaai.org
A preprint submitted to arXiv (arXiv:2605.23035) by Dongxin Guo and colleagues presents a mechanistic interpretability approach connecting large language... Source: letsdatascience.com
Importance:NewsLLM local inference/hardware
Kimi K2.5 runs on RTX 3060 with 768GB Intel Optane memory at 4 tokens per second
A Chinese AI enthusiast known as APFrisco demonstrated Moonshot AI's Kimi K2.5 model, a Mixture-of-Experts (MoE) large language model with 1 trillion total... Source: tradingview.com
Importance:Researchbrain-LLM alignment
Training Data Drives Brain-LLM Alignment Across Languages
An arXiv preprint by Dongxin Guo et al., posted 21 May 2026, reports that brain-LLM alignment was measured using fMRI data from **112 participants** across... Source: letsdatascience.com
Importance:Launchvoice/multimodal model release
StepFun Releases StepAudio 2.5 Realtime: An End-to-End Voice Model with Roleplay-Specific RLHF and Paralinguistic Comprehension
StepFun, the Shanghai-based AI lab, released StepAudio 2.5 Realtime. It is an end-to-end real-time speech large language model with fully customizable... Source: marktechpost.com
Importance:ResearchLLM efficiency
How ‘slimmed-down’ large language models can reduce AI’s environmental and energy footprint
U of T Engineering researchers examine ways to make the use of language models more resource efficient by replacing their high-precision parameters with... Source: news.engineering.utoronto.ca
Importance:NewsLLM inference optimization
Accelerating LLM Inference with Prompt Caching for Open‑Source Models on Databricks
Large language model (LLM) inference often involves repeated prompts—think of the same system or instruction prompt appearing in thousands of requests. Source: databricks.com
VSAS-Bench: Real-Time Evaluation of Visual Streaming Assistant Models
Streaming vision-language models (VLMs) continuously generate responses given an instruction prompt and an online stream of input frames… Source: machinelearning.apple.com
Importance:Opinionmultimodal LLMs overview
The Rise Of The Multimodal LLM
AI leaders discussed multimodal systems, sensory computing, privacy risks, robotics, and future human-machine collaboration possibilities. Source: forbes.com
Importance:ResearchLLMs in healthcare/clinical benchmarks
Evaluating large language models for diagnostic reasoning from unstructured clinical narratives in epilepsy
Large Language Models (LLMs) have been shown to encode clinical knowledge. Many evaluations, however, rely on structured question-answer benchmarks,... Source: nature.com
Importance:LaunchGoogle DeepMind/Gemma security model
Google DeepMind Features Hirundo’s Security-Hardened Gemma 4 Model – Outperforms LLMs 170x Its Size on Security
TEL AVIV, Israel--(BUSINESS WIRE)--May 21, 2026-- Google DeepMind [https://cts.businesswire.com/ct/CT?id=smartlink&url=https%3A%2F%2Fdeepmind.google% Source: venturebeat.com
Importance:NewsAlibaba model update and AI chip
Alibaba Expands AI Push With Model Update, New Chip
Alibaba Group Holding Ltd. has unveiled a new AI chip and an update to its flagship large language model, as the company expands its offerings for AI... Source: caixinglobal.com
Importance:Newsdistributed LLM training
SN9 enables large-scale AI model training using IOTA architecture
Bittensor's Subnet 9 launches IOTA architecture for distributed AI model training, enabling billion-parameter LLMs through collaborative mining pipelines. Source: cryptobriefing.com
Importance:Newsmodel release
Alibaba unveils new Qwen model, custom chips in bid to become China’s AI factory
Chinese tech giant launches Qwen3.7-Max, its new large language model designed as a robust foundation for AI models. Source: scmp.com
Importance:Newsmodel release
Harvard Data Trained This AI Model
Talkie” is a large language model trained on only pre-1931 public domain content from Harvard libraries. Source: harvardmagazine.com
Importance:Newssmall language models
Australia leads rise of the small language model
While large language models have dominated mainstream artificial intelligence since ChatGPT arrived in 2023, their smaller equivalents are about to take... Source: afr.com
Importance:NewsAlibaba AI chip and LLM release
Alibaba Unveils New AI Chip, Upgrades AI Model
The company is pursuing its AI ambitions on multiple fronts. Source: wsj.com
Importance:NewsAlibaba Zhenwu chip and Qwen LLM
Alibaba reveals more powerful Zhenwu AI chip, new LLM
Alibaba announced updates to its AI offerings, including a more powerful chip and a new large language model. Source: cnbc.com