AIskimIQ

Daily AI & tech news brief

Archive/large language models

🧠 Large Language Models

News about foundation models, LLMs, multimodal models, benchmarks, and model releases from OpenAI, Anthropic, Google, Meta, Mistral, and others.

981 articles

Importance:Newsmodel training

The Controversy Behind Anthropic Destroying Physical Books to Train Its AI Models

Anthropic has faced criticism over its practice of physically destroying books to extract training data for its large language models, which the company argues offer the highest-quality text available. The case has reignited debate over the ethics and legality of AI training data sourcing. Source: eu.36kr.com

Importance:Launchmodel applications

Meitu Xiuxiu Rolls Out New AI Assistant Powered by Its Own Large Language Model

Chinese photo-editing app Meitu Xiuxiu has introduced a new AI Assistant feature built on its in-house large language model, according to Sina Technology. The update was announced on September 10, expanding the app's AI capabilities for users. Source: news.futunn.com

Importance:News

GPT-6 Astra meets V2Fun: large models begin orchestrating 3D generation

A new wave of AI tools is combining large language models with 3D generation pipelines, letting models coordinate multiple specialized systems instead of doing the work themselves. This shift points to a broader trend of AI acting as an orchestrator across code, image, and 3D modeling workflows. Source: eu.36kr.com

Importance:Launch

Mercury 2.5 diffusion model claims 1,107 tokens/sec at GPT-5.6 Luna-level quality

Inception unveiled Mercury 2.5 on September 8, 2026, a text-generation model built on diffusion technology rather than the typical autoregressive approach. The company says it matches GPT-5.6 Luna's performance while generating up to 1,107 tokens per second at a cost of roughly 120 yen per million output tokens. Source: gigazine.net

Importance:LaunchConsumer AI

Apple's Siri AI debuts in English on September 14, more languages coming in October

Apple confirmed that its overhauled, LLM-powered Siri AI will launch alongside iOS 27 on September 14, initially supporting only English. Five additional languages are set to follow in October as the rollout expands. Source: gagadget.com

Importance:NewsIndustry news

Another Anthropic employee quits, warns about frontier AI labs' direction

An Anthropic staffer known as Cal has publicly announced leaving the company, citing concerns about the path frontier AI labs are taking. The departure adds to a growing list of insiders voicing doubts about how leading AI companies are handling model development. Source: nationalreview.com

Importance:NewsConsumer AI

SCX.ai builds Swift package to tap Apple's on-device LLMs, with Australia in focus

SCX.ai is developing an MIT-licensed Swift package that lets developers connect to Apple's hosted large language models more flexibly. The project has a particular focus on Australian use cases and aims to simplify access to Apple's LLM capabilities. Source: arnnet.com.au

Importance:LaunchNew Model Release

Inception Debuts Mercury 2.5, Claiming 1,100 Tokens per Second

Inception has released Mercury 2.5, a diffusion-based LLM said to hit over 1,100 tokens per second in real-world use. The launch highlights how speed is becoming a new front in the competition among AI model makers. Source: shattered.io

Importance:NewsMedical LLM

South Korea to Build Medical LLM Trained on Local Clinical Standards

South Korea plans to develop a large language model specifically adapted to its domestic healthcare system. The move follows criticism that foreign AI models fail to properly reflect local medical practices and guidelines. Source: mbiz.heraldcorp.com

Importance:ResearchLLM Evaluation

Amazon Study Uses Ising Models to Expose Bias in LLM Judges

New research from Amazon suggests that apparent agreement among LLM-based judges can be deceptive. Using Ising models borrowed from physics, researchers Krishna Balasubramanian and Sasha found hidden bias patterns behind seemingly strong consensus. Source: quantumzeitgeist.com

Importance:NewsAI Economics

AI Model Prices Plunge Faster Than PCs Ever Did

The cost of running AI models has dropped dramatically over just three years, outpacing even the steep price declines seen during the personal computer boom of the 1980s. Analysts point to this rapid cost collapse as a key driver of wider AI adoption. Source: citybiz.co

Importance:NewsLLM Theory

What Chaos Theory Reveals About Large Language Models

LLMs may look like black boxes, but researchers argue the math behind their behavior is well understood. A new deep-dive explores how chaos and complexity theory can explain patterns inside these models. Source: hackernoon.com

Importance:ResearchLLM metacognition

Study finds language models causally rely on confidence to guide behavior

New research explores metacognition in AI systems, showing that confidence signals extracted from language models can causally influence their outputs. The findings suggest models don't just simulate self-assessment but actually use it to shape decisions, similar to biological cognition. Source: nature.com

Importance:NewsLocal LLM deployment

Windows PCs aim to run 100-billion-parameter LLMs locally, challenging Mac's AI edge

After two years of hype around AI PCs, a new effort claims to move local execution of massive language models from demos into real, usable products on Windows machines. The goal is to close the gap with Apple's Mac lineup in on-device AI performance. Source: eu.36kr.com

Importance:Launch290B parameter model

Cipheras Group's Apex AI Fund rolls out proprietary 290-billion-parameter LLM

A private AI-focused investment fund says it has deployed a custom-built 290-billion-parameter large language model. The move positions it among the first investment vehicles to use such a large proprietary model in its operations. Source: lincolnjournal.com

Importance:LaunchGrok 4.5 release

Grok 4.5 debuts with emphasis on coding and AI agents, heating up rivalry with OpenAI and Anthropic

Elon Musk's AI company xAI has released Grok 4.5, the newest version of its flagship large language model. The update targets coding capabilities and agentic AI features, pushing further competition with OpenAI and Anthropic. Source: ibtimes.com

Importance:NewsGemma model performance

Google Cloud says smaller Gemma 3 12B outperforms 27B model on TPU throughput

Google Cloud reports that under heavy load, the smaller Gemma 3 12B model kept scaling efficiently on TPU v6e hardware. In contrast, the larger 27B version hit a throughput ceiling once user load passed 64. Source: itbrief.asia

Importance:NewsOpenAI security incident

Questions raised over whether OpenAI's new model behaved unexpectedly

The article revisits a recent security incident where Hugging Face reported a breach of its production infrastructure. It raises the question of whether an OpenAI model played an unintended role in the episode. Source: predictiveanalyticsworld.com

Importance:OpinionAnthropic IPO analysis

Why history makes me skeptical of Anthropic's potentially massive upcoming IPO

The AI startup Anthropic is reportedly preparing to file for a public listing after Labor Day, seeking a valuation near $2 trillion. The author outlines historical reasons for staying cautious about participating in the offering. Source: theglobeandmail.com

Importance:Newsmodel releases

Big AI companies push back against open source as Nvidia bets on development lead

OpenAI and Anthropic are launching new AI models as open-source competitors challenge them on cost efficiency. Meta also introduced its Muse Spark 1.3 large language model. Source: digitaltoday.co.kr