AIskimIQ

Daily AI & tech news brief

Archive/large language models

🧠 Large Language Models

News about foundation models, LLMs, multimodal models, benchmarks, and model releases from OpenAI, Anthropic, Google, Meta, Mistral, and others.

987 articles

Importance:NewsAnthropic Claude deployment

Australia now has access to Anthropic’s Claude Mythos. It may improve cyber safety – but not for everyone

The AI era has fundamentally changed the risks associated with poor cybersecurity practices. Source: theconversation.com

Importance:ResearchLLM benchmarks/evaluation

Law Professors Rate AI Answers Higher in Blinded Study

A blinded study from Stanford Law School, now posted as a working paper on SSRN (Salinas, Frieders, Guha, and colleagues, with Julian Nyarko),... Source: letsdatascience.com

Importance:OpinionAI consciousness/Anthropic

No, Artificial Intelligence Is Not Conscious

Anthropic is regarded as a giant among AI companies, but perhaps what it really excels in is anthropomorphism. Earlier this year, the company released an... Source: theatlantic.com

Importance:ResearchLLM agents/continual learning

LifeSkill: LLM Agents Learn Continuously

The imperative for Large Language Model (LLM) agents to adapt and learn continuously in dynamic, interactive environments is clear. Source: startuphub.ai

Importance:ResearchLLM benchmarks/attention testing

ChatGPT, Claude stumble in attention test, raising questions for AGI

ChatGPT and Claude, among the latest large language models, performed worse than expected in the Stroop test, a psychology experiment used to gauge human... Source: digitaltoday.co.kr

Importance:ResearchLLM efficiency/multilingual

Making LLMs faster and more efficient across multiple languages

Large language models (LLMs), which are the artificial intelligence (AI) systems behind modern chatbots, translation tools, and virtual assistants,... Source: techxplore.com

Importance:LaunchGoogle Gemma open source model

Google's new open source Gemma 4 12B analyzes audio, video — and runs entirely locally on a typical 16GB enterprise laptop

While many AI open source model providers are pursuing larger and more powerful models, Google is still giving attention to the smaller, more local side of... Source: venturebeat.com

Importance:LaunchMicrosoft first-party LLM

Microsoft unveils first LLM, vibe coding model, and speech/text updates

Microsoft has unveiled its first large language model called MAI-Thinking-1, which it says matches the performance of Anthropic's Claude Opus 4.6. Source: legaltechnology.com

Importance:Launchregional LLM release

Hong Kong unveils HKGAI-V3 model built on domestic tech and DeepSeek V4

Agent Workshop operated stably for up to 28 hours without interruption in a single session to produce a research report. Source: scmp.com

Importance:NewsLLM governance/policy

US Treasury Secretary Bessent meets with LLM labs in San Francisco

Scott Bessent's weekend sit-down with major AI companies signals the Treasury Department's deepening focus on large language model governance. Source: cryptobriefing.com

Importance:Launchregional LLM release

SAR updates its first homegrown AI model

Hong Kong on Wednesday unveiled the latest version of its homegrown large language model, HKGAI V3, marking a major step in efforts to build artificial... Source: global.chinadaily.com.cn

Importance:NewsLLM security/malware

Researchers build self-replicating AI worm with BYO LLM

A team of researchers at the University of Toronto in Canada has assembled a self-replicating malware - a worm - that is able to reason its way through... Source: itnews.com.au

Importance:PolicyAI regulation/executive order

Trump’s new AI executive order drastically shifts the administration’s stance on the tech

This order asks artificial intelligence companies to give the U.S. government up to 30 days to assess frontier models before they are released. Source: scientificamerican.com

Importance:ResearchLLM benchmarks/evaluation

Stroop Test Exposes Inherent LLM Flaw

Summary: A new cognitive evaluation of artificial intelligence has unmasked a fundamental, systemic flaw running through large language model (LLM)... Source: neurosciencenews.com

Importance:LaunchMiniMax M3 model release/benchmarks

MiniMax M3 debuts, eclipsing GPT-5.5 and Gemini 3.1 Pro on key benchmark performance for just 5-10% of the cost

Big news in enterprise AI broke over the weekend as Chinese AI startup MiniMax released its highly anticipated M3 large language model on Sunday evening... Source: venturebeat.com

Importance:ResearchLLM efficiency/quantization

Tether AI open-sources TurboQuant, reducing LLM KV cache memory use by 5x

The stablecoin giant's AI division adapts a Google Research algorithm into a production-ready tool that could make running large language models on phones... Source: cryptobriefing.com

Importance:NewsMiniMax M3 analyst coverage

M Stanley: MINIMAX-W (00100.HK) Releases M3 in Major Upgrade; Reiterates Overweight

M Stanley issued a research report stating that MINIMAX-W (00100.HK) has launched its new flagship large language model (LLM), M3, marking a major upg... Source: aastocks.com

Importance:OpinionLLM robustness/adversarial

How human error became a weapon against large language models

Alan Turing proposed a test for machine intelligence: could a computer convince a human it was human? We have begun conducting the same test on ourselves,... Source: newscientist.com

Importance:VideoLLM release

Introduction Video: Large Language Model “Fujitsu Generative AI For Cohere Takane” Gwyneth Paltrow (0Uo7khfQEq)

In this video, we introduce “Takane,” an enterprise-grade large language model (LLM) jointly developed by Fujitsu and Cohere.With high-performance ... Source: mshale.com

Importance:NewsLLM company/IPO

AI start-up MiniMax plans secondary listing after HK IPO

Chinese large-language model (LLM) company MiniMax is planning a secondary listing on the A-share market following its HK$106.7 billion ($13.6 billion) IPO... Source: globaltimes.cn