AIskimIQ

Daily AI & tech news brief

Archive/large language models

🧠 Large Language Models

News about foundation models, LLMs, multimodal models, benchmarks, and model releases from OpenAI, Anthropic, Google, Meta, Mistral, and others.

693 articles

Importance:ResearchLLM applications in transportation

LLM-assisted screening method for large-scale transportation model calibration

Sustainable mobility transitions require a comprehensive understanding of transport dynamics. To this end, Activity-Based Models (ABMs) provide a... Source: nature.com

Importance:Launchcountry-specific LLMs

PM Presents First Portuguese AI Model

Prime Minister Luís Montenegro presented AMÁLIA, the first large artificial intelligence (AI) language model developed specifically for European Portuguese,... Source: portugaldecoded.substack.com

Importance:ResearchLLM training techniques

LLM Data Mixture Breaks When Training Pools Shift: Causal Inference Offers Fix

Every large language model begins as a question of proportion. How much web text? How much code? How much scientific literature? The ratio of training data... Source: techtimes.com

Importance:LaunchClaude Sonnet 5

Anthropic launches Claude Sonnet 5 AI model with coding, safety upgrades as Fable and Mythos controls lifted

Anthropic PBC today debuted Claude Sonnet 5, a midrange large language model that outperforms its predecessor in several areas. The LLM will be the default... Source: siliconangle.com

Importance:NewsChinese AI models

A new, inexpensive Chinese AI model is catching up with Anthropic, OpenAI on their home turf

Since DeepSeek shocked markets early last year with its cheap but powerful AI model, global consumers have been faced with a choice: Chinese offerings with... Source: reuters.com

Importance:PolicyEuropean open-source LLM

EU Backs Open-Source AI Model Covering 24 EU Languages

The European Commission has selected the EUROPA consortium, led by Italian AI company Domyn, to develop an open-source large language model (LLM) covering... Source: slator.com

Importance:LaunchNVIDIA Nemotron

NVIDIA Diffusion LLM Hits 2.42x Throughput Without Retraining: Nemotron TwoTower Released

NVIDIA diffusion language model Nemotron TwoTower achieves 2.42x LLM inference throughput without a full retraining run, retaining 98.7% of autoregressive... Source: techtimes.com

Importance:Researchhybrid LLM workflows

Hybrid LLM Workflows Blend Local Privacy with Cloud Reasoning

A detailed field guide published on Towards Data Science explores hybrid local-cloud large language model (LLM) workflows, offering a practical framework... Source: techgig.com

Importance:Researchclinical AI

Clinical drug report generation using multi-phase prompt large language models

Accurate and timely synthesis of clinical drug information is essential for pharmacists engaged in evidence-based practice and formulary evaluation. Source: nature.com

Importance:OpinionRAG parsing

The Untaught Lessons of RAG Question Parsing: Structure Before You Search

This article is a manifesto companion to Enterprise Document Intelligence, the series whose philosophy is laid out in Amplify the Expert. Source: towardsdatascience.com

Importance:LaunchAnthropic Claude Science

Anthropic releases Claude Science, a product aimed at researchers, the pharma industry

Last year, Claude Code fundamentally upended programming industry. Anthropic CEO Dario Amodei believes Claude Science will do the same for life sciences. Source: statnews.com

Importance:NewsLLM inference optimization

DeepSeek open sources DSpark, a new framework to speed up LLM inference by up to 85%

DSpark can make decoding faster, but acceptance quality still determines how much speed the system actually realizes. Source: venturebeat.com

Importance:NewsChina AI model on domestic chips

Meituan debuts China’s biggest AI model trained on local chips

As China attempts to move beyond using domestic chips solely for model inference, food delivery giant Meituan has released what it claims is the country's... Source: amp.scmp.com

Importance:Launchopen-source LLM release

Portugal launches its own open-source AI model, “Amália”

Amália comes as an alternative Portuguese-language large language model (LLM) that will be released under an open-source license. Key takeaways:. Source: cybernews.com

Importance:Researchbrain-inspired LLM architecture

EPFL builds a brain-inspired AI model called MiCRo

EPFL researchers have created MiCRo, a brain-inspired large language model divided into specialized “experts” that makes AI more transparent. Source: ggba.swiss

Importance:ResearchDeepSeek DSpark/inference acceleration

DeepSeek has released 'DSpark,' which can improve the speed of AI language model generation by up to 85%.

DeepSeek has released DSpark , a speculative decoding technology that accelerates text generation for large-scale language models. Source: gigazine.net

Importance:Launchnew LLM launch/Base44 Base 1

Base44 Launches Its Own In-House AI Model

Israeli AI-app builder Base44 has launched its first proprietary large language model, “Base 1,” now live in production across its platform. Source: israeldefense.co.il

Importance:LaunchGoogle BigQuery AI/natural language SQL

Google adds AI.AGG for BigQuery natural-language SQL

Google has introduced AI.AGG() in preview for BigQuery. The function lets users run natural-language aggregation queries over unstructured data in SQL. Source: itbrief.asia

Importance:ResearchLLM inference efficiency

Peking University and DeepSeek Open-Source DSpark, Delivering Major Leap in LLM Inference Efficiency

Peking University and DeepSeek jointly open-source DSpark, a speculative decoding framework that boosts LLM inference speed by 60-85% with up to 661%... Source: pandaily.com

Importance:NewsClaude benchmarks

Claude AI Beats Human Robotics Teams 20x: Anthropic Marks Physical AI Turn

Claude AI robotics benchmark shows Opus 4.7 finishing physical robot programming in 9 minutes, against 181 minutes for AI-assisted human teams, in Anthropic... Source: techtimes.com