AIskimIQ

Daily AI & tech news brief

Archive/large language models

🧠 Large Language Models

News about foundation models, LLMs, multimodal models, benchmarks, and model releases from OpenAI, Anthropic, Google, Meta, Mistral, and others.

973 articles

Importance:NewsAI Security

Stolen AI Accounts Flood Dark Web as 'LLM Jacking' Grows

According to the Financial Times, hackers are increasingly hijacking corporate AI model and cloud accounts, then reselling access on dark web marketplaces. The trend, dubbed 'LLM jacking,' highlights growing security risks tied to enterprise AI adoption. Source: biz.chosun.com

Importance:NewsLLM security

AI Agents Expose a Gap Between Reading Data and Acting on It

Discussions about AI security tend to focus on the model itself — whether it's aligned, jailbreak-resistant, or prone to hallucination. But a growing risk lies in the disconnect between what AI agents can read and what systems they're allowed to actually modify. Source: venturebeat.com

Importance:Launchmodel routing

KT Develops Tech That Automatically Picks the Right AI Model

KT built an LLM router that analyzes a user's query and automatically selects the most suitable AI model to answer it. The system, called Router Arena, benchmarks and compares different LLMs to route requests efficiently. Source: mk.co.kr

Importance:Newsmodel benchmarks

KT Places Second in LLM Router Benchmark with AutoModelRouter

KT's in-house tool AutoModelRouter, which automatically selects and routes queries among various AI models, took second place in a public benchmark of LLM routers. The result highlights growing competition around efficient model-selection technology. Source: biz.chosun.com

Importance:NewsLLM comparison

Jev vs. LLMs: Testing AI That Decides Instead of Just Generates

TypeSafe AI's Jev was tested across 3,080 classification tasks to compare its accuracy, latency, calibration, and confidence against standard LLMs. The goal was to see whether a decision-focused AI system can outperform generative models on structured tasks. Source: towardsdatascience.com

Importance:Newsmodel deployment challenges

Experts Warn Model Skew Is an Overlooked Risk for Telecom AI

Boost Mobile data scientist Priyank Jain says the telecom industry underestimates unglamorous data issues that quietly determine whether AI models keep working correctly. He argues model skew deserves far more attention than it currently gets. Source: fiercewireless.com

Importance:ResearchLLM evaluation

Can LLMs Reliably Automate Meta-Analysis Generation?

Researchers are increasingly testing large language models for tasks like interpreting scientific literature, extracting data, and supporting statistical analysis in evidence synthesis. The study examines how reliably LLMs can handle automated meta-analysis generation. Source: nature.com

Importance:Newsmodel selection

Why Small Language Models Can Be the Smarter Choice for Federal AI

Government agencies could save money and gain better oversight by choosing model size based on the actual task rather than automatically picking the largest available model. Smaller, purpose-fit models can offer comparable results with lower cost and risk. Source: fedtechmagazine.com

Importance:NewsLLM societal impact

What Are the Broader Societal Effects of Large Language Models?

IBM describes large language models as advanced AI systems trained on massive datasets to understand and generate human-like text. The piece looks at how their widespread adoption is reshaping communication, work, and society more broadly. Source: daily-sun.com

Importance:Researchmodel architecture

TimeBraid merges time-series and language models for forecasting

Researchers introduced TimeBraid, a family of models that align pretrained language models with pretrained time-series foundation models. The goal is to combine strengths of both approaches for better understanding and forecasting of temporal data. Source: alphaxiv.org

Importance:NewsAI in science

Anthropic's CRISPR-like discovery is notable, but no Nobel-level breakthrough

Anthropic announced that its Claude model helped identify a gene-editing system resembling CRISPR. Commentators caution this is unlikely to be a revolutionary scientific leap, more likely just another useful addition to the gene-editing toolbox. Source: newscientist.com

Importance:NewsAI ethics

Can large language models act as ethical advisers for healthier urban design?

LLMs are increasingly used to offer text-based recommendations in fields like urban planning, design and public health. The growing role raises questions about whether these models can be trusted to provide sound ethical guidance in areas affecting people's wellbeing. Source: techxplore.com

Importance:Newsgovernment applications

Why small language models can be the smarter choice for government AI

Federal agencies don't always need the largest available AI model — picking a size that matches the actual task can lower costs and improve oversight. Experts argue this approach offers a more practical path to responsible AI adoption in government. Source: fedtechmagazine.com

Importance:NewsAI in healthcare

AI helps uncover hidden links between herbs and diseases in ancient medical texts

Traditional herbal medicine has built up a vast but scattered body of knowledge about which plants may treat which illnesses. Language models are now being applied to sift through this sparse, fragmented historical data and reveal previously hidden herb-disease connections. Source: bioengineer.org

Importance:PolicyAI evaluation

Taiwan's Digital Ministry Pushes 'Sovereign AI' Standard to Ensure Models Truly Understand the Country

Taiwan's Ministry of Digital Affairs released evaluation results assessing how well AI models genuinely understand Taiwan, not just its language. The initiative aims to bring government and industry together to build a trustworthy, locally grounded AI ecosystem. Source: moda.gov.tw

Importance:LaunchLLM optimization

Liner Debuts Model-Routing API to Slash Enterprise LLM Bills by Over 50%

Liner has launched a new API that automatically directs each request to the cheapest model capable of handling it, rather than always using expensive frontier models. The company says this approach can cut enterprise LLM costs by more than half. Source: manilatimes.net

Importance:ResearchLLM applications

LLM-Based System Beats Traditional Tools at Spotting Smart Grid Cyberattacks

Researchers found that a large language model can detect cyberattacks on smart power grids more effectively than conventional security defenses. The finding matters as energy infrastructure remains one of the most frequently targeted digital systems worldwide. Source: bioengineer.org

Importance:NewsAI startups

DeepMind Alumnus Raises Funds for New Reasoning-Focused Startup

A former DeepMind researcher is reportedly raising tens of millions of dollars for a new venture called Metis Reasoning. The move signals growing investor interest in AI approaches that go beyond traditional large language models. Source: pymnts.com

Importance:NewsAnthropic models

Engineer Speculates Anthropic's Opus 5.5 Could Be First Model Trained via Self-Improvement

A Google engineer named Patrick suggests that Anthropic's Opus 5.5 may have been created through recursive self-improvement, possibly distilled from a much more powerful internal model referred to as Model 2. The claim remains an unconfirmed hypothesis rather than an official statement from Anthropic. Source: eu.36kr.com

Importance:ResearchLLM security

NVIDIA Confidential Computing Aims to Secure Sensitive AI Inference Workloads

NVIDIA is pushing confidential computing as a way to protect sensitive data and proprietary model context during large language model inference. The technology targets enterprise and personal use cases where privacy and security are critical as LLM deployment scales up. Source: developer.nvidia.com