News about foundation models, LLMs, multimodal models, benchmarks, and model releases from OpenAI, Anthropic, Google, Meta, Mistral, and others.
719 articles
Importance:NewsAI investment
Wang Huiwen's family office backs broad AI push from LLMs to agents
A detailed report examines an AI investment strategy backed by Wang Huiwen's family office, covering the full industry chain from foundational technology research to AI agents. Source: eu.36kr.com
Importance:ResearchAI data integration
MCP pilots aim to bridge trust gap between LLMs and open data portals
Pilot projects with the governments of Uruguay and Brazil tested Model Context Protocol (MCP) as a structured way for AI systems to access current, reliable open data. Source: govinsider.asia
Importance:Launchacademic access
Advanced Research Computing broadens access to large language models
In October, Advanced Research Computing launched a new set of LLM services now open to all students, faculty, and staff. The expansion aims to give the wider university community easier access to AI tools for research and coursework. Source: news.vt.edu
Importance:Researchtraining optimization
Host Offloading Eases GPU Memory Strain in JAX-Based LLM Training
Training large language models increasingly hits GPU memory limits before compute capacity is fully used, due to the size of model weights, gradients, and optimizer states. A new approach uses host offloading in JAX to reduce these high-bandwidth memory bottlenecks. Source: developer.nvidia.com
Importance:Researchagents
ARTEM Gives LLM Agents a Sense of Space and Time
Researchers unveiled ARTEM (Agentic Retrieval with Temporal-Episodic Memory), a hybrid agent architecture that combines LLMs with a self-organizing neural memory system. The design aims to give AI agents better spatial and temporal episodic memory for more context-aware reasoning. Source: ojs.aaai.org
Importance:Newsbenchmarks
Rethinking How We Measure LLM Performance
As large language models continue to spread worldwide, experts are reassessing which benchmarks actually capture meaningful performance. The piece takes stock of current LLM evaluation methods and their limitations. Source: cacm.acm.org
Importance:Researchmedical applications
CancerLLM: A New Language Model Built for Oncology
Researchers introduced CancerLLM, a 7-billion-parameter language model designed specifically for the cancer domain. It aims to boost LLM performance on oncology-related tasks and assist medical professionals in their work. Source: nature.com
Importance:Researchagents
A Closer Look at LLM-Powered Intelligent Agents
A new overview examines intelligent agents, a longstanding goal in AI research, in light of recent progress in large language models. It covers key definitions, methods, and future directions for LLM-based agents. Source: wires.onlinelibrary.wiley.com
Importance:Newsresources
5 Books to Deepen Your Understanding of LLMs
A curated list highlights five books covering how to build, fine-tune, and deploy large language models. They're aimed at readers who want a deeper technical grasp of LLM development. Source: kdnuggets.com
Importance:Newsspace LLM deployment
NASA sends Google's Gemma language model into orbit
NASA has deployed Google's Gemma LLM in space, testing the concept amid ongoing debate over whether orbital data centers can realistically host the largest and most capable AI models. Source: aol.com
Importance:Newsrobotics AI model
TurboVLA matches 7B robot AI performance without a language model, runs at 32Hz on consumer GPU
For three years, vision-language-action (VLA) research has assumed that top-tier robot AI models need a large language model at their core. TurboVLA challenges that assumption by achieving comparable results without one. Source: techtimes.com
Importance:NewsLLM security vulnerability
A core design flaw leaves LLMs surprisingly easy to exploit
Researchers found a fundamental weakness that makes it simple to manipulate LLMs into producing harmful outputs, including instructions for sabotaging an aircraft's navigation system. Source: technologyreview.com
Importance:Launchconversational AI
PolyAI unveils real-time voice model aiming for more human-like AI calls
PolyAI Ltd. announced Dialog-RSN-1, a new voice dialog AI model designed to make conversations with AI-driven phone agents sound more natural and human. Source: siliconangle.com
Importance:LaunchLLM security
Cequence AI Gateway adds LLM governance to close model access gaps
The new governance layer gives enterprises a single controlled path for every tool call, API request, agent interaction, and LLM prompt, as companies increasingly seek to manage the tools and APIs used by AI agents. Source: securityboulevard.com
Importance:PolicyLLM usage policy
Debian weighs five competing proposals on AI and LLM use in the project
Debian developers last week opened discussion on a general resolution addressing how AI large language models should be used within the project, with five different proposals now on the table. Source: phoronix.com
Importance:Newsopen-source models
After Kimi K3's open-source release, top-tier open models look increasingly realistic
The open-sourcing of Kimi K3 suggests that AI developers' strategy of releasing highly optimized, high-performance models as open source is gaining real momentum. Source: eu.36kr.com
Importance:LaunchLLM runtime tool
JetBrains open-sources KotlinLLM, a runtime code generator
JetBrains released an experimental IntelliJ IDEA plugin that lets Kotlin/JVM projects hand off runtime logic to an LLM directly from Kotlin code. Source: infoworld.com
Importance:Researchbenchmarks
New African LLM Benchmark Stress-Tests AI Safety Across Languages
Backed by the GSMA, the African Trust & Safety LLM Challenge has released a benchmark of 4,216 verified, reproducible tests designed to probe AI safety across diverse African languages and cultural contexts. The project aims to fill gaps in safety evaluation for regions underrepresented in existing LLM benchmarks. Source: thefastmode.com
Researchers introduced a Dual-Population Co-Evolutionary (DPEC) framework that lets an LLM-driven population of algorithms learn from expert-designed strategies to solve complex scheduling problems. The approach was applied to agile earth observation satellite scheduling, combining adaptive large neighborhood search with LLM assistance. Source: eurekalert.org
Importance:Launchhealthcare AI
Truveta's Language Model Extracts Cancer Staging Data from Millions of Clinical Notes
A study published in JCO Clinical Cancer Informatics shows that the Truveta Language Model can accurately pull critical cancer staging details from vast collections of unstructured clinical records. The research highlights a scalable way to turn raw medical notes into research-ready data. Source: globenewswire.com