Alibaba unveils Zhenwu M890 chip and Qwen3.7-Max LLM
CNBC reports Alibaba unveiled a more powerful AI processor, the Zhenwu M890, and a next-generation large language model, Qwen3.7-Max. Source: letsdatascience.com
News about foundation models, LLMs, multimodal models, benchmarks, and model releases from OpenAI, Anthropic, Google, Meta, Mistral, and others.
987 articles
CNBC reports Alibaba unveiled a more powerful AI processor, the Zhenwu M890, and a next-generation large language model, Qwen3.7-Max. Source: letsdatascience.com
New research reveals a sigmoid scaling law for LLM factual recall, driven by model size and training data composition, explaining up to 94% of performance... Source: startuphub.ai
There's a growing assumption that if you connect a large language model (LLM) to your production system or application, it will simply “know” how to answer... Source: towardsdatascience.com
SandboxAQ, the Alphabet spinout chaired by former Google CEO Eric Schmidt, has partnered with Anthropic to embed its large quantitative models directly. Source: theaiinsider.tech
Quantitative models in drug discovery, materials discovery, science and other sectors will now have much wider distribution via Claude. Source: hpcwire.com
Berkeley Lab researchers developed MatterChat, a versatile AI “bridge model” that translates between specialized physics models and conversational Large... Source: newscenter.lbl.gov
Baidu(09888.HK) Chairman and CEO Robin Li stated that the landscape of large-scale models is evolving rapidly. Participants across China and globally are... Source: news.futunn.com
Per reporting by The Next Web and Ars Technica, the preprint server arXiv will ban authors for one year if moderators find "incontrovertible evidence" that... Source: letsdatascience.com
This article shows a full working implementation in pure Python, with real benchmark numbers. Most teams evaluate LLM responses by reading them and guessing... Source: towardsdatascience.com
SandboxAQ integrates its Large Quantitative Models (LQMs) with Anthropic's Claude, unlocking $50T+ quantitative economy access via simple prompts. Source: quantumzeitgeist.com
Millions of people around the world query large language models (LLMs) for information. Although several studies have compellingly documented the persuasive... Source: nature.com
As LLM-powered applications move into production — and as AI agents take on more consequential tasks like browsing the web, writing and executing code,... Source: marktechpost.com
Ask an AI model the same political question in two different languages, and you may get two very different responses. Our team's new research,... Source: goodauthority.org
Researchers at Google say they have uncovered the first known case of hackers using AI to develop a zero-day cyber exploit. Source: computing.co.uk
Microsoft is actively seeking AI startups for potential acquisitions. This move aims to bolster its AI talent and develop a cutting-edge AI model... Source: m.economictimes.com
China's generative AI sector is seeing another wave of aggressive fundraising, with leading large language model (LLM) developers rapidly securing capital... Source: digitimes.com
AI vulnerability exploitation is advancing fast. Criminal groups are using AI to find zero-days, build evasive malware, and automate attacks. Source: helpnetsecurity.com
A team researchers from China have released AntAngelMed, a large open-source medical language model that the team describes as the largest and most capable... Source: marktechpost.com
Learn how understanding LLM distillation techniques improves model training through innovative teacher-student approaches. Source: marktechpost.com
In the past two years, if you've been following the research on the interpretability of large models, you'll notice a phenomenon: in this field,... Source: eu.36kr.com