
Topic
LLMs
Large language model releases, benchmarks, and capability research
Featured


Alibaba's Qwen3.8-Max claims agentic AI lead, plans open-weight release

OpenAI cuts Luna prices 80% as AI competition shifts to cost
All Stories

Benchmark Scores Hide the Real Cost of Reasoning Models
Alibaba's Qwen 3.8-Max and Claude Opus 5 demonstrate that raw benchmark scores mask critical differences in time and…

Liquid AI brings edge AI to Raspberry Pi with 2.6B parameter model
Liquid AI, a startup founded by former MIT computer scientists, released LFM2.5-2.6B, a 2.6 billion parameter language…
Alibaba's Qwen3.8-Max Challenges US AI Leadership
Alibaba released Qwen3.8-Max, claiming it is its most capable AI model to date with performance comparable to…

Moonshot AI Releases Kimi K3, First Open 3T-Parameter Model
Moonshot AI released Kimi K3 on July 27, 2026, a 2.8 trillion parameter open-weight model that is the first in its…

Fundamental LLM flaw makes security impossible, researchers argue
Researchers presented a paper at the International Conference on Machine Learning arguing that large language models…

Amazon Cuts Staff From Homegrown LLM Division
Amazon has cut staff from its division developing proprietary large language models, according to a company…
Moonshot's Kimi 3 aims to match Anthropic's Opus 4.8
Moonshot's upcoming Kimi 3 model is expected to narrow the performance gap with Anthropic's Claude Opus 4.8, according…
Open Models Give Enterprises AI Control Closed Systems Cannot
NVIDIA's Nemotron open models enable enterprises to customize, inspect, and control AI systems for domain-specific…
Musk Directs Tesla Staff to Adopt xAI's Grok Model
Elon Musk sent a memo to Tesla staff directing them to adopt Grok, the AI model developed by xAI, citing lower token…

Startup Shrinks 27B-Parameter Model to iPhone
PrismML, a Khosla Ventures-backed startup, claims to have compressed Alibaba's Qwen 3.6 large language model, which…

OpenAI Researcher: GPT-5.6 Beats Human Interns on Most Tasks
At the International Conference on Machine Learning in Seoul, OpenAI senior researcher Noam Brown stated that GPT-5.6…
Nemotron 3 Ultra Matches Closed Models at 10x Lower Cost
NVIDIA's Nemotron 3 Ultra model, tuned through LangChain's Deep Agents harness, achieved benchmark-leading performance…

Anthropic finds consciousness-like structure in Claude
Anthropic published research showing that Claude language models have spontaneously developed an internal structure…

Tencent's Hy3 removes licensing barrier, but GLM-5.2 keeps coding crown
Tencent released Hy3, a 295-billion-parameter open-weight model under Apache 2.0 license, removing regional…
Mistral AI raises funding to democratize frontier AI models
Mistral AI, founded in 2023, has secured significant funding to develop open source AI models with the stated goal of…

Z.ai launches ZCode to undercut Cursor and Claude Code
Z.ai, a Beijing-based AI lab, launched ZCode, a free desktop application designed as an agent-first development…

Why Every LLM Gives You the Same Answer
Large language models exhibit severe homogeneity in their responses to open-ended questions, converging on predictable…
Anthropic Cuts Prices on Claude Sonnet 5 to Challenge Agent Market
Anthropic has launched Claude Sonnet 5, a model positioned as a more affordable alternative to its Opus offering and…