📚 All Articles
80 guide(s) — regularly updated
Xiaomi MiMo-V2.6: Live-streamed RL training that propels an open-source model to #1 on Artificial Analysis
Discover Xiaomi MiMo-V2.6: live RL training, open-source models up to 1T parameters, #1 on Artificial Analysis. Weights available on Hugging Face.
Anthropic accuses GLM-5.3 of "Mythos-class" hacking capabilities — and reignites the war against open weights
Anthropic accuses GLM-5.3 of Mythos-class hacking capabilities and reignites the war against open weights models. Analysis of the explosive report.
MiniMax M3.1 Flash Preview: China continues its war of small, fast models
Discover MiniMax M3.1 Flash Preview, the new fast small Chinese model quietly launched on September 27, 2026, without a model card or press release.
Claude Sonnet 5.5: Anthropic targets the mid-market with a model that's 30% faster and 30% cheaper
Claude Sonnet 5.5: Anthropic targets the mid-market with a model 30% faster and 30% cheaper. An analysis of the new model in the Claude 5.5 family.
Naive-N0.5-Flash: the open-weight 309B MoE with no full-attention layers at all aims for 2,000 tokens/s per user
Naive-N0.5-Flash: NaiveAI's open-weight 309B-parameter MoE without full-attention aims for 2,000 tokens/s per user on Hugging Face.
Gemini 3.8 Live with Live Avatar: Google adds real-time visual presence to its conversational assistant
Google announces Gemini 3.8 Live with Live Avatar: a conversational assistant with real-time visual presence that listens, sees and speaks.
The price war: GPT-6 Sol and Luna at half price, dropped 90 minutes after Claude Opus 5.5
AI price war: OpenAI launches GPT-6 Sol and Luna at half price, just 90 minutes after Claude Opus 5.5's release. Full analysis.
Gemini 3.8 Live: Google's real-time voice at $1.38 per hour — background reasoning and tool execution without interrupting the conversation
Gemini 3.8 Live: Google's real-time AI voice at $1.38/hr. Background reasoning and tool execution without interrupting the conversation.
DeepSeek V4.1-Flash: MIT, 552B, a quarter of the KV-cache, and the end of the Pro model behind the deepseek-flash endpoint as of September 14
DeepSeek V4.1-Flash: MIT weights on Hugging Face, 552B parameters, KV-cache cut to a quarter, and the Pro model ending as of September 14. Full analysis.
OpenAI GPT-6 Astra: 64.6% on Terminal-Bench-Science and ARC-AGI-3 nearly complete, but its written reasoning is becoming harder to monitor
OpenAI GPT-6 Astra scores 64.6% on Terminal-Bench-Science and nears ARC-AGI-3, but its written reasoning becomes harder to monitor.
DeepSeek swaps V4-Pro for V4.1-Flash behind the same endpoint: why you should re-test your pipelines before September 14
DeepSeek replaces V4-Pro with V4.1-Flash behind the same endpoint: find out why you should re-test your pipelines before September 14.
Best Local LLMs (September 2026)
Discover the best local LLMs of September 2026: complete ranking to run AI locally with privacy and no subscription.