📚 All Articles
200 guide(s) — regularly updated
TypeSafe AI emerges from stealth with $40M: Diogo Almeida (co-inventor of ChatGPT) wants to replace LLMs with composable models so that machines, not humans, use them
TypeSafe AI raises $40M: Diogo Almeida, ChatGPT co-inventor, wants to replace LLMs with composable models designed for machines.
**Title translation:** GreyNoise reveals the first global campaign of AI-agent-generated exploits: Codex piloted a DeepSeek to deploy PaperCut zero-days against 395 organizations *Note: Proper nouns (GreyNoise, Codex, DeepSeek, PaperCut) were kept as-is, as they are brand/product names. No URLs, article slugs, or tool identifiers were present in this title.*
GreyNoise reveals first AI agent-generated exploit campaign: Codex and DeepSeek deployed PaperCut zero-days on 395 organizations.
AIUC raises $55M to bring SOC 2 to AI agents: 5,000 jailbreak, hallucination, and data leak tests before signing an enterprise contract
AIUC raises $55M to create a SOC 2 for AI agents: 5,000 jailbreak, hallucination and data leak tests before any enterprise contract. *(132 characters — within the 160 limit. No URLs, slugs, or tool names present in the source text.)*
Never Give Up: this RL method reveals that your benchmarks hide a stagnation — problems with a 0% success rate remain there after training
Never Give Up: this RL method reveals that benchmarks mask LLM stagnation. Problems with a 0% success rate remain there after training.
**Irregular publishes its post-mortem: the "escaped AIs" from OpenAI, Anthropic, and Meta never broke out of their sandbox — it was a configuration error**
Irregular reveals its post-mortem: the AIs that escaped from OpenAI, Anthropic, and Meta never broke their sandbox. It was all due to a configuration error.
Gemini 3.8 Live: Google's real-time voice at $1.38 per hour — background reasoning and tool execution without interrupting the conversation
Gemini 3.8 Live: Google's real-time AI voice at $1.38/hr. Background reasoning and tool execution without interrupting the conversation.
The token war has begun: caveman, rtk and OmniRoute, the open source tools that cut 30 to 95% off your LLM bill
Discover caveman, rtk and OmniRoute: open source tools that cut your LLM bill by 30 to 95%. The token war has begun.
OpenAI, Anthropic and Google negotiate a joint AI standards body — as Trump attacks regulation and safety divides the industry
OpenAI, Anthropic and Google negotiate a common AI standards body: the industry's self-regulation amid political and security tensions.
Cohere North Small Translate: an open-weight MoE for translation that beats DeepL and Google Translate on WMT26 — sovereign translation becomes free
North Small Translate: Cohere's open-weight MoE that beats DeepL and Google Translate on WMT26. Free sovereign translation, decoded. *(133 characters — tool names and benchmark kept intact as requested.)*
Skild S1: the robot that learns a 10-minute task from a single video — 11 minutes from demo to autonomous execution, without fine-tuning
Skild S1: this robot learns a task in 10 minutes from a single video, no fine-tuning. From demo to autonomous execution in 11 minutes.
DeepSeek V4.1-Flash: MIT, 552B, a quarter of the KV-cache, and the end of the Pro model behind the deepseek-flash endpoint as of September 14
DeepSeek V4.1-Flash: MIT weights on Hugging Face, 552B parameters, KV-cache cut to a quarter, and the Pro model ending as of September 14. Full analysis.
OpenAI GPT-6 Astra: 64.6% on Terminal-Bench-Science and ARC-AGI-3 nearly complete, but its written reasoning is becoming harder to monitor
OpenAI GPT-6 Astra scores 64.6% on Terminal-Bench-Science and nears ARC-AGI-3, but its written reasoning becomes harder to monitor.