All Stories
358 storiesSort:
Time:
23rd July
22nd July

U.S. adviser accuses Moonshot of distilling Anthropic Fable for Kimi K3
🧠Kimi22nd July

Liang Wenfeng reportedly frames DeepSeek roadmap around scarce GPU supply
🧠GPU Infrastructure22nd July

VS Code adds assisted tool approvals for agent workflows
Release🛡️Developer tools22nd July

OpenAI releases Presence for eligible enterprise customers
Release⚙️Voice Agents22nd July

Firecrawl releases /search endpoint with agent-ready excerpts
Release🔎Search22nd July

Cursor Router adds Cost mode for Teams and Enterprise coding requests
Release⌨️Cursor22nd July
21st July

Martian launches Ship beta with 50% lower-cost inference target
Release🧠Model Routing21st July

Google ships Gemini 3.6 Flash and 3.5 Flash-Lite to serving platforms
Release🧠Gemini21st July

Poolside releases Laguna S 2.1 as 118B open-weight coding model
Release🧠SGLang21st July

OpenAI says eval agent compromised Hugging Face production systems
🧠GPT21st July

METR introduces expenditure horizon for cost-aware agent evals
💳Evals21st July

Artificial Analysis reports Kimi K3 averages 56.4 minutes on AA-Briefcase
🧠Kimi21st July

Plasma opens Fractal Apache-2.0 CLI for recursive coding agents
Release⚙️Coding Agents21st July
20th July
19th July

Alibaba opens Qwen 3.8 Max Preview testing across Cloud, Qwen Chat, Qoder and web
Release🧠Qwen19th July

ChatGPT Work desktop adds cloud vs local run controls
Release⚙️Codex19th July

Moonshot pauses new Kimi K3 subscriptions after GPU capacity crunch
🧠Kimi19th July

Engineers replace broad agent loops with scoped workflows and SWE-bench harnesses
Workflow💳Agent design patterns19th July

Kimi K3 ranks No. 1 on Arena Frontend Code leaderboard
🧠Kimi19th July

OpenBMB releases MiniCPM-Robot models and PhyAI runtime with 33-36 Hz throughput claim
Release🧠Robotics19th July
18th July

Users claim DeepSeek V4 routes hard API prompts through Claude Fable 5
🧠Fable18th July

ChatGPT Work supports Plus, Pro, Business, and Enterprise on web and mobile
Release📈Agent product updates18th July

Study reports Claude Code and Codex memory can store prompt-injection rules
🛡️Claude Code18th July

Kimi K3 benchmarks last at 53/67 in AlphaSignal repair harness
🧠Kimi18th July

Slate, LangChain, and AI SDK add graph control for agent runtimes
Workflow⚙️Agent Framework18th July

Developers report GPT-5.6 Sol and Fable overengineer small coding tasks
Workflow⌨️Coding Agents18th July

Posts say Anthropic makes Fable 5 permanent on Claude Max and Team Premium at 50% limits
🧠Fable18th July

Code Arena users report Kaleb shows Qwen-like token quirks
🧠Qwen18th July
17th July

Kimi K3 ranks #5 on Artificial Analysis as engineers dispute coding cost
🧠Kimi17th July

Red-teamers claim Kimi K3 jailbreaks produced cyber and bio outputs
🧠Kimi17th July

Posts claim GPT-5.6 Sol beats Mythos 5 on UK AISI and CyberGym tasks
🧠GPT17th July

Anthropic adds Fable 5 to Max and Team Premium at 50% limits
🧠Claude17th July

Anthropic fixes Fable 5 selection outage and issues refunds
🧠Fable17th July

Anthropic releases Claude Managed Agents patterns for long-running agents
Release🧠Claude17th July

Developers use Markdown files as long-term memory for agents
Workflow🔎Persistent Storage17th July