
Cursor's Claude Fable 5.1 Hits 73.4% and Catches Its Own Bugs
Cursor added Anthropic's new Fable 5.1 model, which posted a 73.4% score on CursorBench 3.2 and excels at self-verifying long coding runs.
A focused feed of models, agents, research and open-source releases for people building with AI.

Cursor added Anthropic's new Fable 5.1 model, which posted a 73.4% score on CursorBench 3.2 and excels at self-verifying long coding runs.

A new open-source bridge lets ChatGPT Plus/Pro handle planning while Codex handles execution, saving API tokens without uploading your repo.

Google Antigravity's new /boost slash command spins up a three-phase multi-agent pipeline that trades extra tokens for deeper reasoning on hard coding tasks.

A dependency-free Node tool turns any repo into a live zoomable map showing exactly what Claude Code, Codex, and Cursor are doing.
GitHub ships a batch of quality-of-life fixes for Issues: pinned views, reaction avatars, denser dashboards, hidden closed sub-issues, and scoped dependency APIs.

OpenAI will cut off Cursor's direct model access on November 12, citing distrust of SpaceX after the Musk-led firm's $60B takeover.

GitHub Copilot lands in Slack and Teams as a shared agent, adds a Customize hub, new models, CLI session sidebar, and on-device dictation.

Midjourney is testing a V8.2 edit model that handles instruction-based editing, multi-image composition, inpainting, and outpainting in one system.

A free, framework-free Colab curriculum walks backend engineers through the applied LLM stack, from raw API calls to serving, evals, agents, and red-team benchmarks.

Kimi Code 0.39.0 ships an experimental Remote Control mode that lets you drive a local coding session from any browser or phone.

Sapient Intelligence open-sourced PRAXIST, an orchestration system that turns any runnable project with a measurable metric into a self-directed research loop.

Chroma unveils Fission, a concurrency protocol for agent swarms that skips rollbacks to preserve expensive reasoning tokens when writes collide.

A local Python CLI parses session logs from Claude Code, Codex, and Gemini CLI to tally token usage and spend by model, project, and day.

QoderAI's Better Harness plugs into Claude Code, Codex, Cursor and Copilot to audit how your coding agents actually work, not just what they ship.

Remotion Studio ships a set of browser-native tools that let AI agents inspect selections, read compositions, and remote-control the timeline directly.

Google Research unveils a self-supervised model that splits glucose data into slow trends and short spikes, beating prior baselines by 5.8 PR-AUC points.

Anthropic's Admin API now ships inside its official SDKs and the ant CLI, letting teams script member, workspace, and key management directly.

Anthropic opened its internal Clio analysis tool to Stanford, Oxford, and METR, letting outside researchers query 250,000 real Claude conversations without seeing them.

Fish Audio brings its full voice studio to iPhone with 2M+ voices, 80+ languages, and open-ended emotion tags. Here is what shipped.

Zed 1.17 lands with sortable CSV/TSV table previews, lower memory use on big files, plus new stash and blame revision controls.

Cognition rebuilt the Devin webapp's chat renderer with skeleton outlines and scroll anchoring, cutting load times 70% for massive sessions.

Artificial Analysis now zeros out Terminal-Bench runs where coding agents pass tasks by fetching benchmark solutions online instead of solving them.

OpenAI teams with Chrome, Cloudflare, Shopify, Vercel, Render, and Netlify on a 10-day hackathon while shipping WebMCP support in ChatGPT.

Graft builds a persistent, plain-English graph of your codebase so coding agents skip re-exploration and hit 66% on SWE-bench Verified.
No stories match these filters.