
Moonshot AI's Kimi Code Now Lets You Control Coding Sessions From Any Device
Kimi Code 0.39.0 ships an experimental Remote Control mode that lets you drive a local coding session from any browser or phone.
hermes-ai.net
Read docs →A focused feed of models, agents, research and open-source releases for people building with AI.

Kimi Code 0.39.0 ships an experimental Remote Control mode that lets you drive a local coding session from any browser or phone.

Sapient Intelligence open-sourced PRAXIST, an orchestration system that turns any runnable project with a measurable metric into a self-directed research loop.

Chroma unveils Fission, a concurrency protocol for agent swarms that skips rollbacks to preserve expensive reasoning tokens when writes collide.

A local Python CLI parses session logs from Claude Code, Codex, and Gemini CLI to tally token usage and spend by model, project, and day.

QoderAI's Better Harness plugs into Claude Code, Codex, Cursor and Copilot to audit how your coding agents actually work, not just what they ship.

Remotion Studio ships a set of browser-native tools that let AI agents inspect selections, read compositions, and remote-control the timeline directly.

Google Research unveils a self-supervised model that splits glucose data into slow trends and short spikes, beating prior baselines by 5.8 PR-AUC points.

Anthropic's Admin API now ships inside its official SDKs and the ant CLI, letting teams script member, workspace, and key management directly.

Anthropic opened its internal Clio analysis tool to Stanford, Oxford, and METR, letting outside researchers query 250,000 real Claude conversations without seeing them.

Fish Audio brings its full voice studio to iPhone with 2M+ voices, 80+ languages, and open-ended emotion tags. Here is what shipped.

Zed 1.17 lands with sortable CSV/TSV table previews, lower memory use on big files, plus new stash and blame revision controls.

Cognition rebuilt the Devin webapp's chat renderer with skeleton outlines and scroll anchoring, cutting load times 70% for massive sessions.

Artificial Analysis now zeros out Terminal-Bench runs where coding agents pass tasks by fetching benchmark solutions online instead of solving them.

OpenAI teams with Chrome, Cloudflare, Shopify, Vercel, Render, and Netlify on a 10-day hackathon while shipping WebMCP support in ChatGPT.

Graft builds a persistent, plain-English graph of your codebase so coding agents skip re-exploration and hit 66% on SWE-bench Verified.

Anthropic unified Claude's memory across chat and Cowork, added editable topic files, and gave users a sensitive-topics toggle.

Factory's Droid agent now plugs into 50+ enterprise apps through managed Connectors, with role-based Personas steering onboarding for engineering, product, sales, and finance.

A new open-source Rust tool renders Claude Code sessions as a live, scrubbable flow graph inside your terminal or browser.

OpenAI's Sol, Terra, and Luna models land in AWS's agentic IDE, with Terra cutting successful-task costs by roughly 82% on Terminal-Bench 2.1.

Prime Intellect's open-source harness turns agents into programmable systems, driving ARC-AGI-3 scores from 30% to 95.5% and enabling week-long autonomous runs.

ElevenLabs ships a terminal-first CLI that treats voice agents as version-controlled config files, with structured JSON and skills for coding agents.

A new Zig-built plugin turns x64dbg into an MCP endpoint, letting Claude and other AI agents drive reverse engineering sessions through 71 tools.

A weekend reverse-engineering project turns a compiled macOS app into 440K lines of readable TypeScript, then bolts on a four-provider inference router.

GitHub added a three-day default cooldown on Dependabot version update pull requests, aiming to filter out short-lived poisoned package releases before they land in your repo.
No stories match these filters.