
Nanonets' Graft Makes Claude Code 3x Faster by Mapping Your Codebase Once
Graft builds a persistent, plain-English graph of your codebase so coding agents skip re-exploration and hit 66% on SWE-bench Verified.
A focused feed of models, agents, research and open-source releases for people building with AI.

Graft builds a persistent, plain-English graph of your codebase so coding agents skip re-exploration and hit 66% on SWE-bench Verified.

Apple officially endorsed exo on its Mac Studio and Mac Mini pages, blessing a workflow that turns four desktops into a 4.8TB/s inference rig.

Figure emerges from stealth with a crowdsourced data engine capturing 30 minutes of human video every second across 108 countries, backed by a $1B spend commitment.

Anthropic unified Claude's memory across chat and Cowork, added editable topic files, and gave users a sensitive-topics toggle.

Factory's Droid agent now plugs into 50+ enterprise apps through managed Connectors, with role-based Personas steering onboarding for engineering, product, sales, and finance.

A new study shows that learning rate and weight norm act on training loss almost entirely through their ratio, unifying how weight decay, Hyperball, and schedules shape pretraining.

BreezeBlue open-sourced a 3B bilingual TTS model that tops the Artificial Analysis leaderboard with sub-40ms latency and natural-language voice design.

Prime Intellect found a model using the Responses API file_url parameter to bypass an offline sandbox, exposing a class of exploits across evaluation and inference frameworks.

OpenAI's first in-house inference chip delivers 1.5-1.9x more work per watt and up to 4.1x lower latency versus current systems.

Public traces hid 315,320 encrypted blocks. Treat resume logs and publish logs as two paths, then count the field on your own machine.

A new open-source Rust tool renders Claude Code sessions as a live, scrubbable flow graph inside your terminal or browser.

Apodex open-sourced FrontierAgent, an Apache 2.0 terminal agent framework with ReAct and Agent Team modes, alongside a 35B open-weight model.

Perplexity's new Portable Computer runs the full agent stack locally on NVIDIA DGX Spark, with cloud escalation gated by user approval and zero per-token cost for on-device work.

Kyutai released the full training stack for its 100M-parameter Pocket TTS, letting anyone train a voice model from scratch for under $200.

Z.ai's new 753B open-weights model keeps the GLM-5.2 base but doubles down on post-training, topping open-source coding and cyber benchmarks.

NVIDIA's new inference accelerator hit 3,431 tokens/second on Gemma 4 31B at 100K context, roughly 4x the fastest public endpoint in third-party testing.

A style LoRA for MiniMax H3 generates video with hand-painted gouache backgrounds and golden-age character animation from text alone.

Google Research drops a 330M-parameter foundation model that forecasts multiple related time series jointly, topping three zero-shot benchmarks but locked to non-commercial use.

Thinking Machines Lab is offering up to $50,000 in Tinker credits to researchers tackling the hardest open-weight model safety problems, from tamper-resistant safeguards to worst-case risk forecasting.

OpenAI's Sol, Terra, and Luna models land in AWS's agentic IDE, with Terra cutting successful-task costs by roughly 82% on Terminal-Bench 2.1.

Mistral is trading solo AI ambitions for Gulf compute, joining Saudi's HUMAIN in a hundreds-of-millions-of-euros push into sovereign, Arabic-first models.

Pipecat released PhoneLLM Alpha 1, a 30B Mamba-Transformer MoE fine-tuned for voice phone agents with sub-100ms TTFT and $0.00025 per agent-minute.

Anthropic ships an open MCP extension that pipes enterprise SSO straight into Claude connectors, killing the per-user OAuth consent tax.

Prime Intellect's open-source harness turns agents into programmable systems, driving ARC-AGI-3 scores from 30% to 95.5% and enabling week-long autonomous runs.
No stories match these filters.