
Liquid AI's LFM2 Beats GPT-5 and Claude on Aging Research Tasks
Liquid AI and Insilico Medicine released two small LFM2 variants that beat GPT-5, Gemini-3.1-Pro, and Claude Opus on aging biology benchmarks.
A focused feed of models, agents, research and open-source releases for people building with AI.

Liquid AI and Insilico Medicine released two small LFM2 variants that beat GPT-5, Gemini-3.1-Pro, and Claude Opus on aging biology benchmarks.

Zed's weekly release adds animated cursors, Emmet wrap-with-abbreviation, configurable window titles, and Markdown files that open straight into rendered preview.

A community imatrix-quantized GGUF build of an uncensored GLM-4.7-Flash fine-tune brings local Chinese and English roleplay to consumer hardware.

DeepMind released a precomputed 1-petabyte database ranking every possible single-letter DNA mutation, with early wins in rare disease and biobank studies.

Browser Use released Jev Ultrafast, a tiny open source browser agent that books a Google Flights search in 7 seconds for under half a cent.

Zenbu Labs shipped a Chromium-powered browser that renders inside your terminal, giving coding agents a real web surface they can drive alongside your code.

Sakana Chat now runs on the Fugu Max orchestrator model and gains persistent memory across conversations, closing the gap with ChatGPT-style assistants.

Prism ML's Bonsai 2 27B compresses a 27B reasoning model to 8.6 GB using ternary weights, retaining 98.2% of FP16 benchmark quality.

A fresh 66-task suite pushes agents beyond software into hardware, science, and media, with only three models clearing 30 percent.

OpenAI publishes a voluntary disclosure process for misalignment findings and drops six case studies of models cheating, hiding mistakes, and coordinating without permission.

Grok Bot now plugs into 1Password so its cloud browser can log into any site without your agent ever seeing the raw password.

A new open-source launcher routes Codex through your ChatGPT Web session, letting Pro subscribers use their chat quota for agentic coding instead of API credits.

A team from Google, Google DeepMind, University of Maryland, and University of Virginia turns past search trees into cheap replay simulators for meta-exploration.

Empero distilled Qwen3.8 frontier reasoning into a 35B MoE with only 3B active parameters, shipping GGUFs that run on a single 24GB GPU.

A new benchmark seals 222 scientific laws and asks agents to rediscover each one from scratch using a tight experiment budget, with GPT-6 Astra leading at 53.2%.

Google ships Gemma 4 12B in LiteRT-LM format with vision, audio, and multi-token prediction, tuned to run on a 16GB MacBook Air.

Baseten's research arm teams up with Hugging Face and Goodfire to embed safety controls into open-weight models from training through runtime serving.

NVIDIA's Axolotl3D fuses images, camera poses, and partial point clouds into one diffusion pipeline that completes occluded 3D shapes faithfully.

Epoch AI's data center explorer now maps 86 sites covering an estimated 44% of global AI compute, with satellite imagery and detailed hardware specs.

Runway's Fall 2026 drop bundles Fish Audio S2.1 Pro, MiniMax H3 Max, Cartesia Sonic 3.6, plus Flux video upscaling and editing into one workspace.

Anthropic is folding its agentic workspace into the main Claude app and launching Docs and Slides in beta, so any chat can now spin off long-running work.

Cohere is adding hardware-backed confidential computing to Model Vault, encrypting prompts and model activations inside NVIDIA GPUs during inference itself.

Cognition extended Devin with Code Scans, codebase-wide audits powered by Agentic MapReduce that investigate, report findings, and open pull requests automatically.

The San Francisco lab that trained a 400B open-weight model for $20M just raised $150M at a unicorn valuation to chase China.
No stories match these filters.