
Soup Trains Llama-3.1-8B on a 4 GB Laptop GPU for Free
Layer streaming plus 4-bit quantization pins an 8B model into 3.32 GB of VRAM, turning laptop GPUs into real fine-tuning boxes.
hermes-ai.net
Read docs →A focused feed of models, agents, research and open-source releases for people building with AI.

Layer streaming plus 4-bit quantization pins an 8B model into 3.32 GB of VRAM, turning laptop GPUs into real fine-tuning boxes.

Meta previews WildArtifactBench, an evaluation that judges agents on messy real-world tasks using human and AI preference votes instead of rigid rubrics.

Google's mid-tier reasoning model nearly clears the 85% target on ARC-AGI-2 at a quarter per task, redrawing the cost-performance frontier.

Google's agent-first coding platform now embeds directly in VS Code, JetBrains, Zed, and Visual Studio, sharing one account and context across editors.

Liquid AI ships DSpark draft models for its LFM2.5 family, delivering up to 3.18x GPU throughput and cutting agent latency by nearly half.

AT&T takes an equity stake in Brett Adcock's stealthy AI hardware startup Hark, providing cellular connectivity for phone-free AI devices.

Meta shared new benchmarks, robotics demos, and a real-world agent evaluation for Muse Spark 1.2, its coding-focused multimodal model ahead of an open-weights release.

Exa's search API now plugs directly into Codex and ChatGPT, giving OpenAI's coding agent live access to over 100 billion web sources.

Microsoft Research updates its deep-learning DFT functional with 2.5x more training data, native CP2K integration, and a public performance benchmark.

Kyutai's open-source music transcription model can now export sheet music PDFs, editable MusicXML, and guitar tabs alongside MIDI output.

Black Forest Labs shipped a dedicated video super-resolution endpoint that regenerates any clip up to native 4K with two quality-vs-fidelity modes.

Kyutai and Mirelo's open transcription model now exports editable sheet music and guitar tabs, on top of streaming per-instrument MIDI.

Kakao researchers show how a two-step transfer trick predicts the optimal learning rate for a 10-trillion-token MoE run using tiny proxy models.

A new open-source Codex Skill orchestrates voice, avatar, lip-sync, subtitles and QA into a single one-shot pipeline for vertical presenter videos.

AWS ships a governed bridge between coding agents and 300+ services, replacing AWS Labs tooling with plugins, skills, and a managed MCP server.

AMAP-ML released an open-source orchestration layer that lets Claude Code, Codex, and OpenClaw agents complete multi-hour computer tasks without state drift.
No stories match these filters.