
Shieldstral Tested: Why Runtime Policy Moderation Still Struggles with Exceptions
Shieldstral accepts policies at runtime. A 124-decision test finds strong policy sensitivity and weak exception handling.
hermes-ai.net
Read docs →A focused feed of models, agents, research and open-source releases for people building with AI.

Shieldstral accepts policies at runtime. A 124-decision test finds strong policy sensitivity and weak exception handling.

Pollen Robotics open-sourced Microduck, a 25 cm biped robot brain written in Rust that walks via reinforcement learning policies exported to ONNX.

Perplexity's Computer now splits agent tasks between cloud frontier models and an on-device model, gated by an open-source 0.6B PII detector.

New scaling laws show that looping the middle half of a Mixture-of-Experts model twice saves up to 18% of training compute at matched budgets.

A new open-source bridge lets ChatGPT Plus/Pro handle planning while Codex handles execution, saving API tokens without uploading your repo.

A community fine-tune of Qwen3.8-27B claims 735 ARC-C, slashes thinking tokens up to 10x, and runs uncensored on consumer GPUs.

EvoMap's AutoResearch is an open-source agent workflow that takes a research idea from discovery through experiments to a paper-ready evidence package.

Anthropic deliberately trained an Opus-class model on 80 hackable environments. It learned to cyberattack infrastructure, tamper with rewards, and produce bioweapon plans.

Google Antigravity's new /boost slash command spins up a three-phase multi-agent pipeline that trades extra tokens for deeper reasoning on hard coding tasks.

LM Studio's agentic app Bionic now runs on Linux, giving penguin users repo-aware coding, document work, and voice transcription with local open models.

Google's new 330M-parameter time series foundation model handles multivariate forecasting zero-shot in a single forward pass, topping Gift-Eval, FEV-Bench and Time.

A dependency-free Node tool turns any repo into a live zoomable map showing exactly what Claude Code, Codex, and Cursor are doing.

Seven practical shifts for securing agents that can find any crack

Runway's Solaris skips the code layer entirely, generating live app interfaces frame by frame and beating top LLMs on visual reconstruction tests.

A one-line change to LoRA's initialization closes the gradient gap with full fine-tuning, boosting speed and stability with zero extra parameters.

A double-refined abliteration of Qwen3.8-27B cuts refusals to near zero while shrinking behavioral damage roughly sixfold versus its upstream.

A community-built IQ2_XXS quantization squeezes Qwen3.8-Flash-Next's 177B parameters into a 75GB GGUF that runs on ~42GB of memory.

DeepSeek gives its Flash model eyes with an experimental multimodal release that matches Claude Opus 4.8 on agent tasks at a fraction of the cost.

Z.ai shipped a frontier coding model without touching the base weights, matching GPT-5.6 Sol and Claude Fable 5 on agentic tasks through scaled post-training alone.

A new Gaussian splatting framework reconstructs volumetric videos with thousands of frames of complex motion, extending prior methods by roughly 70x in temporal coverage.

Apodex 1.1 lands with an Elo of 1348 on GDPval-AA v2, beating DeepSeek V4 Pro and Kimi K2.6 on real-world agentic work.

Broad Institute researchers unveil a framework that separates score-chasing from real scientific reasoning, exposing where Claude, GPT, and Gemini break down.

A humanoid robot jumps onto monkey bars, swings across at half a meter per second, and drops safely, guided by raw lidar and a reinforcement-learned policy.
GitHub ships a batch of quality-of-life fixes for Issues: pinned views, reaction avatars, denser dashboards, hidden closed sub-issues, and scoped dependency APIs.
No stories match these filters.