
Mistral Bets Hundreds of Millions on Saudi Arabia's HUMAIN for Sovereign AI
Mistral is trading solo AI ambitions for Gulf compute, joining Saudi's HUMAIN in a hundreds-of-millions-of-euros push into sovereign, Arabic-first models.
hermes-ai.net
Read docs →A focused feed of models, agents, research and open-source releases for people building with AI.

Mistral is trading solo AI ambitions for Gulf compute, joining Saudi's HUMAIN in a hundreds-of-millions-of-euros push into sovereign, Arabic-first models.

Pipecat released PhoneLLM Alpha 1, a 30B Mamba-Transformer MoE fine-tuned for voice phone agents with sub-100ms TTFT and $0.00025 per agent-minute.

Artificial Analysis and Liquid AI launched a joint benchmark measuring how quantized small models actually perform on iPhone 17 Pro and Galaxy S26 Ultra.

Jeremy Avigad argues the fixation on neural theorem provers hides a far richer landscape of ways AI is reshaping how mathematics gets done.

Qwen ships an FP8 preview of the architecture behind Qwen4, pairing sparse attention, n-gram embeddings and 125B params with 6B active.

A compact 4B open-source model with hybrid sliding-window attention, native 1M-token context, and agent-focused benchmarks that top comparable small models.

A community quantization pairs a 27B Qwen model with multi-token prediction and AMD's IU4 matrix path, hitting ~49 tokens per second on a single Strix Halo APU.

Google Research unveils ME-POIs, a framework that fuses anonymized foot-traffic patterns with text embeddings to improve place understanding by up to 81.9%.

Anthropic opens its most capable cybersecurity model to Enterprise customers through scans, partner integrations, and $35M in open-source credits.

A new leaderboard scores frontier models on synthesizing 70 to 150 page medical case files, with Claude Fable 5 leading at 64.4 percent.

DeepSeek's new experimental vision model bolts image understanding onto V4-Flash at the same price, closing in on Claude Opus 4.8 on multimodal agent tasks.

Sakana AI upgraded its free JP-EN-ZH translator to the new Namazu model, beating Google Translate, DeepL, and Claude Opus 4.8 in head-to-head evaluations.

Google's mid-tier reasoning model nearly clears the 85% target on ARC-AGI-2 at a quarter per task, redrawing the cost-performance frontier.

Liquid AI ships DSpark draft models for its LFM2.5 family, delivering up to 3.18x GPU throughput and cutting agent latency by nearly half.

Kakao researchers show how a two-step transfer trick predicts the optimal learning rate for a 10-trillion-token MoE run using tiny proxy models.
No stories match these filters.