
Ant Group's Ling-3.0-flash-VL Adds Vision to a 124B Open Reasoning Model
Ant Group's InclusionAI lab dropped an MIT-licensed 124B mixture-of-experts vision model that activates just 5.5B parameters per token and ingests text, images, and video.
A focused feed of models, agents, research and open-source releases for people building with AI.

Ant Group's InclusionAI lab dropped an MIT-licensed 124B mixture-of-experts vision model that activates just 5.5B parameters per token and ingests text, images, and video.

OpenAI launches a vertical ChatGPT for banks that bundles Daloopa, PitchBook, and LSEG data with GPT-6 Astra reasoning and firm-specific templates.

humans& releases Persimmon, a 550B user model built to simulate real people in group chats, fooling AI judges 20 percent of the time.

Google folded the full Gemini API documentation directly into AI Studio, so developers can read reference material without leaving the build surface.

Cohere's new 218B MoE translation model tops WMT26 against DeepL, Google Translate, and open alternatives across 50+ languages under a non-commercial license.

Google launches a native Gemini desktop app for Windows 10 and 11 with an Alt+Space hotkey, Spark agent access, and image and video generation.

DeepSeek's 552B mixture-of-experts model splits input and output compute asymmetrically, natively handles vision, and undercuts its own flagship on price.

World Labs unveiled Atlas, an omni world model that unifies text, images, video, and 3D into a single 3D-grounded spatial context.

Inception's new diffusion-based LLM hits 1,107 tokens per second on NVIDIA GPUs while boosting quality 40% over Mercury 2.

Magic claims a pretraining recipe that matches DeepSeek V4 Pro Base with roughly 50x fewer FLOPs, hitting frontier quality for around $0.5M on GB200.

UkisAI's Swift-Qwen3.8-27B cuts thinking tokens by 58% while keeping accuracy within 1% of the base, delivering roughly 2x faster reasoning.

Sony AI built a system that reads the biomedical literature and predicts unpublished gene interactions, with two confirmed in wet-lab experiments.

Samsung leads Europe's largest tech equity round ever, doubling Mistral's valuation to over €21 billion and cementing a chip-supplier-turned-shareholder alliance.

Fine-tuned 0.8B Qwen beat GPT-5.6 Sol xhigh, 2 million buyer profiles a day became 72 million, and GraphQL serving fell from $27 million to $1 million

OpenBMB's 2.6B dense reasoning model tops the sub-4B open weights leaderboard, punching well above its weight on agentic benchmarks while staying token-efficient.

A year-long Stanford study of 1,182 Character.AI users finds heavier chatbot engagement predicts lower well-being, largely by crowding out face-to-face contact.

TokenRhythm fine-tunes Qwen3.5-4B with a routing harness that turns agent traces into training data, gaining 5.93 points across ten benchmarks.

Japan's national LLM project releases an Apache-2.0 vision-language model with reasoning traces, trained on a scrubbed 29M-sample dataset.

OpenBMB's compact 2B model hits open-source SOTA in its class, beating several 4B rivals on code, math, and agent benchmarks while running locally.

Artificial Analysis pushes Intelligence Index v4.2 with private test sets, a 4,592-page PDF reasoning benchmark, and a new agentic knowledge-work suite.

IFM released six fully open models from 0.9B to 375B parameters, complete with weights, code, training data, and recipes under Apache 2.0.

InclusionAI open-sources a 124B-parameter mixture-of-experts vision-language model with 5.5B active parameters, 256K context, and MIT license.

OpenAI's new flagship saturates ARC-AGI-3 and FrontierMath Tier 4, drives a browser at 1.9x speed, and crosses the Critical cybersecurity threshold.

Keep the project on disk, treat lunch as a bill, and score the first finished tasks rather than dollars per million tokens
No stories match these filters.