←Back to NewsAI News/InfranewsInfraGpusIntelligent Internet's Meta-Zenith Agent Rewrote vLLM Kernels for a 4x SpeedupAn autonomous agent ran 111 trials to optimize vLLM for Qwen3.8-27B on a single RTX 5090, hitting 4.1x throughput at 65K context.SourceIntelligent InternetPublishedSep 9, 2026, 9:40 AMAuthorAlphaSignal NewsroomRead1 min readAn autonomous agent ran 111 trials to optimize vLLM for Qwen3.8-27B on a single RTX 5090, hitting 4.1x throughput at 65K context.Reporting is indexed from AlphaSignal. Rights remain with the original publisher and cited sources.Read original report ↗Next readsTogether AI · newsTogether AI's ThunderKittens Hits 22.4 PFLOPS on NVIDIA's RubinPyTorch · newsNVIDIA's CUDA Python 1.0 Makes Python a First-Class GPU CitizenAlphaSignal · paperFlashAttention-4's Direct-P Finally Unlocks 2.13x FP4 Speedup on NVIDIA GB200