hermes-ai.net

Read docs →
HermesHermes Agent Docs
Back to News

OrcaRouter's OrcaSAQ2 Squeezes a 27B Coding Agent Into 16 GB GPUs

OrcaRouter shrinks Qwen3.8-27B from 54GB to 12.3GB using 3-bit mixed-precision quantization, keeping 70% SWE-bench Verified on a single 16GB GPU.

OrcaRouter's OrcaSAQ2 Squeezes a 27B Coding Agent Into 16 GB GPUs
Source
AlphaSignal
Published
Author
AlphaSignal Newsroom
Read
1 min read

OrcaRouter shrinks Qwen3.8-27B from 54GB to 12.3GB using 3-bit mixed-precision quantization, keeping 70% SWE-bench Verified on a single 16GB GPU.

Reporting is indexed from AlphaSignal. Rights remain with the original publisher and cited sources.

Read original report