# ExpGSI cuts estimated computation for reward-guided language-model inference

_Published Monday, September 21, 2026 at 2:06 AM EDT · AI, Science · Latest · Tier 2 — Notable_

Researchers introduced ExpBoN, a preprint method that adds exponential noise when selecting the best response from multiple language-model samples, allowing finer control over reward and deviation from the model's original output distribution. They integrated it with guided speculative inference to create ExpGSI, which uses reward signals while reducing estimated computation.

Tests on MATH500, MMLU-STEM and Minerva Math found that ExpGSI maintained comparable accuracy while cutting estimated computation by 14%-39% across candidate budgets for Qwen2.5-Math and by up to 45% at 16 candidates for Qwen3. The findings are limited to the reported models, benchmarks and estimated compute measure.

## Sources

- [cs.LG updates on arXiv.org](https://arxiv.org/abs/2609.21899)

---
Canonical: https://techandbusiness.org/newswire/Xu-yTxv8Q-3lEOR1gWXeQk
Published: 2026-09-21T06:06:22.303Z
Story chronology: 2026-09-21T04:00:00.000Z
Retrieved: 2026-09-21T07:57:11.314Z
Publisher: Tech & Business (techandbusiness.org)
