Explore Broadly, Reason Sharply: Push Small Models toward the Frontier via Sampling
Power‑sharpened sampling is presented as an inference‑time method to improve reasoning in small language models without reinforcement‑learning updates.
Power‑sharpened sampling is presented as an inference‑time method to improve reasoning in small language models without reinforcement‑learning updates.