← Back to the wire

HeaPA: Difficulty-Aware Heap Sampling and On-Policy Query Augmentation for LLM Reinforcement Learning

AnnouncementResearchJul 10, 2026

Weiqi Wang and co-authors submitted the paper "HeaPA: Difficulty-Aware Heap Sampling and On-Policy Query Augmentation for LLM Reinforcement Learning" to arXiv on January 30, 2026, with a revised version posted on July 9, 2026. The paper is categorized under cs.LG and cs.CL. It was assigned an arXiv-issued DOI via DataCite and is available in PDF and HTML formats.

Evidence

1source· awaiting independent confirmation

No score is assigned. Sources and their independence are shown in the citation chain below.

Citation chain · 1 source

01medPRIMARY
Weiqi WangPerson
Canonical: https://arxiv.org/abs/2601.22448