Apriel-Reasoner: RL Post-Training for General-Purpose and Efficient Reasoning research paper by ServiceNow, 2026
ServiceNow · Apr 2, 2026 · Reasoning · 12 upvotes 5 months ago
What it shows
Apriel-Reasoner is a 15B-parameter language model trained with reproducible multi-domain reinforcement learning to improve reasoning efficiency and accuracy across diverse tasks while reducing inference costs.
Hugging Face's summary; not yet checked by hand.
More from ServiceNow
All 20Other reasoning papers
TopicAbout this paper
- Authors
- Rafael Pardinas, Ehsan Kamalloo, David Vazquez and 1 more
- arXiv
- 2604.02007 · PDF
- Citations
- Not counted yet · Semantic Scholar
- Upvotes
- 12 · Hugging Face
- Lab
- ServiceNow · on Companies · on Acquisitions · on Quarterly · on Paydays · on TechConf
Changes
| What changed | |
|---|---|
| Sep 26, 2026today | New paperAdded to the listSep 26, 2026today |