Webscale-RL: Automated Data Pipeline for Scaling RL Data to Pretraining Levels research paper by Salesforce, 2025
Salesforce · Oct 7, 2025 · Training and scaling · 33 upvotes 11 months ago
What it shows
A scalable data engine converts large-scale pre-training documents into diverse question-answer pairs for reinforcement learning, significantly improving model performance and efficiency.
Hugging Face's summary; not yet checked by hand.
More from Salesforce
All 33Other training and scaling papers
TopicAbout this paper
- Authors
- Zhepeng Cen, Haolin Chen, Shiyu Wang and 7 more
- arXiv
- 2510.06499 · PDF
- Citations
- Not counted yet · Semantic Scholar
- Upvotes
- 33 · Hugging Face
- Code
- github.com/SalesforceAIResearch/PretrainRL-pipeline
- Lab
- Salesforce · on Companies · on Acquisitions · on Quarterly · on Paydays · on TechConf
Changes
| What changed | |
|---|---|
| Sep 26, 2026today | New paperAdded to the listSep 26, 2026today |