How to Train Your LLM Web Agent: A Statistical Diagnosis research paper by ServiceNow, 2025
ServiceNow · Jul 5, 2025 · Agents and evaluation · 52 upvotes · unverified 1 year ago
What it shows
A study on compute allocation for post-training LLM-based web agents finds that combining supervised fine-tuning with on-policy reinforcement learning improves performance and reduces computational costs compared to using either method alone.
UnverifiedHugging Face's summary; not yet checked by hand.
More from ServiceNow
All 20Other agents and evaluation papers
TopicAbout this paper
- Authors
- Dheeraj Vattikonda, Santhoshi Ravichandran, Emiliano Penaloza and 13 more
- arXiv
- 2507.04103 · PDF
- Citations
- Not counted yet · Semantic Scholar
- Upvotes
- 52 · Hugging Face
- Lab
- ServiceNow · on Companies · on Acquisitions · on Quarterly · on Paydays · on TechConf
Changes
| What changed | |
|---|---|
| Sep 26, 2026today | New paperFound by the weekly scan, unverifiedSep 26, 2026today |