Future Optical Flow Prediction Improves Robot Control & Video Generation research paper by Salesforce, 2026
Salesforce · Jan 15, 2026 · Multimodal and robotics · 19 upvotes 8 months ago
What it shows
A novel language-conditioned optical flow forecasting model combines Vision-Language Model and Diffusion architecture to predict future motion from noisy web-scale video data, demonstrating versatility in robotic manipulation and video generation tasks.
Hugging Face's summary; not yet checked by hand.
More from Salesforce
All 33Other multimodal and robotics papers
TopicAbout this paper
- Authors
- Kanchana Ranasinghe, Honglu Zhou, Yu Fang and 7 more
- arXiv
- 2601.10781 · PDF
- Citations
- Not counted yet · Semantic Scholar
- Upvotes
- 19 · Hugging Face
- Code
- github.com/SalesforceAIResearch/FOFPred
- Lab
- Salesforce · on Companies · on Acquisitions · on Quarterly · on Paydays · on TechConf
Changes
| What changed | |
|---|---|
| Sep 26, 2026today | New paperAdded to the listSep 26, 2026today |