DUET-VLM: Dual stage Unified Efficient Token reduction for VLM Training and Inference research paper by AMD, 2026
AMD · Feb 21, 2026 · Inference and efficiency · 0 citations · 5 upvotes · unverified 7 months ago
What it shows
DUET-VLM presents a dual compression framework that reduces visual tokens while maintaining high accuracy in vision-language models through coordinated vision and language backbone processing.
UnverifiedHugging Face's summary; not yet checked by hand.
More from AMD
All 6Other inference and efficiency papers
TopicAbout this paper
- Authors
- Aditya Kumar Singh, Hitesh Kandala, Pratik Prabhanjan Brahma and 2 more
- arXiv
- 2602.18846 · PDF
- Venue
- arXiv.org
- Citations
- 0, 0 influential · Semantic Scholar
- Upvotes
- 5 · Hugging Face
- Code
- github.com/AMD-AGI/DUET-VLM
- Lab
- AMD · on Companies · on Acquisitions · on Quarterly · on Paydays · on TechConf
Changes
| What changed | |
|---|---|
| Sep 28, 2026today | Influential citationsfirst count: 0Sep 28, 2026today |
| Sep 28, 2026today | Citationsfirst count: 0Sep 28, 2026today |
| Sep 28, 2026today | New paperFound by the weekly scan, unverifiedSep 28, 2026today |