Instella-MoE Technical Report research paper by AMD, 2026
AMD · Sep 1, 2026 · Foundation models · 0 citations · 3 upvotes · unverified 3 weeks ago
What it shows
Instella-MoE is an open Mixture-of-Experts language model trained on AMD GPUs using sparse activation, Gated Multi-head Latent Attention, and multi-stage post-training to achieve strong benchmark performance with full reproducibility.
UnverifiedHugging Face's summary; not yet checked by hand.
More from AMD
All 6Other foundation models papers
TopicAbout this paper
- Authors
- Jiang Liu, Sudhanshu Ranjan, Prakamya Mishra and 10 more
- arXiv
- 2609.00791 · PDF
- Citations
- 0, 0 influential · Semantic Scholar
- Upvotes
- 3 · Hugging Face
- Lab
- AMD · on Companies · on Acquisitions · on Quarterly · on Paydays · on TechConf
Changes
| What changed | |
|---|---|
| Sep 28, 2026today | Influential citationsfirst count: 0Sep 28, 2026today |
| Sep 28, 2026today | Citationsfirst count: 0Sep 28, 2026today |
| Sep 28, 2026today | New paperFound by the weekly scan, unverifiedSep 28, 2026today |