Skip to content
Papers.

Designer-RSI: Evolving Procedural Memory from User Traffic for Agentic Graphic Design research paper by Adobe, 2026

Adobe · Sep 18, 2026 · Agents and evaluation · 0 citations · 33 upvotes · unverified

Read on arXiv

What it shows

Professional graphic design is a long-horizon agentic task in which structured, editable artifacts emerge from many interdependent actions, yet outcomes admit no reliable programmatic oracle.

UnverifiedHugging Face's summary; not yet checked by hand.

More from Adobe

All 8
PaperCitations
Beyond Starry Night: Shortcut-Aware Control-State Planning for Artist-Grounded Text to Image GenerationAtelier improves artist-grounded image generation by translating vague artistic intent into explicit control states that separate scene content from style, reducing reliance on stereotypical shortcuts.Multimodal and robotics · Aug 2026 · Unverified1
LoST: Level of Semantics Tokenization for 3D ShapesLevel-of-Semantics Tokenization (LoST) improves 3D shape generation by ordering tokens based on semantic salience and using a novel relational alignment loss for better reconstruction and efficiency.Multimodal and robotics · Mar 2026 · Unverified3
WorldCam: Interactive Autoregressive 3D Gaming Worlds with Camera Pose as a Unifying Geometric RepresentationVideo diffusion transformers enhanced with camera pose representation enable precise action control and long-term 3D consistency in interactive gaming environments through physics-based action spaces and geometric grounding.Multimodal and robotics · Mar 2026 · Unverified10
Memory-V2V: Augmenting Video-to-Video Diffusion Models with MemoryMemory-V2V enhances multi-turn video editing by maintaining cross-consistency through explicit memory mechanisms and efficient token compression in video-to-video diffusion models.Multimodal and robotics · Jan 2026 · Unverified1
Both Semantics and Reconstruction Matter: Making Representation Encoders Ready for Text-to-Image Generation and EditingLatent diffusion models using representation encoder features face challenges in semantic compactness and pixel-level reconstruction, which are addressed through a semantic-pixel reconstruction objective that enables compact yet semantically rich representations for unified text-to-image and image editing tasks.Multimodal and robotics · Dec 2025 · Unverified22
V-RGBX: Video Editing with Accurate Controls over Intrinsic PropertiesV-RGBX presents an end-to-end framework for intrinsic-aware video editing that combines video inverse rendering, photorealistic video synthesis, and keyframe-based editing with physically grounded intrinsic channel manipulation.Multimodal and robotics · Dec 2025 · Unverified5
MotionStream: Real-Time Video Generation with Interactive Motion ControlsMotionStream enables real-time video generation with sub-second latency and up to 29 FPS by distilling a text-to-video model with motion control into a causal student using Self Forcing with Distribution Matching Distillation and sliding-window causal attention with attention sinks.Multimodal and robotics · Nov 2025 · Unverified76
Topic
PaperCitations
Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic CapabilitiesGemini 2.X model family, including Gemini 2.5 Pro and Flash, offers superior coding, reasoning, and multimodal understanding capabilities across a range of computational efficiencies.Google · Jul 2025 · Unverified4,134
DeepSeek-V3.2: Pushing the Frontier of Open Large Language ModelsDeepSeek-V3.2 introduces DeepSeek Sparse Attention and a scalable reinforcement learning framework, achieving superior reasoning and performance compared to GPT-5 and Gemini-3.0-Pro in complex reasoning tasks.DeepSeek · Dec 2025 · Unverified761
WebWatcher: Breaking New Frontier of Vision-Language Deep Research AgentWebWatcher, a multimodal agent with enhanced visual-language reasoning, outperforms existing agents in complex visual and textual information retrieval tasks using synthetic trajectories and reinforcement learning.Alibaba (Qwen) · Aug 2025 · Unverified117
SkillOpt: Executive Strategy for Self-Evolving Agent SkillsSkillOpt introduces a systematic text-space optimizer for agent skills that trains skills as external agent state with stable updates and zero deployment inference overhead, achieving superior performance across multiple benchmarks and execution environments.Microsoft · May 2026 · Unverified85
Nemotron 3 Nano: Open, Efficient Mixture-of-Experts Hybrid Mamba-Transformer Model for Agentic ReasoningWe present Nemotron 3 Nano 30B-A3B, a Mixture-of-Experts hybrid Mamba-Transformer language model.NVIDIA · Dec 2025 · Unverified81
AgentFold: Long-Horizon Web Agents with Proactive Context ManagementAgentFold, a novel proactive context management paradigm, enhances long-horizon task performance through dynamic context folding, achieving superior results on benchmarks compared to larger models and proprietary agents.Alibaba (Qwen) · Oct 2025 · Unverified77
About this paper
Authors
Hongyang Du, Lan Yan, Christian Flores and 1 more
arXiv
2609.22086 · PDF
Citations
0, 0 influential · Semantic Scholar
Upvotes
33 · Hugging Face
Lab
Adobe · on Companies · on Acquisitions · on Quarterly · on Paydays · on TechConf

Changes

What changed
Influential citationsfirst count: 0Sep 25, 2026
Citationsfirst count: 0Sep 25, 2026
New paperFound by the weekly scan, unverifiedSep 25, 2026

Sources: each lab's own papers and arXiv, with citation and upvote counts from Semantic Scholar and Hugging Face. One-line summaries are for orientation, not a substitute for the paper. Logos via logo.dev; trademarks belong to their owners.

New papers by email

Monday afternoons, only in weeks with new papers from the labs.

Double opt-in. Unsubscribe any time.