Skip to content

RPTune: Learned Context Curation for LLM Catalog Search research paper by Google, 2026

Google · Oct 1, 2026 · Retrieval and data · 31 upvotes · unverified 4 days ago

Read on arXiv

What it shows

For small merchant businesses (SMBs) whose catalogs fit within a long-context LLM, full-catalog prompting offers a compelling alternative to multi-stage retrieval designed primarily for large marketplaces with...

By Chuxuan Hu, Hejie Cui, Norman Huang and 3 more · arXiv 2610.00964 · PDF

UnverifiedHugging Face's summary; not yet checked by hand.

More from Google

All 36
PaperCitations
AIM: Agentic Idea Management for Automated ResearchFrontier LLMs are increasingly used to automate scientific research through iterative search.Agents and evaluation · Sep 2026 · Unverified6 days ago-
Selecting Diverse SFT Traces Improves Post-RL GeneralizationVerified solutions are not equally useful for preparing reasoning models for reinforcement learning (RL).Reasoning · Sep 2026 · Unverified8 days ago-
RRSI: Regularized Recursive Self-Improvement of Agent HarnessesRRSI keeps self-improving agent harnesses from memorising their training tasks, so gains carry over to new benchmarks.Agents and evaluation · Sep 20262 weeks ago3
Verifiable Social Reasoning for LLM AssistantsLLM assistants are widely used for daily social advice, yet evaluating their social reasoning in such consultation settings remains challenging since (i) it requires setups where the assistant learns about social...Reasoning · Sep 2026 · Unverified2 weeks ago1
Dream-RSI: Recursive Self-Improvement through Evolving WorldsDream-RSI enables scalable recursive self-improvement by using historical discovery replay to evaluate exploration policies offline, reducing costly online evaluations.Retrieval and data · Sep 2026 · Unverified3 weeks ago9
Procedural Graphs: Self-Evolving Execution Structures for LLM AgentsA procedural graph framework organizes agent actions into structured relational triplets, providing situational guidance and self-evolving topology to improve long-horizon tool use.Foundation models · Sep 2026 · Unverified3 weeks ago0
WikiSkill: Compiling Agent Experience into Persistent Knowledge for Skill EvolutionWikiSkill co-evolves reusable agent skills with a persistent knowledge base to systematically accumulate experience and improve performance across models.Agents and evaluation · Aug 2026 · Unverified5 weeks ago6
EnvHarness: Awakening Static Worlds for Agent LearningEnvHarness and EnvRigger dynamically reshape static environments via programmable plugins to target agent weaknesses and improve reinforcement learning co-evolution.Agents and evaluation · Aug 2026 · Unverified6 weeks ago8
Topic
PaperCitations
From Local to Global: A Graph RAG Approach to Query-Focused SummarizationGraphRAG builds a knowledge graph of a document set so a model can answer questions about the whole collection.Microsoft · Apr 20242 years ago2,267
DINOv3DINOv3, a self-supervised learning model, achieves superior performance across various vision tasks by scaling datasets and models, addressing dense feature degradation, and enhancing flexibility with post-hoc strategies.Meta · Aug 2025 · Unverified1 year ago1,608
DeepSeek-Prover: Advancing Theorem Proving in LLMs through Large-Scale Synthetic DataGenerating extensive Lean 4 proof data from mathematical competition problems improved DeepSeekMath 7B's theorem-proving capabilities over GPT-4 and other methods.DeepSeek · May 2024 · Unverified2 years ago264
Qwen3-VL-Embedding and Qwen3-VL-Reranker: A Unified Framework for State-of-the-Art Multimodal Retrieval and RankingThe Qwen3-VL-Embedding and Qwen3-VL-Reranker models form an end-to-end multimodal search pipeline, leveraging multi-stage training and cross-attention mechanisms to achieve high-precision retrieval across diverse modalities.Alibaba (Qwen) · Jan 2026 · Unverified9 months ago255
Aya Dataset: An Open-Access Collection for Multilingual Instruction TuningThe initiative builds a human-curated instruction-following dataset spanning 65 languages and creates the largest multilingual collection of instruction-following instances through templating and translating existing datasets across 114 languages, contributing datasets and platforms for participatory research.Cohere · Feb 20242 years ago236
DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree SearchDeepSeek-Prover-V1.5 improves theorem proving by optimizing training and inference, utilizing reinforcement learning, and proposing RMaxTS for diverse proof paths, achieving state-of-the-art results on miniF2F and ProofNet benchmarks.DeepSeek · Aug 2024 · Unverified2 years ago210
About this paper
Authors
Chuxuan Hu, Hejie Cui, Norman Huang and 3 more
arXiv
2610.00964 · PDF
Citations
Not counted yet · Semantic Scholar
Upvotes
31 · Hugging Face
Lab
Google · on Companies · on Acquisitions · on Paydays · on TechConf · on Releases

Changes

What changed
New paperFound by the weekly scan, unverifiedOct 5, 2026today

New papers by email

Monday afternoons, only in weeks with new papers from the labs.

Double opt-in. Unsubscribe any time.