Skip to content
Papers.

CHIMERA: Compact Synthetic Data for Generalizable LLM Reasoning research paper by Apple, 2026

Apple · Mar 1, 2026 · Reasoning · 2 citations · 57 upvotes · unverified

Read on arXiv

What it shows

A synthetic reasoning dataset called CHIMERA is introduced to overcome data-centric challenges in training large language models for cross-domain reasoning, achieving performance comparable to much larger models.

UnverifiedHugging Face's summary; not yet checked by hand.

More from Apple

All 9
PaperCitations
MintAct: A Unified Visual Agent for Digital EnvironmentsWe present MintAct, a family of vision-language models that unifies UI grounding, multi-step navigation across mobile, desktop, and web, and visual tool use, trained at 2B, 4B, and 8B scales.Agents and evaluation · Sep 2026 · Unverified0
It Takes Two to Match: Co-Evolving Generative Retriever with Reinforcement LearningCoGR trains LLMs to generate compact keywords for both queries and items, enabling direct inverted-index retrieval optimized via co-evolving reinforcement learning.Retrieval and data · Sep 2026 · Unverified1
Embarrassingly Simple Self-Distillation Improves Code GenerationSimple self-distillation improves code generation in large language models by fine-tuning on model-generated samples, effectively addressing precision-exploration trade-offs in decoding.Foundation models · Apr 2026 · Unverified33
Sharp Monocular View Synthesis in Less Than a SecondSHARP synthesizes photorealistic views from a single image using a 3D Gaussian representation, achieving state-of-the-art results with rapid processing.Multimodal and robotics · Dec 2025 · Unverified21
One Layer Is Enough: Adapting Pretrained Visual Encoders for Image GenerationFAE, a framework using a feature auto-encoder and dual decoders, adapts pre-trained visual representations for generative models, achieving high performance in image generation tasks.Multimodal and robotics · Dec 2025 · Unverified24
STARFlow-V: End-to-End Video Generative Modeling with Normalizing FlowSTARFlow-V, a normalizing flow-based video generator, offers end-to-end learning, robust causal prediction, and high-quality video generation with practical sampling efficiency.Multimodal and robotics · Nov 2025 · Unverified9
CLaRa: Bridging Retrieval and Generation with Continuous Latent ReasoningCLaRa enhances retrieval-augmented generation by introducing unified embedding-based compression and joint optimization, achieving state-of-the-art performance in QA benchmarks.Reasoning · Nov 2025 · Unverified11
Pico-Banana-400K: A Large-Scale Dataset for Text-Guided Image EditingPico-Banana-400K is a large-scale, high-quality dataset for instruction-based image editing, featuring diverse edit pairs, multi-turn editing, preference subsets, and long-short instruction pairs, enabling comprehensive research and benchmarking.Retrieval and data · Oct 2025 · Unverified62
Topic
PaperCitations
Chain-of-Thought Prompting Elicits Reasoning in Large Language ModelsAsking a model to write out its intermediate steps (chain of thought) sharply improves its math and logic answers.Google · Jan 202222.1k
DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language ModelsDeepSeekMath 7B improves mathematical reasoning through enhanced data pre-training and Group Relative Policy Optimization, achieving high scores on MATH benchmark without external tools.DeepSeek · Feb 2024 · Unverified8,994
DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement LearningReinforcement learning taught a model long step-by-step reasoning on par with OpenAI o1.DeepSeek · Jan 20255,694
DeepSeek-Coder-V2: Breaking the Barrier of Closed-Source Models in Code IntelligenceDeepSeek-Coder-V2, a Mixture-of-Experts language model, excels in code-specific tasks by enhancing coding and mathematical reasoning capabilities while expanding language support and context length.DeepSeek · Jun 2024 · Unverified502
MolmoAct: Action Reasoning Models that can Reason in SpaceAction Reasoning Models (ARMs) integrate perception, planning, and control to enable adaptable and explainable robotic behavior, achieving superior performance across various tasks and settings.Ai2 · Aug 2025 · Unverified185
ThinkAct: Vision-Language-Action Reasoning via Reinforced Visual Latent PlanningThinkAct, a dual-system framework, uses reinforced visual latent planning to enable few-shot adaptation, long-horizon planning, and self-correction in embodied AI tasks by bridging high-level reasoning with low-level action execution.NVIDIA · Jul 2025 · Unverified172
About this paper
Authors
Xinyu Zhu, Yihao Feng, Yanchao Sun and 5 more
arXiv
2603.00889 · PDF
Venue
arXiv.org
Citations
2, 0 influential · Semantic Scholar
Upvotes
57 · Hugging Face
Lab
Apple · on Companies · on Quarterly · on Paydays · on TechConf

Changes

What changed
Influential citationsfirst count: 0Sep 25, 2026
Citationsfirst count: 2Sep 25, 2026
New paperFound by the weekly scan, unverifiedSep 25, 2026

Sources: each lab's own papers and arXiv, with citation and upvote counts from Semantic Scholar and Hugging Face. One-line summaries are for orientation, not a substitute for the paper. Logos via logo.dev; trademarks belong to their owners.

New papers by email

Monday afternoons, only in weeks with new papers from the labs.

Double opt-in. Unsubscribe any time.