| Sep 24, 2026 | Towards Physics of Multimodal Pretraining: Knowledge Flow, Modality Synergy, Early Unification, and RecipesSep 24, 2026 | New paperAdded to the list |
| Sep 24, 2026 | StudentSim: Training LLM-based Student SimulatorsSep 24, 2026 | New paperAdded to the list |
| Sep 24, 2026 | MolmoAct2: Action Reasoning Models for Real-world DeploymentSep 24, 2026 | New paperAdded to the list |
| Sep 24, 2026 | On the Design of Qwen3.8-Next Architecture: Evaluation, Efficiency, and Training StabilitySep 24, 2026 | New paperAdded to the list |
| Sep 24, 2026 | Strategic Navigation or Stochastic Search? How Agents and Humans Reason Over Document CollectionsSep 24, 2026 | New paperAdded to the list |
| Sep 24, 2026 | RRSI: Regularized Recursive Self-Improvement of Agent HarnessesSep 24, 2026 | New paperAdded to the list |
| Sep 24, 2026 | Agents' Last ExamSep 24, 2026 | New paperAdded to the list |
| Sep 24, 2026 | DeepSeek-V4.1-Flash: Pushing the Limits of KV Cache CompressionSep 24, 2026 | New paperAdded to the list |
| Sep 24, 2026 | Cosmos 3: Omnimodal World Models for Physical AISep 24, 2026 | New paperAdded to the list |
| Sep 24, 2026 | Gemma 4 Technical ReportSep 24, 2026 | New paperAdded to the list |
| Sep 24, 2026 | Mamba: Linear-Time Sequence Modeling with Selective State SpacesSep 24, 2026 | New paperAdded to the list |
| Sep 24, 2026 | Direct Preference Optimization: Your Language Model is Secretly a Reward ModelSep 24, 2026 | New paperAdded to the list |
| Sep 24, 2026 | Efficient Memory Management for Large Language Model Serving with PagedAttentionSep 24, 2026 | New paperAdded to the list |
| Sep 24, 2026 | FlashAttention: Fast and Memory-Efficient Exact Attention with IO-AwarenessSep 24, 2026 | New paperAdded to the list |
| Sep 24, 2026 | DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement LearningSep 24, 2026 | New paperAdded to the list |
| Sep 24, 2026 | SwiftKV: Fast Prefill-Optimized Inference with Knowledge-Preserving Model TransformationSep 24, 2026 | New paperAdded to the list |
| Sep 24, 2026 | Arctic-Embed: Scalable, Efficient, and Accurate Text Embedding ModelsSep 24, 2026 | New paperAdded to the list |
| Sep 24, 2026 | LoRA Learns Less and Forgets LessSep 24, 2026 | New paperAdded to the list |
| Sep 24, 2026 | Nemotron-4 340B Technical ReportSep 24, 2026 | New paperAdded to the list |
| Sep 24, 2026 | Megatron-LM: Training Multi-Billion Parameter Language Models Using Model ParallelismSep 24, 2026 | New paperAdded to the list |
| Sep 24, 2026 | From Local to Global: A Graph RAG Approach to Query-Focused SummarizationSep 24, 2026 | New paperAdded to the list |
| Sep 24, 2026 | Phi-3 Technical Report: A Highly Capable Language Model Locally on Your PhoneSep 24, 2026 | New paperAdded to the list |
| Sep 24, 2026 | LoRA: Low-Rank Adaptation of Large Language ModelsSep 24, 2026 | New paperAdded to the list |
| Sep 24, 2026 | Segment AnythingSep 24, 2026 | New paperAdded to the list |
| Sep 24, 2026 | The Llama 3 Herd of ModelsSep 24, 2026 | New paperAdded to the list |
| Sep 24, 2026 | LLaMA: Open and Efficient Foundation Language ModelsSep 24, 2026 | New paperAdded to the list |
| Sep 24, 2026 | Gemini 1.5: Unlocking multimodal understanding across millions of tokens of contextSep 24, 2026 | New paperAdded to the list |
| Sep 24, 2026 | Chain-of-Thought Prompting Elicits Reasoning in Large Language ModelsSep 24, 2026 | New paperAdded to the list |
| Sep 24, 2026 | Training Compute-Optimal Large Language ModelsSep 24, 2026 | New paperAdded to the list |
| Sep 24, 2026 | Why Language Models HallucinateSep 24, 2026 | New paperAdded to the list |
| Sep 24, 2026 | Constitutional Classifiers: Defending against Universal Jailbreaks across Thousands of Hours of Red TeamingSep 24, 2026 | New paperAdded to the list |
| Sep 24, 2026 | Alignment faking in large language modelsSep 24, 2026 | New paperAdded to the list |
| Sep 24, 2026 | Sleeper Agents: Training Deceptive LLMs that Persist Through Safety TrainingSep 24, 2026 | New paperAdded to the list |
| Sep 24, 2026 | Constitutional AI: Harmlessness from AI FeedbackSep 24, 2026 | New paperAdded to the list |
| Sep 24, 2026 | GPT-4 Technical ReportSep 24, 2026 | New paperAdded to the list |
| Sep 24, 2026 | Training language models to follow instructions with human feedbackSep 24, 2026 | New paperAdded to the list |
| Sep 24, 2026 | Scaling Laws for Neural Language ModelsSep 24, 2026 | New paperAdded to the list |
| Sep 24, 2026 | Language Models are Few-Shot LearnersSep 24, 2026 | New paperAdded to the list |
| Sep 24, 2026 | BERT: Pre-training of Deep Bidirectional Transformers for Language UnderstandingSep 24, 2026 | New paperAdded to the list |
| Sep 24, 2026 | Attention Is All You NeedSep 24, 2026 | New paperAdded to the list |