Efficient Reasoning Training Does Not Always Harm CoT Faithfulness and Monitorability Paper • 2610.03509 • Published 9 days ago • 17
Synthetic Pre-pretraining Survives Scale, but Not as a Grammatical Prior Paper • 2609.39827 • Published 11 days ago • 13
Logic Before Language: Pre-pretraining on Formal Derivations Fosters Skill Acquisition and Compressibility Paper • 2608.03930 • Published Aug 4 • 10
view article Article There is no such thing as a tokenizer-free lunch catherinearnett • Sep 25, 2025 • 103
EmpiriGraph-Psy: A Dataset and LLM Pipeline for Extracting Empirical Relation Graphs from Psychology Abstracts Paper • 2606.08362 • Published Jun 6 • 2
Inference Optimized Checkpoints (with Model Optimizer) Collection A collection of generative models quantized and optimized for inference with Model Optimizer. • 98 items • Updated 23 days ago • 190
Compliance versus Sensibility: On the Reasoning Controllability in Large Language Models Paper • 2604.27251 • Published Apr 29 • 10
Fundamental Reasoning Paradigms Induce Out-of-Domain Generalization in Language Models Paper • 2602.08658 • Published Feb 9 • 13
Beyond RAG for Agent Memory: Retrieval by Decoupling and Aggregation Paper • 2602.02007 • Published Feb 2 • 20
No Shortcuts to Culture: Indonesian Multi-hop Question Answering for Complex Cultural Understanding Paper • 2602.03709 • Published Feb 3 • 8
Youtu-LLM: Unlocking the Native Agentic Potential for Lightweight Large Language Models Paper • 2512.24618 • Published Dec 31, 2025 • 156
An Empirical Study on Preference Tuning Generalization and Diversity Under Domain Shift Paper • 2601.05882 • Published Jan 9 • 21
Enhancing Linguistic Competence of Language Models through Pre-training with Language Learning Tasks Paper • 2601.03448 • Published Jan 6 • 13
Olmo 3 Pre-training Collection All artifacts related to Olmo 3 pre-training • 10 items • Updated Dec 23, 2025 • 36