The Obfuscation Atlas: Mapping Where Honesty Emerges in RLVR with Deception Probes Paper • 2602.15515 • Published Feb 17 • 2
The Obfuscation Atlas Collection Obfuscated Policy, Obfuscated Activations, Blatant Deception, and Honest models trained in the Obfuscation Atlas paper. • 490 items • Updated Feb 20 • 1
Blockwise Advantage Estimation for Multi-Objective RL with Verifiable Rewards Paper • 2602.10231 • Published Feb 10 • 13
view article Article Self-Hosting LLaMA 3.1 70B (or any ~70B LLM) Affordably abhinand • Aug 20, 2024 • 27
Olmo 3 Post-training Collection All artifacts for post-training Olmo 3. Datasets follow the model that resulted from training on them. • 32 items • Updated Dec 23, 2025 • 59
RLCR Collection Collection of models and datasets for Beyond Binary Rewards: Training LMs to Reason about their Uncertainty • 10 items • Updated Aug 6, 2025 • 7
Model Merging Collection Model Merging is a very popular technique nowadays in LLM. Here is a chronological list of papers on the space that will help you get started with it! • 30 items • Updated Jun 12, 2024 • 253
view article Article Everything You Need to Know about Knowledge Distillation Kseniase • Mar 6, 2025 • 86
🚀 [NeurIPS 2025] RPC Resources Collection Sampled Reasoning Paths for NeurIPS 2025 Paper: A Theoretical Study on Bridging Internal Probability and Self-Consistency for LLM Reasoning • 6 items • Updated Apr 28 • 8
view article Article Training and Finetuning Reranker Models with Sentence Transformers tomaarsen • Mar 26, 2025 • 196
view article Article Illustrating Reinforcement Learning from Human Feedback (RLHF) +2 natolambert, LouisCastricato, lvwerra, Dahoas • Dec 9, 2022 • 419
view article Article KV Caching Explained: Optimizing Transformer Inference Efficiency not-lain • Jan 30, 2025 • 378
view article Article PyTorchModelHubMixin: Bridging the Gap for Custom AI Models on Hugging Face not-lain • Nov 11, 2024 • 21