Running 44 T-Search: an open agentic retriever 🔎 44 Open agentic retriever for hard multi-step search
Rank-Then-Act: Reward-Free Control from Frame-Order Progress Paper • 2607.01897 • Published Jul 2 • 7
Qantara: Bridge-Flow Training for Multi-Paradigm JEPA Control Paper • 2607.04978 • Published Jul 6 • 7
Qantara: Bridge-Flow Training for Multi-Paradigm JEPA Control Paper • 2607.04978 • Published Jul 6 • 7
Unstable Features, Reproducible Subspaces: Understanding Seed Dependence in Sparse Autoencoders Paper • 2606.12138 • Published Jun 10 • 8
Interpreting and Steering a Text-to-Speech Language Model with Sparse Autoencoders Paper • 2606.10029 • Published Jun 8 • 12
Interpreting and Steering a Text-to-Speech Language Model with Sparse Autoencoders Paper • 2606.10029 • Published Jun 8 • 12
Train One Sparse Autoencoder Across Multiple Sparsity Budgets to Preserve Interpretability and Accuracy Paper • 2505.24473 • Published May 30, 2025
Small Vectors, Big Effects: A Mechanistic Study of RL-Induced Reasoning via Steering Vectors Paper • 2509.06608 • Published Sep 8, 2025
Running Featured 25 Chasing the Counting Manifold in Open LLMs 📚 25 Counting manifolds in open LLMs from behavior to SAEs.
F-GRPO: Don't Let Your Policy Learn the Obvious and Forget the Rare Paper • 2602.06717 • Published Feb 6 • 76
F-GRPO: Don't Let Your Policy Learn the Obvious and Forget the Rare Paper • 2602.06717 • Published Feb 6 • 76
T-pro 2.0: An Efficient Russian Hybrid-Reasoning Model and Playground Paper • 2512.10430 • Published Dec 11, 2025 • 121