Control-Diverse Reinforcement Fine-Tuning: Decoupling the Shared Control Bottleneck of RL Post-Training
Published in arXiv preprint (under review, AAAI 2027), 2026
RL post-training is usually explained by which circuits it activates. We separate activation from control, and find that control concentrates on a shared set of components across tasks even when activations look diverse.
Recommended citation: Tan, B., Wang, J., Hou, D., Jiang, L., Wu, Z., Shen, Y., Lin, F., Yamada, K., & Koike, A. (2026). Control-Diverse Reinforcement Fine-Tuning: Decoupling the Shared Control Bottleneck of RL Post-Training. arXiv preprint arXiv:2608.08224.
Download Paper
