
Community Discussion · Policy
What Open-Weight LLMs Mean for Reinforcement Learning Researchers
What caught my attention most about the release of Kimi K3 wasn't the rumored 2.8 trillion parameter count, but the statement that "weights will be released soon." In the field of reinforcement learning (RL), we've been plagued by two issues for a long time: first, we can't get the base model weights, making it impossible to do RL training in real-world scenarios; second, even if we get API access, we can't control the internal representations of the model. The open weights of K3 might be the friendliest signal for RL researchers in recent years.
Physix Frontier