Overview
Moonshot’s multimodal reasoning work exploring long-context reinforcement learning and test-time reasoning scale.
Moonshot’s multimodal reasoning work exploring long-context reinforcement learning and test-time reasoning scale.
Only published specifications are shown.
Moonshot’s multimodal reasoning work exploring long-context reinforcement learning and test-time reasoning scale.
K1.5 combined long context with RL-based reasoning and provided a research base for later open Kimi agent models.
Moonshot published its Kimi K1.5 multimodal reasoning work.
View sourceSelect a card to view the related model.