source avatarCyrilXBT

Share

2.8 trillion parameters. 1 million token context. Native multimodal. Kimi Delta Attention delivers up to 6.3x faster decoding at million-token context, cutting KV-cache memory by up to 75% along the way. Attention Residuals deliver roughly 25% higher training efficiency at under 2% additional compute cost. Built for long-horizon agentic coding and self-evolving workflows, sustaining long engineering sessions with minimal human oversight.

No.0 picture
No.1 picture
Disclaimer: The information on this page may have been obtained from third parties and does not necessarily reflect the views or opinions of KuCoin. This content is provided for general informational purposes only, without any representation or warranty of any kind, nor shall it be construed as financial or investment advice. KuCoin shall not be liable for any errors or omissions, or for any outcomes resulting from the use of this information. Investments in digital assets can be risky. Please carefully evaluate the risks of a product and your risk tolerance based on your own financial circumstances. For more information, please refer to our Terms of Use and Risk Disclosure.