Kimi.ai {bot}

@kimi-moonshot-x.bsky.social

Unofficial mirror account of https://x.com/kimi_moonshot from Twitter Built by Moonshot AI to empower everyone to be superhuman. ⚡️API: https://platform.moonshot.ai/ @KimiProduct where we share cool use cases and prompts.

Kimi Work for Financial Analysts - Tutorial #2 Use Kimi Work for 3 common investment research tasks: - Build a live investor dashboard - Update financial models in spreadsheet - Process and generate reports in batch Stay tuned for more Kimi Work workflows!

Build Slides with Kimi Work - Tutorial #1. Kimi Slides handles the entire slide-building process: - Clear structure and research, powered by Kimi K3 - Cohesive design, including polished charts and SmartArts - Editable and ready to download Let us know what you'd like to see next in the comments!

We've open-sourced MoonEP, our high-performance communication library for distributed MoE workloads. Built to make expert-parallel communication more efficient at scale, MoonEP helps reduce communication overhead in large MoE training and inference systems. Explore on GitHub:

GitHub - MoonshotAI/MoonEP: MoonEP: A Perfectly Balanced Expert Parallelism Library via Dynamic Redundant Experts

MoonEP: A Perfectly Balanced Expert Parallelism Library via Dynamic Redundant Experts - MoonshotAI/MoonEP

github.com

We've open-sourced AgentENV in collaboration with kvcache-ai. AgentENV is a distributed system for running agent environments at scale. Its components power agentic RL training for Kimi K3, with fast snapshot, resume, and fork support for large-scale parallel agent workflows. Explore on GitHub:

GitHub - kvcache-ai/AgentENV: AgentENV (AENV) is a distributed platform for running agent environments at scale.

AgentENV (AENV) is a distributed platform for running agent environments at scale. - kvcache-ai/AgentENV

github.com

We've open-sourced FlashKDA, our high-performance CUTLASS-based implementation of Kimi Delta Attention kernels. It delivers 1.72×–2.22× prefill speedup over the flash-linear-attention baseline on H20, and works as a drop-in backend for flash-linear-attention. Explore on GitHub:

GitHub - MoonshotAI/FlashKDA: FlashKDA: high-performance Kimi Delta Attention kernels

FlashKDA: high-performance Kimi Delta Attention kernels - MoonshotAI/FlashKDA

github.com

Releasing the model weights and technical report of Kimi K3. Kimi K3 is our most capable model: a 2.8T MoE model with native visual understanding and a 1M-token context window. New model architecture: 2.5x the intelligence per unit of compute, not just more params. (1/3)

Image from Twitter

The Kimi API is now live on AWS Marketplace. 🚀 If your team is already running on AWS, you can now access Kimi with consolidated billing. Plus, eligible customers can apply Kimi API usage directly toward their AWS EDP commitments. (1/2)

Image from Twitter

🌘 Meet Kimi K2.7 Code HighSpeed! A high-speed mode of our latest open-source multimodal coding model, Kimi K2.7 Code. ⚡️ Up to 6× faster: Around 180 tok/s on coding tasks with median-length inputs, and up to 260 tok/s on shorter-context tasks. (1/3)

Kimi API Platform

Kimi K2.7 Code Open Platform, providing trillion-parameter K2.7 Code large language model API, supporting 256K long context and Tool Calling. Professional code generation, intelligent dialogue, visual reasoning, helping developers build AI applications.

platform.kimi.ai

Extra API quota for Kimi K2.7 Code builders 🎉 If you're building with Kimi API, get 20%–30% extra quota when you top up $100+ by July 2! 🔷 $100–$299 → +20% quota 🔷 $300–$999 → +25% quota 🔷 $1,000+ → +30% quota (One bonus per account.) (1/2)

Image from Twitter