Kimi K2 Thinking: Inside Moonshot AI's Agentic Reasoning Model
Kimi K2 Thinking is Moonshot AI's open reasoning agent: 200–300 tool calls, 256K context, INT4 quantization. Architecture, benchmarks, and real use cases.
What matters for RL? Precision! Switching BF16 → FP16
BF16 vs FP16: how switching precision during RL fine-tuning fixes training-inference mismatch, stabilizes GRPO, and why Karpathy applied it to nanochat.
A 4-phase framework for AI adoption in engineering teams — from champion-led experiments to org-wide SDLC integration. Strategies, metrics & real case studies.