Kimi K3 tops frontend code rankings but trails badly on hard math
Moonshot's Kimi K3 sits at #1 on the Code Arena Frontend leaderboard, but scores only ~39% on FrontierMath Tier 4 versus ~90% for top OpenAI and Anthropic models.
Moonshot's Kimi K3 sits at #1 on the Code Arena Frontend leaderboard, but scores only ~39% on FrontierMath Tier 4 versus ~90% for top OpenAI and Anthropic models.
Running Qwen3.6 27B with Opencode? The model's appetite for context tokens can gut your usable window fast. Here's what's happening and how to push back.
xAI's coding agent got caught uploading home directories, SSH keys and all, to cloud buckets. The lesson is not "never trust xAI." It is that "local" describes what leaves your machine, and almost nobody checks.