This Skill Just Made Kimi K3 A 10x Better Designer
🔗 Link do vídeo: https://www.youtube.com/watch?v=EwTOiqWWqEc
🆔 ID do vídeo: EwTOiqWWqEc
📅 Publicado em: 2026-07-25T17:16:26Z
📺 Canal: AI LABS
⏱️ Duração (ISO): PT15M15S
⏱️ Duração formatada: 00:15:15
📊 Estatísticas:
– Views: 4.315
– Likes: 224
– Comentários: 11
🏷️ Tags:
Kimi K3 review: Moonshot AI's new model tops LMArena for frontend design. Kimi K3 vs Claude, the Kimi K3 benchmarks, why Kimi Code is slow, and the skill that strips AI slop out of every design. We tested it all in Claude Code.
Hallmark Skill: https://github.com/Nutlope/hallmark
Community with All Resources: http://ailabspro.io
The Roundup, our daily newsletter covering the AI stories that matter. Join now: https://www.theroundup.so/
Every AI model has its own design style, and you don't notice it until you've used one enough. Kimi K3 is genuinely good at frontend, but it still has patterns it defaults to. So this is the full Kimi K3 test: what Moonshot AI actually shipped, how to run it properly, and the skill that pushes any model off its defaults.
The Kimi K3 benchmarks (Kimi K3 vs Fable 5, GPT 5.6 and Claude)
– A million-token context window, and on Artificial Analysis intelligence it sits ahead of Opus 4.8 and Gemini 3.6, just behind Claude Fable 5 and GPT 5.6, by a margin small enough to call it the same tier.
– On frontend design it beats both on LMArena, largely because of vision-in-the-loop: it screenshots what it built, looks at the result, and adjusts, instead of guessing from the code.
– Price is where it wins outright: $3/$15 per million tokens vs $5/$30 for GPT 5.6 and $10/$50 for Fable 5.
Why we didn't use Kimi Code
– Kimi Code is Moonshot's own terminal agent, launched with K2.5. A task that takes Claude Code or Codex ~3 minutes took it closer to 10.
– The weights aren't open yet, so every request goes back to Kimi's own overloaded servers, and Moonshot admits in their docs the harness isn't built to bring out K3's full potential.
– Caching breaks when you switch models, and its sub-agent reporting didn't match what we saw run.
How to use Kimi K3 in Claude Code (the Kimi K3 installation)
– Kimi K3 Claude Code setup: point the Anthropic base URL at the Kimi endpoint, swap the auth token for your key, set the model to kimi-k3, run claude.
– Looking for free Kimi K3 or a free Kimi K3 API? There isn't one, but you don't have to pay per token either. CLIProxyAPI turns the subscription you already pay for into a local API on your own machine, so you point Claude Code at localhost instead of the direct API and never think about a bill.
– All of it lives in that one terminal session, so closing it puts you straight back on your normal Claude subscription. Max reasoning is on by default.
The Hallmark skill (Kimi K3 design without the slop)
– An anti-AI-slop design skill with four verbs: default, audit, redesign, and study. Study is the interesting one, it treats a reference site as direction instead of cloning it.
– 100+ references and a 58-gate check before it hands anything back, plus a style library you can browse as real landing pages.
– Install it into .agents for Codex and Kimi Code, or .claude for Claude Code. With Kimi, invoke it manually with the slash command, auto-invocation isn't reliable on a non-Claude model.
What we found testing it
– The Kimi K3 website we built showed Opus 4.8's fingerprints (distillation), images behind the hero, oversized offset headings, warm orange and brown palettes, but the copy is far less hype-stuffed than Claude's.
– Auto-compaction doesn't run with Kimi in Claude Code, so watch your context or the answers drift generic.
– Same test on Opus 4.8: without the skill it was heavy slop (gradients, rounded boxes, the usual). With Hallmark, the result was more intentional and, right now, better than Kimi's.
– Codex with the skill dropped its green-and-white default, used SVGs more sparingly, and ran an interactive browser test on its own.
We're a software company running these models on our own products, so this is a real Kimi K3 test, not a launch-day reaction. Whether you searched Kimi K3 vs Claude, Kimi K3 vs Fable 5, how to use Kimi K3, Kimi K3 benchmarks, or free Kimi K3, this is the setup to copy. Works the same whether you build with Claude Code, Codex, or ChatGPT. Everything we use is in AI Labs Pro. Subscribe and hit the hype button.
Hashtags:
#ai #claude #claudeCode #fable5 #kimiK3 #gpt56 #chatgpt #claudeFable5