Tutorials Self-Host Kimi K3 Open Weights With vLLM
Run Moonshot's 2.8T open-weight model on your own GPUs with vLLM and MXFP4.
How-to content for builders, indie hackers, and AI engineers. Less theory, more shipped code.
Tutorials Run Moonshot's 2.8T open-weight model on your own GPUs with vLLM and MXFP4.
Tutorials Point the OpenAI SDK at Meta's agent model, add tools, let it self-manage a 1M-token context.
Tutorials Install Nous Research's self-improving agent, run it on a local model, and build a reusable skill.
Tutorials Self-host OpenClaw, wire up Telegram, and ship a SKILL.md skill your agent triggers on its own.
Tutorials Build a semantic cache that reuses answers for similar prompts and slashes LLM API costs.
Tutorials Enable Claude Code's experimental Agent Teams: peer messaging and a self-claiming shared task list.
Tutorials Build a tool-using agent on Anthropic's Claude Fable 5 that plans, acts, and verifies its own work.
Tutorials Deploy any agent framework to Microsoft Foundry's managed sandbox runtime (Build 2026 launch).
Tutorials Build a safe local agent harness with shell, files, approvals, and logs in Python.
Tutorials Use Gemini 3.5 Flash code execution to write Python that fixes itself in a loop.
Machine Learning Fine-tune 7B LLMs on one 24GB GPU with 70% less VRAM
Machine Learning Use GRPO to teach a 0.5B model multi-step math reasoning end to end.