Articles, one at a time.
Every piece cites primary sources and adds something you can act on: a worked example, a tested command, a documented limitation. Drafts that only summarize what already exists stay unpublished.
LLM Inference Engines Compared 2026: vLLM vs SGLang vs TGI vs MAX
A source-verified 2026 decision guide for vLLM, SGLang, TGI, and MAX, with use/skip guidance and deployment tradeoffs.
Read →
Qwen3.6-Plus: 1M Token Context and Claude-Level Performance
Alibaba's Qwen3.6-Plus: 1M token context, agentic coding, hybrid MoE, ~$0.29/M input. Sourced benchmarks vs Claude Opus 4.7 and a when-to-skip guide.
Read →
Fine-Tune LLMs with LoRA and QLoRA: 2026 Guide
Learn to fine-tune LLMs with LoRA and QLoRA in 2026. VRAM requirements, dataset prep, Unsloth/Axolotl setup, hyperparameters, and evaluation.
Read →
Langfuse: Self-Host LLM Observability for Free — 2026 Guide
Deploy Langfuse free with Docker Compose. Open-source LLM observability covering traces, evals, prompt management, and Kubernetes scaling.
Read →
LiteLLM: One Proxy for 140+ LLMs — Setup & Cost Guide
LiteLLM unifies 100+ LLM APIs behind one OpenAI-compatible endpoint. Learn to self-host, control costs, and set provider fallbacks in 2026.
Read →
Goose by Block: A Free, Open-Source AI Agent Review 2026
An in-depth review of Goose, Block's Apache 2.0 AI agent. Compare it to Claude Code, explore MCP extensions, Recipes, and local Ollama setup.
Read →
OpenAI Codex CLI: Terminal Coding Agent Setup Guide 2026
Complete guide to OpenAI Codex CLI — setup, safety modes, sandboxing, and how it compares to Claude Code in 2026.
Read →
Qwen3 Review: Hybrid Thinking Modes and MoE Architecture Explained
Qwen3 ships hybrid thinking/non-thinking modes, MoE variants up to 235B, and Apache 2.0 licensing. Developer guide with benchmarks, setup, and API pricing.
Read →
MCP Ecosystem 2026: What the Adoption Numbers Actually Show
MCP adoption traced to primary sources, plus a census of the official registry: 22,408 active servers, and how many are remote-only.
Read →
Cloud Dev Environments Compared: Codespaces vs Gitpod vs CodeSandbox
Compare GitHub Codespaces, Gitpod, and CodeSandbox for cloud development. Pricing, features, performance, and which to choose in 2026.
Read →
GLM-5: The Open-Source Frontier Model You Can Self-Host
GLM-5 is an MIT-licensed frontier model with top-5 benchmark scores. Learn how to self-host it and compare it with GPT-5 and Claude.
Read →
AI Agent Frameworks Compared 2026: LangGraph vs CrewAI vs OpenAI SDK
Compare the top AI agent frameworks in 2026 — LangGraph, CrewAI, OpenAI Agents SDK, and more — with code examples and use-case recommendations.
Read →