LLM Engineering
How LLM applications get built and shipped: inference and serving, structured output, token formats, evaluation, and the libraries powering production language-model systems.
149 articles
LLM Engineering
OSWorld-V2: The GUI Agent Benchmark That Hides Its Answers
LLM Engineering
Running a 2.78-Trillion-Parameter Model on a Laptop: Inside WASTE's NVMe-Streaming Engine
LLM Engineering
Running 753B MoE Models on Consumer GPUs: Hand-Written SASS Kernels for 2-Bit Experts
LLM Engineering
LLM Checker: Hardware-Aware Model Selection for Local Inference
LLM Engineering
Fighting LLM Hallucinations in 2026: llm-council vs claude-octopus vs LLM-Check
LLM Engineering
Running Large LLMs on Limited Hardware in 2026: AirLLM vs llm-checker
LLM Engineering
The Fragile Economics of Free AI: Mapping the Great Inference Subsidy Wars
LLM Engineering
Shard: Proving LLM Inference Can Work Across Scattered GPUs and Terrible Internet
LLM Engineering
Nanocoder: The Terminal Coding Agent That Lets You Switch Models Mid-Conversation
LLM Engineering
ds4: The SSD-Streaming Inference Engine That Treats Your Mac's NVMe Like RAM
LLM Engineering
Harness-1: Training Search Agents with State Externalization
LLM Engineering
SichGate Methodology: When Healthcare CISOs Need to Red-Team 4-Bit Llama Without Hiring Offensive Security
LLM Engineering
ModelRegression: Building a Daily LLM Benchmark That Tests What Developers Actually Use
LLM Engineering
Inside AI Product Bench: Why Two LLMs Disagree on Half Their Product Recommendations
LLM Engineering
Neuromod-LLM: Treating Language Models Like Brains on Drugs
LLM Engineering
makemore: Understanding Language Models by Implementing Them Seven Different Ways
LLM Engineering
FastLLM: How a Single Line of Code Can Sabotage AI Reasoning While Improving Benchmarks
LLM Engineering
whichllm: Hardware-Aware LLM Selection Using Evidence-Graded Benchmarks
LLM Engineering
JARVIS: The LLM-Orchestrated AI System That Pioneered Multi-Model Task Automation
LLM Engineering
Guardrails AI: Building Fail-Safe Layers Around Unpredictable LLMs
LLM Engineering
MosaicML Composer: The PyTorch Training Framework That Makes Checkpoints Hardware-Agnostic
LLM Engineering
Building a Codebase Documentation Engine with LLMs: Lessons from auto_llm_codebase_analysis
LLM Engineering
How NYU Built a Leaderboard to Track LLM Agents Hacking Their Way Through CTF Challenges
LLM Engineering