LLM Engineering
How LLM applications get built and shipped: inference and serving, structured output, token formats, evaluation, and the libraries powering production language-model systems.
149 articles
LLM Engineering
AgentOps: The Missing Observability Layer for Production AI Agents
LLM Engineering
Inside Inspect AI: How the UK Government Built a Framework for LLM Safety Evaluations
LLM Engineering
E2B Desktop Sandbox: Building LLM Agents That Actually Control Computers
LLM Engineering
Removing AI Safety Guardrails with a Single Vector: Inside Refusal Direction Research
LLM Engineering
Building Production LLM Systems: A Deep Dive into the LLM Engineer's Handbook Reference Architecture
LLM Engineering
LLM: The Unix Philosophy Meets Large Language Models
LLM Engineering
Inside the LLM Engineer Handbook: A Production-Ready Mental Map of the AI Toolchain
LLM Engineering
LLaVA-CoT: Teaching Vision-Language Models to Show Their Work
LLM Engineering
Brainstorm: Teaching LLMs to Predict Hidden Web Endpoints
LLM Engineering
HunyuanVideo: Tencent's Diffusion Transformer Architecture for 720p Video Generation
LLM Engineering
Building Statistically Robust LLM Rankings with Pairwise Comparisons
LLM Engineering
BitNet: Running 100B Parameter Models on Your Laptop at Human Reading Speed
LLM Engineering
LLM-Check: Detecting Hallucinations by Reading Your Model's Mind
LLM Engineering
Mapping LLM Safety as a Landscape: How Weight Perturbations Reveal the Fragility of Alignment
LLM Engineering
VERL: The Hybrid-Controller Framework Reshaping How We Train LLMs with Reinforcement Learning
LLM Engineering
IB4LLMs: Using Information Bottleneck Theory to Build Jailbreak-Resistant Language Models
LLM Engineering
SRMT: Teaching Robots to Share Their Thoughts Through Memory
LLM Engineering
Building Resumable LLM Evaluations: A Template for Rate-Limited API Testing
LLM Engineering
Burpference: Adding LLM Intelligence to Your Security Proxy Workflow
LLM Engineering
LLMmap: Fingerprinting Large Language Models Through Behavioral Analysis
LLM Engineering
Inside Transformer Debugger: OpenAI's Circuit Tracing Tool for Mechanistic Interpretability
LLM Engineering
Dora-VAE: How Inference-Time Scalability Solves the 3D Diffusion Training Bottleneck
LLM Engineering
Repomix: Why Packing Your Entire Codebase Into One File Is the Future of LLM Workflows
LLM Engineering