LLM Engineering
How LLM applications get built and shipped: inference and serving, structured output, token formats, evaluation, and the libraries powering production language-model systems.
149 articles
LLM Engineering
CircuitStream: A Lightweight LLM Proxy for Teams Sharing Rate Limits
LLM Engineering
Building a Slack Doppelgänger: Fine-Tuning LLMs on Your Message History with Modal
LLM Engineering
Inside the Foundation Model Transparency Index: How Stanford Scores AI Giants on Disclosure
LLM Engineering
SimplyRetrieve: Building Privacy-First RAG Systems That Treat LLMs as Context Interpreters, Not Oracles
LLM Engineering
Deconstructing Sparse Matrix Performance: A Case Study in Rust Optimization
LLM Engineering
LLM-CLI: A Crystal-Fast Command Generator That Learns What You Mean
LLM Engineering
Instructor: How Pydantic Models Turned LLM JSON into Type-Safe Python Objects
LLM Engineering
Inside Microsoft's LMOps: Research Prototypes That Reveal How LLMs Actually Work
LLM Engineering
Paxml: Google's JAX Framework for Training Models at Trillion-Parameter Scale
LLM Engineering
BishopFox/llm-testing-findings: The Missing Standard for Documenting AI Security Vulnerabilities
LLM Engineering
LitGPT: The Zero-Abstraction Framework for Production LLM Training
LLM Engineering
getML: The C++ Engine That Makes Feature Engineering on Relational Data 1000x Faster
LLM Engineering
Weaponizing Machine Learning Models: How Pickle Deserialization Turns TensorFlow into a Trojan Horse
LLM Engineering
Inside GCG: The Gradient-Based Attack That Broke LLM Alignment
LLM Engineering
AutoRedTeam: Training Language Models to Attack Other Language Models
LLM Engineering
MiniHF: Building Domain-Specific Language Models Through Constitutional AI and Tree Search
LLM Engineering
Zep: Why Temporal Knowledge Graphs Beat Vector Databases for AI Agent Memory
LLM Engineering
Red-Teaming LLMs with Systematic Prompt Perturbation: Inside Fiddler Auditor
LLM Engineering
How Large Language Models Are Learning to Think in Graphs: A Research Taxonomy
LLM Engineering
Running Mixtral-8x7B on Consumer Hardware: Expert Offloading and Mixed Quantization
LLM Engineering
Mangio-RVC-Fork: When Voice Conversion Meets Ensemble Pitch Detection
LLM Engineering
Axolotl: The Config-Driven LLM Fine-Tuning Framework Racing Ahead of Research
LLM Engineering
LLM Sherpa: How Smart Chunking Fixes RAG's Biggest Problem
LLM Engineering