> your AI agent picks dependencies from memory; give it dated facts — try starlog.dev ↗ vet your agent's deps ↗ vibe-coding is fine. vibe-importing isn’t. — try starlog.dev ↗ vibe-importing isn’t fine ↗ your agent has never seen your private packages — try starlog.dev ↗ facts for private packages ↗ a linter for the dependencies your AI agent picks — try starlog.dev ↗ a linter for agent deps ↗

All articles

LLM Engineering

How LLM applications get built and shipped: inference and serving, structured output, token formats, evaluation, and the libraries powering production language-model systems.

149 articles

LLM Engineering

OSWorld-V2: The GUI Agent Benchmark That Hides Its Answers

★ 229 Python Aug 5, 2026
LLM Engineering

Running a 2.78-Trillion-Parameter Model on a Laptop: Inside WASTE's NVMe-Streaming Engine

★ 145 C Jul 31, 2026
LLM Engineering

Running 753B MoE Models on Consumer GPUs: Hand-Written SASS Kernels for 2-Bit Experts

★ 253 Sass Jul 10, 2026
LLM Engineering

LLM Checker: Hardware-Aware Model Selection for Local Inference

★ 2.8k JavaScript Jun 29, 2026
LLM Engineering

Fighting LLM Hallucinations in 2026: llm-council vs claude-octopus vs LLM-Check

Various Jun 22, 2026
LLM Engineering

Running Large LLMs on Limited Hardware in 2026: AirLLM vs llm-checker

Various Jun 22, 2026
LLM Engineering

The Fragile Economics of Free AI: Mapping the Great Inference Subsidy Wars

★ 696 Jun 20, 2026
LLM Engineering

Shard: Proving LLM Inference Can Work Across Scattered GPUs and Terrible Internet

★ 19 Python Jun 18, 2026
LLM Engineering

Nanocoder: The Terminal Coding Agent That Lets You Switch Models Mid-Conversation

★ 2.1k TypeScript Jun 14, 2026
LLM Engineering

ds4: The SSD-Streaming Inference Engine That Treats Your Mac's NVMe Like RAM

★ 13.4k C Jun 11, 2026
LLM Engineering

Harness-1: Training Search Agents with State Externalization

★ 390 Python Jun 9, 2026
LLM Engineering

SichGate Methodology: When Healthcare CISOs Need to Red-Team 4-Bit Llama Without Hiring Offensive Security

★ 2 Python Jun 8, 2026
LLM Engineering

ModelRegression: Building a Daily LLM Benchmark That Tests What Developers Actually Use

★ 12 Python Jun 6, 2026
LLM Engineering

Inside AI Product Bench: Why Two LLMs Disagree on Half Their Product Recommendations

★ 23 HTML May 27, 2026
LLM Engineering

Neuromod-LLM: Treating Language Models Like Brains on Drugs

★ 6 Python May 24, 2026
LLM Engineering

makemore: Understanding Language Models by Implementing Them Seven Different Ways

★ 4.0k Python May 24, 2026
LLM Engineering

FastLLM: How a Single Line of Code Can Sabotage AI Reasoning While Improving Benchmarks

★ 3 Python May 17, 2026
LLM Engineering

whichllm: Hardware-Aware LLM Selection Using Evidence-Graded Benchmarks

★ 516 Python May 15, 2026
LLM Engineering

JARVIS: The LLM-Orchestrated AI System That Pioneered Multi-Model Task Automation

★ 24.7k Python May 15, 2026
LLM Engineering

Guardrails AI: Building Fail-Safe Layers Around Unpredictable LLMs

★ 6.9k Python May 15, 2026
LLM Engineering

MosaicML Composer: The PyTorch Training Framework That Makes Checkpoints Hardware-Agnostic

★ 5.5k Python May 15, 2026
LLM Engineering

Building a Codebase Documentation Engine with LLMs: Lessons from auto_llm_codebase_analysis

★ 27 Python May 15, 2026
LLM Engineering

How NYU Built a Leaderboard to Track LLM Agents Hacking Their Way Through CTF Challenges

★ 5 Shell May 15, 2026
LLM Engineering

Terminal-Bench: Why Evaluating LLM Agents on Real Command-Line Tasks Is Harder Than You Think

★ 2.2k Python May 15, 2026