Diagram of an LLM serving as the final ranking stage in a recommender pipeline.

GenRec: An LLM-Backed Recommendation Ranker

GenRec replaces a stack of hand-engineered ranking models with a single fine-tuned LLM that scores items from natural-language context. Here is the architecture, the training loop, and what production teams should know before shipping it.

September 5, 2026 · 9 min · 1827 words · martinuke0
Abstract diagram of an agent loop with tool calls, scratchpad, and termination gates.

Claude Loop Engineering: Designing Agentic Workflows That Actually Converge

Claude loop engineering is the practice of building bounded, observable, tool-using agentic workflows on top of Anthropic’s Claude. This post covers loop anatomy, stop conditions, tool design, and the production patterns that keep agents from drifting, looping forever, or burning budget.

September 5, 2026 · 11 min · 2230 words · martinuke0
An abstract diagram of tokens flowing into a model window.

Context Engineering: The Discipline Behind Reliable LLM Applications

Prompt engineering is shrinking inside the model. Context engineering — choosing what goes into the window — is now the lever for quality, cost, and reliability.

September 5, 2026 · 10 min · 2081 words · martinuke0
Terminal screenshot of a fine-tuning training run showing loss curves.

Fine-Tuning a Small LLM: A Hands-On Experiment

A working engineer’s journal of fine-tuning a small open-source LLM with LoRA on a custom dataset — the full pipeline, the surprises, and the measurable results.

September 5, 2026 · 9 min · 1816 words · martinuke0
Diagram of a RAG ingestion pipeline showing document parsing, chunking, embedding, and indexing stages.

The Hardest RAG Ingestion Problems and How Engineers Actually Solve Them

Most RAG failures start long before the LLM is ever called. This post walks through the five hardest ingestion problems and the patterns teams use to fix them in production.

September 5, 2026 · 10 min · 2107 words · martinuke0
Feedback