Short description of the cover image subject.

Hands‑On Build Guide: LoRA‑Based LLM Fine‑Tuning for Your Portfolio CV

A hands‑on guide to building a LoRA‑based LLM fine‑tuning system you can ship to GitHub and discuss in interviews.

September 11, 2026 · 7 min · 1471 words · martinuke0
A diagram of a neural network with attention branches

Build a Mini LLM Inference Engine with Tree-Attention Speculative Decoding

Learn to implement a compact LLM inference engine that showcases advanced systems skills, from KV cache reuse to tree-attention speculative decoding.

September 10, 2026 · 8 min · 1577 words · martinuke0
A conceptual illustration of a token sampling distribution

Building a Production-Grade Top‑k/Top‑p Sampler with NumPy

This tutorial shows how to build a high‑performance top‑k/top‑p token sampler with temperature scaling and per‑token log‑probability tracking using NumPy. It demonstrates real systems skills relevant to ML inference roles.

September 9, 2026 · 2 min · 216 words · martinuke0
Python code on a laptop screen with a diagram of a circular KV cache

Pure‑Python LLM Inference Engine with Entropy‑Guided Sliding‑Window KV Cache

A practical, runnable pure‑Python LLM inference engine that uses an entropy‑guided sliding‑window KV cache backed by hash‑indexed circular pages – perfect for showcasing systems engineering chops.

September 8, 2026 · 11 min · 2237 words · martinuke0
A sleek laptop screen displaying code, a terminal, and a small generated poem.

From-Scratch Transformer Language Model: A Hands‑On Portfolio Project

A hands‑on guide to building a minimal GPT‑style model from scratch, with runnable code, testing, and extension ideas for a standout portfolio project.

September 8, 2026 · 9 min · 1885 words · martinuke0
Feedback