Tech Blog
  • Home
  • Posts
  • About
Categories
All (5)
agents (2)
agentx (1)
architectures (1)
benchmarks (1)
docker (1)
dotnet (1)
evaluation (1)
gemma (1)
hardware (1)
homelab (1)
hugo (1)
llama (1)
local-ai (3)
mcp (1)
nginx (1)
pcie (1)
quantization (1)
qwen (1)
research (1)

Posts

Local AI Model Test Bank

local-ai
evaluation
agents
benchmarks

Scoring a local model is easy. Scoring it well enough to rank one above another is a different problem — and my previous instrument could not resolve its own top two. Here is what I rebuilt and why.

13 Aug 2026
15 min

My Hardware

local-ai
hardware
pcie
homelab

Two identical 16 GB GPUs, and one of them is worth a quarter of the other. The parts list is the boring half of this post — the interesting half is what the PCIe topology of a consumer board does to a dual-GPU inference box.

13 Aug 2026
9 min

Three Architectures for Local Inference

local-ai
architectures
qwen
llama
gemma
quantization

Qwen, Llama and Gemma solve the same problem three different ways — MoE, a mature baseline, and sliding-window attention. A walk through what each one actually changes, and what fits in the VRAM you have.

13 Aug 2026
10 min

What AgentX Is For

agentx
agents
research
mcp
dotnet

It started as a merge of three tools I was tired of maintaining separately. It turned into a research assistant, and the reason is that a hypothesis and a requirement are the same shape — something stated, refined, and eventually resolved.

13 Aug 2026
9 min

Deploying Hugo with Docker on a VPS

hugo
docker
nginx

How to build a Hugo site into a multi-stage image and serve it with Nginx, without shipping the compiler inside the final image.

31 Jul 2026
1 min
No matching items
 

© 2026 Javier Iracheta. Built with Quarto.