Filter All Philosophy Fringe Tech Rant Essay Things I Love
Laguna-XS runs beautifully on my 5090, and I am keeping the model everyone calls outdated
A newer, well-recommended coding model fits my 32GB GPU at max quality and runs faster than the one I keep. I kept the old one anyway, and the numbers say I was right to.
7 min read
Qwen3.6-27B Scores Higher on BFCL, Ties on My Workload, and Runs 4.9x Slower
Qwen3.6-27B scored higher on BFCL. Both models scored 58/59 on the real workload. One answers in 0.87s, the other in 4.27s.
10 min read
GLM-4.7-Flash on One Consumer GPU: Why It Runs My Homelab Agent
A ~31B Mixture-of-Experts model at 4-bit, on one 32GB card, that beat a dense 32B on the real BFCL function-calling benchmark. What it is, what it is good for, and the numbers.
8 min read
LLMs from the Ground Up: From One Number to a Fine-Tuned Model
Every core LLM idea, defined in order and built on the one before it: weights, tensors, attention, quantization, LoRA.
17 min read
A Raspberry Pi Agent You Text Over Signal: The Complete Build
signal-cli from scratch on an ARM64 Pi, an agent runtime built from source for the Signal channel, and a read-only security model enforced in layers. Everything needed to rebuild it, including every trap.
10 min read
The AIPI Lite, Rebuilt as a Fully Local Voice Assistant
A cheap cloud-locked AI gadget, reflashed to answer from your own LLM. No cloud, no subscription.
12 min read
The whole library, offline, on one small computer
One small computer, a few watts, and a 290-source offline library that keeps working when the internet is off, and comes with you when you travel.
15 min read
A Private Coding LLM I Rent Flat-Rate Through Proton Lumo.
I point my coding agent at Proton Lumo for a private model with no per-token bill, riding a subscription I already pay for. The provider block, the flat-rate plans, a forty-line proxy that proves which model and tier Proton served, and an honest read on where it works.
8 min read
Build Your Own Weather Service: A Self-Hosted App That Stays Up When Its Upstreams Don't
Forecasts, radar, severe alerts, and a burn-window verdict from free public data, cached locally and reachable from anywhere, on hardware you already own.
15 min read
A Private Coding LLM on My Own Network.
I run Qwen3-Coder on a spare RTX 5090 and point my coding agent at it over the LAN. Private, fast, no per-token bill, secured with an API key, and watched on a dashboard. Here is the whole build.
14 min read
A Cistern Dashboard That Warns Me Before It Runs Dry
I run my place off an underground cistern, and a dry run can boil the pump and flood the basement. Here is the whole chain that put the tank on a dashboard, from the 4-20mA probe to the gallons reading on my phone.
17 min read
My Home Dashboards, From Anywhere. No Ports Opened.
Reaching a LAN-only Grafana dashboard from anywhere, behind a login, without opening a single inbound port at home. A self-hosted WireGuard tunnel with Pangolin.
7 min read
Stop Buying Sensors. Read the Ones You Already Paid For.
My two gas heat pumps were measuring a hundred things and showing me four. Here is how I pulled the rest onto my own dashboards, and why it now warns me before the heat goes out.
9 min read
You Bought the Mac. The Org Still Owns the Serial.
You bought a Mac and cannot get past its Remote Management lock. A bypass clears it in minutes, then one reset brings it back. The lock was never really on the Mac.
3 min read
Every Device I Own, on One Dashboard.
Most of my gear ships with its own dashboard, so over a few evenings I wired the whole network into one Prometheus and Grafana stack. The fun part was the two 60GHz radios with no API, where the exporter just logs in the same way the web page does.
7 min read
WiFi Sensing Is Real. The Room-Scale Version Is Marketing.
I built a nine-node Channel State Information rig to watch a room without cameras. It taught me exactly where the hype stops and the physics starts.
7 min read
The Model Isn't Thinking. Neither Were You, Most of the Time.
Stochastic parrot or emerging mind, both cliches are wrong in opposite directions. What language models actually do is internalize the shape of every structured thing we write, contracts, proofs, code, and the unsettling part is how much of our own work runs on the same map.
6 min read