Independent publication // local-first AI // field reports from owned systems
Nodehome
A running publication about local models, private inference, self-hosted agents, weird hardware, research sweeps, and the builders wiring their own AI stack together.
AI is getting physical again. It shows up in terminals, racks, side projects, and ugly little workflows people actually control.
Latest
Daily Sweep - Sep 23, 2026Someone bought a Jensen-signed RTX 5090 second-hand, now it may sell for $15,000
GPU hardware mention - directly relevant to the build or resale market.
WordPress: Unauthenticated path traversal leading to conditional RCE
Show HN: Training a model to identify AI web content from structure alone
Fine-tuning or training technique - relevant if local training becomes part of the workflow.
**Know Who Spoke When: Build Real-Time, Multi-Speaker AI with NVIDIA Nemotron 3 Diarization**
Hugging Face blog post - check for new model releases, library updates, or ecosystem shifts.
Field Reports
builds, experiments, notesThree RTX 3090s, One 32B Model: A Pipeline-Parallel Canary
A current field note on why the 3x3090 serving path moved through pipeline parallelism, not tensor parallelism, for the tested 32B AWQ model.
Gemma 4 12B And The Sensory Agent Lane
A public-safe read on Gemma 4 12B as a local sensory preprocessor: useful for seeing, hearing, and structuring observations without turning into an action system.
Hardware
machines, thermals, economicsPower Caps On Three RTX 3090s: Bursts Versus Sustained Load
A measured note on 300W bursty inference, lower caps for sustained runs, and why power-cap sweet spots are workload-specific.
Parallel Agent Serving Is A Hardware Shape Now
A field-report read on 14x RTX 3090 agent serving, EXL3, FP8 KV cache, Aphrodite, and why concurrency is becoming the local hardware metric.