# This solo builder runs 24/7 local AI on his own hardware | Alex Finn Page: https://stenobird.com/podcast/how-i-ai-7304222/this-solo-builder-runs-24-7-local-ai-on-his-own-hardware-alex-finn Text version: https://stenobird.com/podcast/how-i-ai-7304222/this-solo-builder-runs-24-7-local-ai-on-his-own-hardware-alex-finn.md Podcast: [How I AI](https://stenobird.com/podcast/how-i-ai-7304222) Published: 2026-07-13T11:00:00+00:00 Episode link: https://podcasters.spotify.com/pod/show/pen-name/episodes/This-solo-builder-runs-247-local-AI-on-his-own-hardware--Alex-Finn-e3lo46b Audio file: https://anchor.fm/s/1035b1568/podcast/play/122474123/https%3A%2F%2Fd3ctxlq1ktw2nl.cloudfront.net%2Fstaging%2F2026-6-6%2F427473939-44100-2-a9f059103163d.mp3 Processing state: not_requested JSON: https://stenobird.com/v1/public/podcasts/how-i-ai-7304222/episodes/this-solo-builder-runs-24-7-local-ai-on-his-own-hardware-alex-finn Duration seconds: 2150 ## Resource Alex Finn is an AI builder, YouTuber, and the creator of Vibe Code Academy, a community for people learning to build with AI tools. He runs one of the most ambitious local AI setups I’ve come across: three Mac Studio 512 GB machines, a DGX Spark, and a custom RTX 5090 build, all coordinated through a fleet dashboard he built himself. He’s spent five months figuring out which local models belong on which machines, how to wire them to Claude Code loops, and how to get a software factory running without babysitting it. What you’ll learn: How Alex chose between a Mac Studio (512 GB unified memory), DGX Spark, and RTX 5090, and what each is actually good for Why Tailscale is worth installing even on a single machine, and how it lets one agent manage your entire hardware fleet How the build loop and review loop in Claude Code work How to allocate tasks by machine and model Why unlimited local inference changes the use-case math in a way a $20 cloud subscription never can What OpenClaw and Hermes are each best suited for, and why Alex runs five agents total with failover baked in — Brought to you by: Runway —The creative AI platform for images, video, and more Jira Product Discovery —Prioritize with insights, build with confidence — In this episode, we cover: (00:00) Intro (02:58) Alex's hardware stack (03:48) What "ambient AI" means (04:15) Alex's red-pill moment with OpenClaw (07:04) Mac Studio vs. DGX Spark vs. RTX 5090 (13:24) How to set up local models with no technical knowledge (Tailscale + OpenClaw/Hermes) (17:16) Fleet control dashboard: assigning 24/7 tasks across machines (20:42) Local models as security scanners feeding Claude Code (22:25) How Alex allocates GLM 5.2, Qwen 3.6, and Ornith 1.0 by task (24:28) OpenClaw vs. Hermes: the honest comparison (26:… ## Actions - request_transcript: `POST https://stenobird.com/v1/public/podcasts/how-i-ai-7304222/episodes/this-solo-builder-runs-24-7-local-ai-on-his-own-hardware-alex-finn/transcription-requests` — Idempotently request low-priority transcript generation for this episode. - read_markdown: `GET https://stenobird.com/podcast/how-i-ai-7304222/this-solo-builder-runs-24-7-local-ai-on-his-own-hardware-alex-finn.md` — Read the agent-friendly Markdown representation of this episode resource. A page view does not enqueue transcription. Agents should invoke `request_transcript` explicitly when they need this episode processed. ## Transcript Full transcripts are not published on public pages unless there is a clear rights basis.