the localhost

the localhost

Five tech industry veterans argue about AI that runs on your own hardware. Open weight models, NPUs, home labs, and the real question underneath it all: who controls intelligence, you or a data center? Every week, one host takes the hot seat to defend a provocative claim while the other four try to take it apart. Then everyone votes. We run the models, test the vendor claims ourselves, and share what actually happened, including the failures. No hype, no scripts, and nobody here is pretending to have all the answers. We're learning this stuff alongside you, one episode at a time. New episodes weekly. There's no place like 127.0.0.1.

Episodes

  1. 1d ago

    the localhost:0003 | the ai hardware boom is beginning

    Last week we argued the $25 token was dying. This week: that's exactly when the hardware boom starts. Frank takes the hot seat with a paradox. If intelligence keeps getting cheaper, we don't use less of it — we invent more things to do with it. More agents, more inference, more workloads. Which leaves every enterprise with one question they haven't had to ask before: which intelligence should we rent, and which should we own? Citi and Vercel are reportedly running open-weight models more than half the time. In June it was 29%. By late August, 53%. A new family of models out of the UAE ships six versions built for six classes of hardware — phone to server — with no quantization involved. NVIDIA's Hugging Face acquisition is official. And OpenAI shipped Astra. We also had four people instead of five, a guest from London, and for the first time in three episodes, a completely unanimous vote. ━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━ CHAPTERS 00:00 If intelligence gets cheap, what do you buy?00:47 Welcome back — who's here, who's at Disneyland01:28 Laura Osborn joins from London02:14 Neil, and the Bears/Seahawks problem03:07 Chauncey, and the disclosure03:43 The Rundown: NVIDIA and Hugging Face is official04:25 Neil: what GitHub can tell us about this05:48 Citi and Vercel pass 50% open weights07:50 The three-tier model: endpoint, edge, cloud08:56 Six models for six classes of hardware10:22 Purpose-built beats shrinking things down?11:21 What UK and European customers actually ask for14:41 From cloud architect to local AI15:18 OpenAI ships Astra16:26 Chauncey: be careful what you point it at17:51 Neil: code generation at the edge19:33 Baseline: what should people actually care about?20:20 Open weights are no longer philosophical21:27 What's coming out of IFA22:52 The medical school of the future25:00 Why doesn't anyone know this is possible yet?26:53 "We have more agents than human employees"28:23 Tension of the Week: the hardware boom begins30:27 Chauncey: fixed cost beats a variable one32:39 Laura: I spent five years telling everyone to go cloud34:18 Neil: when pennies turn into a million dollars35:00 The brewery test36:06 The data center next door36:34 The vote38:42 On My Device: the 36B gets the job40:07 Chauncey: Copilot CLI, running locally41:44 The Dream Team and the one-person virtual business41:56 Laura: an on-device speaker coach42:58 Neil: Mac vs the Beast, round two ━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━ WHAT CAME UP • NVIDIA acquiring Hugging Face is confirmed. NVIDIA says it stays open and that NVIDIA hardware won't be required. Neil's read: this looks like Microsoft buying GitHub — the same fears, and possibly the same outcome. His angle is that Hugging Face has been effectively off-limits to a lot of enterprises, and being inside a vendor with a security and compliance apparatus may be what finally gets those models sanctioned. • Citi and Vercel are reported to be running open-weight models for more than half their AI workloads — 29% in June, 53% by late August. Ten weeks. • The MBZUAI Institute of Foundational Models released six models spanning roughly 1B to 375B parameters, each trained for a specific class of hardware rather than quantized down from something larger. Frank ran the 36B on his dual-3090 workstation against Qwen3.8-27B: about 68 tokens/sec against 40, on a comparable qualifying profile. • Neil's three tiers: the endpoint, the cloud, and an emerging middle — departmental edge servers. Local doesn't grant you HIPAA or FERPA compliance by itself, but no round trip to the cloud is a smaller attack surface. • The economics nobody models: an insurance agent filing 50 claims a day, times 10,000 agents. Even at pennies per call, that's over a million dollars a year in workloads that could run at the edge. • Laura, five years a cloud architect, on doing a full 180: the constraint in Europe isn't just sovereignty rules, it's that cloud capacity is at its limits. • And Neil's field test of public sentiment, conducted on a bartender who asked him point blank whether he was anti-AI. The vote was 4–0. First unanimous verdict of the series. ━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━ THIS WEEK'S CREW Frank Buchholz — host, independentLaura Osborn — guest, joining from LondonNeil MisakChauncey Larsen Jacob Rhoades and Robert Henry are off this week. Jacob is celebrating his first anniversary. Robert was on a plane, sending Teams messages anyway. DISCLOSURE: Laura, Neil and Chauncey are employed by Microsoft. Frank is independent. All views expressed are their own, nothing discussed is unannounced or non-public, and none of this is financial advice. ━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━ localhost is a weekly show about running AI on hardware you own. Independent and unsponsored. New episodes weekly: https://thelocalhost.show There's no place like 127.0.0.1. #LocalAI #OpenWeights #AI #EdgeAI #AIHardware #Sovereignty #OnDeviceAI #AIPodcast

    the localhost:0003 | the ai hardware boom is beginning
  2. Aug 22

    the localhost:0001 | open is a position

    Five tech industry veterans, one hot seat, and a vote. In our first episode: Meta gave away a new AI model and promised the big one soon, NVIDIA's CEO used his first ever post on X to defend open weights, and 270 companies signed on. One did not. So Frank takes the hot seat to argue that openness is positional, not principled, and the crew tries to take him apart. Along the way: what open weight actually means (up to the secret sauce), the Flock camera controversy as edge AI's cautionary tale, and a job interview for five AI models on a used $700 graphics card, with a winner nobody was talking about. Plus the stat of the night: one host's router setup runs 86 percent of his AI locally and has already saved him $1,600. We run the models, test the claims, and share what actually happened, including our own mistakes. New episodes weekly. There's no place like 127.0.0.1. Creators & Guests Chauncey Larsen - Host Frank Buchholz - Host Jacob Rhoades - Host Neil Misak - Host Robert Henry - Host (00:00) - cold open: from cost to control (01:05) - meet the hosts (and who writes the paychecks) (04:50) - the rundown: qwen 3.8 lands (07:55) - open weight vs open source, explained (14:45) - flock cameras: edge capture, central control (17:40) - this week's tension: open is a position (21:05) - the debate: manifestos, chess games, and IPOs (29:15) - the vote (31:40) - on my device: five models walk into a job interview (34:45) - what's running on the crew's devices (37:45) - the 86 percent stat (40:20) - wrap: there's no place like 127.0.0.1 Mark Zuckerberg's manifesto, "The Future Is for Everyone": https://www.meta.com/thefutureisforeveryone/ The Open Weights letter coverage and Spark 1.2 status: https://www.orcarouter.ai/blog/meta-muse-spark-1-2-explained Glimmer launch coverage: https://www.cnbc.com/2026/08/10/meta-muse-glimmer-open-weight-ai.html The May "not suitable for open sourcing" quote: https://www.implicator.ai/meta-releases-30b-open-weight-muse-glimmer-and-promises-spark-1-2-weights/ Flock's announced changes: https://www.nbcnews.com/tech/security/flock-safety-police-abuse-oversight-data-retention-rcna592217 Qwen 3.8-27B: https://huggingface.co/Qwen/Qwen3.8-27B The 2.4T open release: https://huggingface.co/Qwen/Qwen3.8-2.4T-A95B Frank's full benchmark, raw traces and all: https://github.com/frankcx1/bakeoff The results post: https://www.linkedin.com/posts/frankjbuchholz_on-monday-meta-gave-away-a-brand-new-ai-share-7493881833236123648-1kd1/ the localhost is five friends talking about AI on your own hardware: Frank Buchholz (independent), Jacob Rhoades, Robert Henry, Neil Misak, and Chauncey Larsen (Microsoft, opinions their own). The show is independent and unsponsored. Find us at https://thelocalhost.show

    the localhost:0001 | open is a position

About

Five tech industry veterans argue about AI that runs on your own hardware. Open weight models, NPUs, home labs, and the real question underneath it all: who controls intelligence, you or a data center? Every week, one host takes the hot seat to defend a provocative claim while the other four try to take it apart. Then everyone votes. We run the models, test the vendor claims ourselves, and share what actually happened, including the failures. No hype, no scripts, and nobody here is pretending to have all the answers. We're learning this stuff alongside you, one episode at a time. New episodes weekly. There's no place like 127.0.0.1.