Breakeven Brothers

Bradley Bernard, Bennett Bernard

Breakeven Brothers is a practical AI podcast where brothers Bradley and Bennett Bernard build software, compare notes, and tell you what actually worked. Bradley is a software engineer at OpenAI and longtime product builder behind SplitMyExpenses. Bennett is a CPA and hands-on software builder applying AI to accounting and real business workflows. Each episode goes beyond release notes and benchmark hype. They test coding agents, new models, developer workflows, and automation in real projects—then talk honestly about what works, what breaks, what costs too much, and what is worth shipping. If you're building with AI—or trying to figure out where it actually helps—join them for practical experiments, honest tradeoffs, and enough sibling banter to keep it human.

  1. Sep 16

    Nvidia, Hugging Face, And The AI Hardware Shift

    Bradley and Bennett unpack NVIDIA’s $12 billion Hugging Face acquisition and consider what control over a major open-model platform could add to NVIDIA’s hardware strategy. From the programmable MicroDuck and Anthropic’s Model Hardware Standard to Figure, 1X, local compute, and 3D-printing robots, they explore why AI may be moving from cheap software into useful physical machines—and why a small desktop duck feels easier to trust than a full-size humanoid. The conversation then shifts to OpenAI’s Agents API and GPT-Live API. They explain managed Codex harnesses, hosted sandboxes, tool and MCP connections, privacy concerns, and full-duplex voice, while Bennett shares his Tax Avatar client Q&A concept and Bradley sketches a voice-driven bill-splitting workflow. They also weigh “Sent with ChatGPT” messages and AI-written documentation, question when generated context becomes slop, and bookmark AI code-review guidance, Higgsfield’s ChatGPT video plugin, and model-discovery resources from Arena.ai and Hugging Face. Chapters (00:00) - Waffle Buns and Intros (04:22) - NVIDIA Acquires Hugging Face (07:25) - MicroDuck and Physical AI (13:03) - Humanoid Robots at Home (16:57) - Robot Companions and Risks (20:52) - OpenAI’s Agents API (25:20) - Agents API vs. SDK (28:36) - Why Sandboxes Matter (30:45) - Full-Duplex Voice AI (35:40) - Tax Avatar for Clients (38:15) - Voice Apps and Pricing (41:45) - AI-Written Slack Messages (45:52) - Detecting Foisted AI Slop (48:28) - Human Review and Sign-Off (52:45) - Outsourcing Critical Thinking (56:23) - AI Review Guardrails (1:00:10) - Higgsfield Video Editing (1:04:00) - Waymo Routes and Speeds (1:06:00) - Autonomous Safety and Abuse (1:09:04) - Coordinated Fleets and Speed #AIAgents #OpenAI #GPTLive #VoiceAI #AIHardware #AIRobotics #NVIDIA #HuggingFace #OpenSourceAI #AISecurity

  2. Sep 5

    Model Week: Astra, Grok Bot, Gemini, Anthropic, And Meta

    Bradley and Bennett break down a packed wave of AI releases: Google Gemini Flash 3.8 and Flash Cyber, Anthropic Fable and Mythos 5.1, Meta Muse Spark 1.3, and OpenAI GPT-6 Astra. Rather than treating benchmark charts as final verdicts, they compare intelligence, speed, token economics, usage limits, interface quality, and the expensive tradeoff between cloud services and private, open-weight models running on local hardware. Then they move from launch claims to real workflows. Bradley shares Astra tests involving code architecture and product analytics, plus an earlier computer-use workflow that created an eBay listing, while Bennett tests Astra on a browser-based co-op game and explores AI-assisted accounting, video editing, and publishing. They also examine Grok Bot’s chief-of-staff approach and the central tension of autonomous agents: the same proactive browser access that makes them powerful creates serious concerns about credentials, prompt injection, irreversible actions, and the need for read-only or tightly scoped permissions. Finally, they debate Valon’s limits on AI access for new employees, balancing token spend and faster onboarding against learning business context from coworkers. Their bookmarks cover six practical Astra use cases and Shopify’s work on tiny models trained for specialized tasks. Chapters: 00:00 Episode 44: Model Week 01:35 Google Gemini Flash 3.8 and Flash Cyber 05:11 Anthropic Fable and Mythos 5.1 08:40 Meta Muse Spark 1.3 and Local AI Hardware 12:57 GPT-6 Astra Launch and First Impressions 18:54 Building with Astra and Rethinking Work 24:06 Delegating Video Production to AI 29:04 Astra Safety, Ultra Mode, and Resets 30:02 Grok Bot and Always-On Agents 32:58 Grok Bot Security, Access, and Limits 42:11 AI Onboarding: Tokens, Context, and Coworkers 53:30 Six Use Cases for Astra 55:30 Shopify and Specialized Tiny Models Links: - Google Gemini Flash 3.8 & Flash Cyber: https://blog.google/innovation-and-ai/models-and-research/gemini-models/3-8-flash-and-3-8-flash-cyber/ - Anthropic Fable & Mythos 5.1: https://www.anthropic.com/claude-fable-and-mythos-5-1 - Meta Muse Spark 1.3: https://research.meta.ai/blog/introducing-muse-spark-1-3 - OpenAI GPT-6 Astra: https://openai.com/index/gpt-6-astra/ - Grok Bot: https://x.ai/bot - Valon turned off AI for new coworkers: https://superpowerdaily.com/posts/valon-makes-most-new-hires-earn-ai-access-projects-75-cut-in-token-spending - NVIDIA acquires Hugging Face for $12B: https://blogs.nvidia.com/blog/nvidia-to-acquire-hugging-face/ - 6 good use cases for GPT-6 Astra to see it's power: https://x.com/theo/status/2095966874010046621 - Arena.ai Leaderboard: https://arena.ai/leaderboard - Hugging Face trending models: https://huggingface.co/models #GPT6Astra #AIModelReleases #AIAgents #ComputerUse #GrokBot #GeminiAI #AnthropicAI #OpenSourceAI #AISecurity #AIProductivity

  3. Aug 18

    ChatGPT Finances, Cloudflare OS, And Work Mode

    Bradley and Bennett kick off a new season after a travel-and-life reset, with Bradley sharing his Copenhagen, Cannes, and Barcelona sprint before the show turns back to the AI firehose. The core conversation starts with ChatGPT Finances: Bradley walks through linking accounts with Plaid, seeing credit cards, bank transactions, investments, subscriptions, charts, and memory-backed financial conversations, then testing it against real decisions like canceling Whisperflow and reviewing whether his American Express annual fee still makes sense. Bennett brings the CPA lens, comparing the workflow to his Monarch Money setup, stressing the trade-off of putting more data with one vendor, and highlighting how AI can act less like a spreadsheet and more like a financial coach or first-pass advisor. The second half zooms out to the agentic web and AI-native work. The hosts unpack Cloudflare OS, Kite Surf, Slack-style agent workspaces, ChatGPT Work mode, the Codex/ChatGPT desktop merge, Remote mode on mobile, and browser-using agents that can price an eBay listing, create a QR contact card, audit Render hosting costs, move static pages to Cloudflare Pages, or monitor software upgrades. Throughout, Bradley and Bennett keep returning to the practical tension for builders and operators: AI agents are making valuable tasks feel low effort, but the harder questions are shifting toward trust, verification, security, product taste, and how much AI-generated code humans can realistically review. Chapters: 00:00 Season Two Returns 00:42 Bradley’s Europe Reset 02:27 Meet Bradley and Bennett 03:54 Trying ChatGPT Finances 06:52 AI Budgeting, Coaching, and Small Business Use Cases 14:52 Using AI to Cut Hosting Costs 17:14 Cloudflare OS and Agent Workspaces 21:14 Agents Browsing, Listing, and Automating Tasks 24:41 QR Codes, Kite Surf, and Agent-First Browsers 35:02 ChatGPT Work Mode, Codex Remote, and App Merges Links: - Codex Maxxing - Cloudflare OS - Kitesurf Browser - ChatGPT Finances - ChatGPT Work Mode - Show Me Code Changes - How Agents Are Transforming Work - LM Arena Leaderboard - Hugging Face Trending Models #BreakevenBrothers #ChatGPTFinances #PersonalFinanceAI #Codex #AIAgents #AgenticWeb #Cloudflare #FutureOfWork #VibeCoding #AIProductivity Creators & Guests Bennett Bernard - Host Bradley Bernard - Host

  4. May 21

    Codex maxxing and the AI-native workplace

    Brad’s North Carolina Cook Out mint Oreo shake review, and Ben’s Pacific Northwest work-trip update before the brothers dig into the big idea of the week: what it actually means to become “AI native” at work. Using Jason’s Codex maxxing article as the spark, Brad frames the shift from one-off ChatGPT-style prompts to long-running agent threads with memory, tools, voice input, steering, queued follow-ups, automations, and heartbeats. Ben translates that into a practical workplace lens with the “core four” framing—model, prompt, context, and tools—and compares onboarding an AI agent to training an intern in accounting or engineering. From there, the episode gets very hands-on. Brad explains how he thinks about memory files, Slack MCPs, voice workflows with Wispr Flow, recurring heartbeat tasks that check GitHub PRs, and Codex artifact-style outputs like PDFs, websites, and dashboards. Ben stress-tests the hype with very real questions about token budgets, enterprise ROI, Slack drafts that overcommit on your behalf, and the funny moral gray area of using powerful AI to do tiny jobs you could have done yourself in three clicks. The back half turns into a mobile-agent tour: Ben shares his Termius, VPS, SSH, and Tailscale experiment for running Codex or Claude Code from a phone, while Brad walks through Codex Remote in the ChatGPT app and why phone-to-computer agents feel like the future. They close with Brad’s must-watch mobile-app founder bookmark and a note that Episode 42 wraps season one before a short hiatus and refreshed look. Chapters: 00:00 Intro and Travel Catch-Up 02:06 AI-Native Codex Workflows 04:50 Core Four Agent Framework 07:26 Better Loops and Memory 11:07 Slack Drafts and Memory 14:23 Voice Input and Steering 19:24 Heartbeats for Recurring Work 23:52 Artifacts and Personalized Tools 26:56 Token Budgets and ROI 36:34 Bookmarks and Mobile Codex 44:04 Season One Hiatus Links: - Codex maxxing / AI-native workflow article - CEO’s journey of building a mobile app - Codex Remote in the ChatGPT app #BreakevenBrothers #OpenAI #Codex #CodexRemote #AINative #AIAgents #AICoding #ClaudeCode #MCPServers #AIProductivity Creators & Guests Bennett Bernard - Host Bradley Bernard - Host

  5. May 8

    GPT-5.5 made Codex our daily driver

    Episode 41 starts with a very Breakeven Brothers moment: the guys forget what episode they’re on, recap Brad’s family visit, and relive a 150-plus-game Soulcalibur gauntlet before diving into the real headline—why Brad now lives inside the Codex app. He explains what changed with GPT-5.5, why the model feels dramatically better at unblocking itself, and how that has shifted his workflow from small, careful requests to bigger, longer-running tasks. If you’ve been wondering whether Codex is just another AI coding surface or something materially different, this conversation gives a grounded answer from someone using it all day. From there, the episode gets practical fast. Brad walks through plugins, computer use, Git worktrees, the built-in review pane, automations, long chat compaction, and why plan mode has mostly disappeared from his workflow, while Ben stress-tests the ideas from an accounting and knowledge-work angle with Canva brochures, CRM data, email tooling, and plenty of honest beginner questions. They also compare Codex to Claude Code and Claude Co-Work, share Brad’s current default of GPT-5.5 on extra high reasoning, and close with two strong bookmarks: Greg Eisenberg and Riley Brown’s Codex masterclass, plus Evan Bacon’s serve-sim tool for showing an iOS simulator inside Codex’s right-hand pane. Chapters: 00:00 Weekend Recap and Soulcalibur 01:24 GPT-5.5 Powers Up Codex 05:16 Terminal Habits vs App Workflow 08:46 Plugins, MCPs, and Computer Use 11:22 Worktrees and Review Pane 15:05 Life After Plan Mode 20:37 Best Settings for Codex 25:48 Accounting Workflows and Plugins 30:14 Bookmarks and Closing Thoughts Links: - Codex app - Canva MCP - Startup Ideas Podcast — “How to Use Codex: The Codex Masterclass” - Augmented Accounting - serve-sim #BreakevenBrothers #Codex #GPT55 #AICoding #OpenAI #ClaudeCode #MCP #DeveloperTools #AIProductivity #AccountingTech Creators & Guests Bennett Bernard - Host Bradley Bernard - Host

  6. Apr 10

    Cursor 3, Gemma 4 & AI Taxes

    Episode 40 opens with Brad’s dirty-30 recap from Big Sur: hiking, Japanese hot baths, unlimited food, and a first falconry experience before the brothers jump into two meaningful AI releases. They break down Cursor 3’s move to a fully agent-first interface, why that changes the feel of an AI IDE, and how Google’s Gemma 4 models look in the small-model and on-device race. There’s also a practical discussion about Apple, local inference in iOS 26, and what better on-device models could unlock for products like SplitMyExpenses. Then the episode gets into one of the most interesting collisions in tech right now: AI tax prep. Using Daniel Vassallo’s thread as the spark, Ben brings the CPA perspective, Brad brings the builder perspective, and together they unpack customer experience, trust, knowledge work, and what AI still misses. The back half covers why MCP servers suddenly feel useful again, the Axios and LiteLLM supply-chain scares, practical OWASP-style security audits, Gary Tan’s gstack claims, a joking SplitMyExpenses sponsor read, and quick bookmarks on iOS reverse-engineering tools and Google’s latest quantum-and-crypto warning. Chapters: 00:00 Big Sur birthday and falconry 02:38 Cursor 3 goes agent-first 04:04 Gemma 4 on-device push 08:42 AI taxes and CPA backlash 18:09 G-Stack versus code quality 25:23 MCP servers make a comeback 33:10 Axios breach and AI audits 40:50 TBPN jokes and sponsor banter 44:04 Bookmarks: iPhone hacks and quantum Links: - SplitMyExpenses - Feross on the Axios attack - Daniel Vassallo AI tax-prep thread - gstack by Gary Tan - App Store Connect CLI - Hopper disassembler MCP #BreakevenBrothers #AICoding #CursorAI #Gemma4 #LocalAI #MCPServers #TaxTech #SupplyChainAttack #OWASP #DeveloperTools Creators & Guests Bennett Bernard - Host Bradley Bernard - Host

  7. Mar 12

    AI agents, accounting workflows & social contracts

    This week, the brothers dive headfirst into the world of personal AI agents, with Brad detailing his meticulous, multi-step journey of setting up OpenClaw on a virtual private server. The conversation navigates the critical balance between unlocking immense power and managing the significant security risks that come with granting AI access to personal data. From automating mundane tasks like unsubscribing from spam emails to the dream of booking dinner reservations with a simple command, they explore the technical hurdles and trust barriers that stand between us and a truly agentic future. The discussion then shifts to the professional world, tackling the recent panic and excitement in the accounting community as many discover the power of AI tools like Claude. Are accountants officially "cooked," or is this simply the next evolution of the profession? Drawing parallels to software engineering, they debate how AI is modifying jobs rather than eliminating them. Finally, they grapple with a fascinating new social concept: "AIDR" (AI Didn't Read), questioning the authenticity and perceived effort of AI-generated communication and what it means for our human-to-human connections in an increasingly automated world. Chapters: 00:09 Introduction and Personal Catch-Up 03:02 Darren Aronofsky's AI Film Series 06:36 Brad's Experience Setting Up OpenClaw 16:34 The Broader Challenges of AI Security 21:44 Is AI Making Accounting Obsolete? 29:57 Adapting to AI and GPT-4's Power 37:48 AIDR: AI's Impact on Communication Links: - AIDR: AI Didn't Read - Google Workspace CLI - Bradley's blog post #OpenInterpreter #AISecurity #FutureOfWork #AccountingTech #GPT4 #ClaudeAI #DeveloperTools #DigitalEthics #AIpodcast #PromptInjection Creators & Guests Bennett Bernard - Host Bradley Bernard - Host

  8. Feb 28

    AI for taxes & the end of App Stores?

    This week, the Breakeven Brothers dive into the practical (and sometimes frustrating) applications of AI, starting with a hilarious attempt to use AI for taxes and an experiment to recreate their famous intro jingle with Google's new Lyria 3 audio model. The results might surprise you. They also recap a flurry of major releases from the past few weeks, including OpenAI's Codex 5.3 Spark, Anthropic's Claude Sonnet 4.6 with its blazing fast mode, and Google's Gemini 3.1 Pro, debating the emerging market split between raw intelligence and sheer speed. The conversation then shifts to the powerful world of on-premise AI, exploring how tools like OpenCode and local LLMs can offer unparalleled security for sensitive data in industries like accounting and finance. Is it possible to get the power of modern AI without sending your data to the cloud? The hosts discuss the trade-offs and explore game-changing productivity hacks, like using the Codex app's automation features to analyze your own usage patterns. Finally, they tackle a thought-provoking tweet from Karpathy about the potential 'end of the App Store' and what the explosion of AI-generated software means for developers, legacy systems like COBOL, and companies like Apple. Chapters: 00:07 Can AI Do Your Taxes? 03:36 Recap of Recent AI Releases 05:56 Testing Google's Music AI, Lyria 09:01 Exploring Open Code with Local LLMs 14:31 The Case for On-Premise AI Models 24:55 Unlocking Power with Codex Automations 31:36 Will AI Make App Stores Obsolete? 38:43 Weekly Bookmarks and Security Concerns 43:11 Conclusion and Next Episode Teaser Links: - Replit Animation on X - CNET: Hackers Are Trying to Copy Gemini - Ollama #AIPodcast #LocalLLM #OpenCode #AppStore #SoftwareDevelopment #Codex #GoogleGemini #OnPrem #AICyberSecurity #AIMusic Creators & Guests Bennett Bernard - Host Bradley Bernard - Host

Ratings & Reviews

5
out of 5
2 Ratings

About

Breakeven Brothers is a practical AI podcast where brothers Bradley and Bennett Bernard build software, compare notes, and tell you what actually worked. Bradley is a software engineer at OpenAI and longtime product builder behind SplitMyExpenses. Bennett is a CPA and hands-on software builder applying AI to accounting and real business workflows. Each episode goes beyond release notes and benchmark hype. They test coding agents, new models, developer workflows, and automation in real projects—then talk honestly about what works, what breaks, what costs too much, and what is worth shipping. If you're building with AI—or trying to figure out where it actually helps—join them for practical experiments, honest tradeoffs, and enough sibling banter to keep it human.

You Might Also Like