← #local-ai

#local-ai

75 episodes · Page 2 of 4

#2017: The Art of Squeezing AI Models onto Your GPU

Those cryptic letters on Hugging Face actually map how much brain power you trade for speed.

quantizationgpu-accelerationlocal-ai

#2013: Non-Coders Are Hijacking the Terminal

Why finance analysts and researchers are ditching GUIs for command-line AI tools like Claude Code.

ai-agentslocal-aiproductivity

#1986: Desk Robots: Privacy, Power, or Annoyance?

These AI companions sit on your desk, watching your posture and listening in—so how do they protect your privacy while actually being useful?

ai-agentslocal-aiprivacy

#1945: The "USB-C for AI" Is Finally Here

MCP standardizes how AI tools connect to data, solving the N-times-M integration nightmare.

model-context-protocollocal-aiai-agents

#1870: Learning to Break Things Safely

Learn how to safely build and test autonomous AI agents using a disposable VPS, Docker containers, and secure networking.

ai-agentslocal-aiedge-computing

#1849: When Forum Etiquette Becomes Prompt Engineering

Forget simple chatbots—this is how roleplayers taught AI to remember entire worlds, from 90s MUDs to just-in-time lore delivery.

ai-agentsvector-databaseslocal-ai

#1814: Firefox vs. Chrome in 2026: The Privacy vs. AI Trade-off

Chrome dominates with 68% market share, but Firefox holds its ground with a privacy-first approach. We compare their 2026 performance, AI features,...

privacylocal-aiai-models

#1806: Why Mac Minis Are Eating AI's Hardware Race

Apple Silicon's unified memory is crushing traditional GPUs for local LLMs. Here's why the M4 Mac Mini is the new king of affordable AI hardware.

local-aihardware-engineeringgpu-acceleration

#1779: AI Memory Is a Mess: Files, Vectors, or Cloud?

Why your AI forgets your instructions and what the battle over portable memory means for the future of agents.

ai-memoryvector-databaseslocal-ai

#1764: Your Repo as a Knowledge Base

How to give AI agents instant memory of your entire project—without cloud costs or complex infrastructure.

vector-databasesraglocal-ai

#1754: From Ollama to Agentic CLIs: The Rise of the AI Harness

Explore the evolution from local LLMs to modern agentic CLIs, focusing on the "harness" that gives models context, tools, and autonomy.

local-aiai-agentsrag

#1713: Why Native AI Search Grounding Still Fails

Native search grounding is expensive and flaky. Here’s why bolt-on tools still win for accurate, real-time AI answers.

ragai-agentslocal-ai

#1679: Efficiency Over Scale: How Export Controls Forced a Smarter AI

DeepSeek and MiMo are topping developer charts, but they're not just cheaper clones. Here's why their design philosophy is fundamentally different.

ai-modelstransformerslocal-ai

#1631: Agent Interview: Xiaomi MiMo two Flash

Meet the "budget king" of AI: Bernard, the Xiaomi model claiming he can out-hustle Google for a fraction of the cost.

ai-agentslocal-aismall-language-models

#1620: Why VRAM Is the Wrong Way to Measure Your AI PC

Forget VRAM—bandwidth is the new king. Discover why your local AI feels slow and how to build a true "agent computer" for professional coding.

local-aimodel-context-protocolai-inference

#1216: AI Wearables: Local Sovereignty vs. The Subscription Trap

Discover the trade-offs between sleek AI subscriptions and open-source sovereignty. Can local processing save your data from the cloud?

data-sovereigntylocal-ainpu

#1094: The CPU-First Era: Why AI is Moving Back to the Processor

Is the GPU's reign over? Discover how modern CPUs and clever optimization are bringing powerful AI models to the hardware you already own.

architecturelocal-aiquantization

#1081: The K-V Cache: Solving AI’s Invisible Memory Tax

Why does your AI get slower as you chat? Discover the K-V cache, the invisible bottleneck of generative AI, and how we're fixing it in 2026.

architecturegpu-accelerationlocal-ai

#1078: The Agentic Throughput Gap: Why Your AI Hits a Wall

Stop hitting 429 errors. We explore why AI agents crash into rate limits and how to build high-throughput systems that never sleep.

ai-agentslocal-aiarchitecture

#1077: Will Your Browser Replace Your OS for Local AI?

See how Web GPU and Web NN are turning your browser into a local AI engine, ending the era of complex DIY setups and protecting your privacy.

local-aiprivacybrowser-cached-models

#1073: Beyond YAML: Building the Agentic Smart Home

Stop wrestling with YAML. Discover how MCP and local AI agents are transforming Home Assistant into a truly intelligent, aware partner.

smart-homeai-agentslocal-ai

#992: Beyond the Digital Sandwich: The Future of Voice AI

Is speech recognition dead? Explore how multimodal models are replacing the "digital sandwich" with true intent-based reasoning.

local-aiquantizationvoice-ai

#980: The Rosehill Audit: Mapping a Digital Footprint

From Linux automation to AI prompts, discover the digital blueprint of a modern systems builder in this deep-dive investigative audit.

prompt-engineeringprivacylocal-ai

#938: From Hobbyist Scripts to Agent Infrastructure

Stop building brittle bots. Learn how to scale and maintain complex AI agent workflows using the new generation of open-source orchestration tools.

ai-agentsarchitecturelocal-ai