GGUF
file format family for storing large language models
Episodes
-
#5485: Fine-Tuning Parakeet for Hebrew and Your Own JargonNVIDIA's Parakeet beats Whisper on Android — but can you teach it Hebrew, or just your own jargon? Two answers, one much happier. -
#5410: Adapters: 102KB That Reshapes a 403GB ModelA 102KB adapter file changes how a 403GB base model behaves — without ever merging into it. Here's how model adapters actually work. -
#5384: Android ASR Runtimes: LiteRT, ExecuTorch, and Why Your Phone Has No VRAMWhy does your phone have no VRAM number? A tour of Android's runtime layer and what it takes to run ASR locally. -
#5416: Xet, Buckets, and Auto-Pulling Model WeightsHugging Face swapped its storage backend to Xet with barely a ripple. So why can't a bucket auto-pull new upstream weights?