BERT
deep learning artificial neural network language model
Episodes
-
#5438: Why Linguists Left the AI RoomLarge language models grew out of linguistics — so why aren't linguists in the room where they're built?Main topic -
#5431: The Other Half of Hugging Face: Why BERT Still Out-Downloads LlamaEncoder models pull over a billion downloads a month. Decoder models pull 397 million. The AI conversation and the download counter are describing ...Main topic -
#5624: One Forward Pass: Building a Bounded ClassifierNo generation, no parsing — just label scores in a single pass. How to build a small encoder classifier that actually works. -
#5621: Constrained Decoding for AI-Generated Podcast TagsA working AI podcast pipeline has one quiet failure: the tagging. Here's how constrained decoding and a stable taxonomy fix it. -
#5430: Two Boxes: ASR and the Text Fixer Behind ItPunctuation, casing, ITN, disfluency — the four-job layer between raw ASR output and text you can actually read. -
#5697: Decision Models That Return Probabilities, Not TextA new class of model skips text generation entirely and returns calibrated probabilities in one forward pass. Here's what that changes. -
#5466: Hebrew Words Hidden in English TextDaniel wants a classifier that spots Hebrew written in Latin letters — and it turns out nobody's built one. -
#5461: TTS Can't Pronounce Hebrew Inside EnglishYour TTS reads Hebrew words with English phonetics. Here's why — and why the obvious fix doesn't work yet. -
#5396: Teaching a Small Model to Stop Spelling Out NumbersYour ASR pipeline is fine until someone dictates "three point two" and gets "three point two" spelled out. Here's how inverse text normalization ac... -
#5188: DeepSeek's Point Release That Isn'tDeepSeek shipped a whole new architecture and called it a point release. Here's what actually changed inside the model.