GPT-4o
family of multimodal artificial intelligence models developed by OpenAI
Episodes
-
#5774: Reading the Model's Mind Mid-InferenceWhat if you could watch a model decide? Inside the tools that open up inference mid-computation — and the limits they hit. -
#5507: Inside ChatGPT's Hidden Pipeline: 8 Models Per ReplyOne chat turn isn't one model call. It's moderation, PII filtering, a tone-based router, and activation classifiers — here's the documented graph. -
#5503: Inside the Hidden Image Generation PipelineThat one-click image generator is secretly a graph of many models. We reconstruct the hidden pipeline behind Gemini and ChatGPT. -
#5435: When Your TTS Model Eats the NumbersNumbers, dates, and acronyms break text-to-speech in specific, documented ways. Here's where normalization lives — and why it depends on your model. -
#5412: Editing vs. Note-Taking for AI Fine-TunesHand-editing a model's output gives three training signals at once. Writing notes gives one — and a weaker one at that. -
#5169: Fine-Tuning vs From-Scratch for Minor LanguagesWhat 774 experiments and a new Armenian model reveal about the tradeoff between fluency and knowledge in low-resource languages. -
#5771: MCP Connectors: Talk to Your Data vs. Open DataSame protocol, two patterns: pointing agents at your own database, and pointing them at public data — with very different results. -
#5770: Talking to Your Data: What MCP Doesn't SolveMCP standardized the pipe. It has no opinion about what flows through it — and that's where the hard part lives. -
#5407: Hemmingway-1 and the War on WaffleA 27B model promises answers without the preamble. Its benchmark is homegrown — and the behavior it targets has a paper trail.