Ollama
software for running large language models (LLMs) on a local computer instead of in cloud servers
Episodes
-
#5436: Small Models as Rewriters, Not WritersWhy "don't say X" prompts backfire, and how a tiny grammar-constrained model can scrub a script without breaking its grammar. -
#5830: Serving Your Own Fine-Tuned Model in the CloudYou fine-tuned an open-weight model. Now how does anyone actually talk to it? Dedicated GPUs vs serverless inference, and the math that decides it. -
#5465: Packaging AI Pipelines So They Actually Get ReusedDaniel's podcast pipeline works — so why can't he reuse it? Recipes, containers, and Hebrew code-switching TTS. -
#5405: Omarchy: The Linux Distro Built for AI AgentsOmarchy treats AI agents as users of the OS itself — every setting a command, every config a text file. Here's how it works and where it breaks. -
#5393: Fine-Tuning a Model on 100 Hand-Edited AnswersYou don't need 10,000 examples to make a model sound like you. The real number is closer to 100 — if the edits are opinionated.