Llama 3.1 8B
Episodes
-
#5774: Reading the Model's Mind Mid-InferenceWhat if you could watch a model decide? Inside the tools that open up inference mid-computation — and the limits they hit. -
#5408: When Small NLP Models Beat the LLMFeature extraction, fill-mask, token classification — the classic NLP tasks still have a job. Here's when a small model beats a frontier API. -
#5393: Fine-Tuning a Model on 100 Hand-Edited AnswersYou don't need 10,000 examples to make a model sound like you. The real number is closer to 100 — if the edits are opinionated. -
#5770: Talking to Your Data: What MCP Doesn't SolveMCP standardized the pipe. It has no opinion about what flows through it — and that's where the hard part lives.