Gemma
family of large language models by Google
Episodes
-
#5442: When Do LLMs Suddenly Get Good at Things?New capabilities seem to appear out of nowhere as models scale. Whether that's real — or just how we measure — is still an open fight. -
#5697: Decision Models That Return Probabilities, Not TextA new class of model skips text generation entirely and returns calibrated probabilities in one forward pass. Here's what that changes. -
#5431: The Other Half of Hugging Face: Why BERT Still Out-Downloads LlamaEncoder models pull over a billion downloads a month. Decoder models pull 397 million. The AI conversation and the download counter are describing ... -
#5393: Fine-Tuning a Model on 100 Hand-Edited AnswersYou don't need 10,000 examples to make a model sound like you. The real number is closer to 100 — if the edits are opinionated.