Baseten
Episodes
-
#5830: Serving Your Own Fine-Tuned Model in the CloudYou fine-tuned an open-weight model. Now how does anyone actually talk to it? Dedicated GPUs vs serverless inference, and the math that decides it. -
#5465: Packaging AI Pipelines So They Actually Get ReusedDaniel's podcast pipeline works — so why can't he reuse it? Recipes, containers, and Hebrew code-switching TTS.