DeepSeek V4
large language model
Episodes
-
#5154: DeepSeek's Two-Endpoint PhilosophyDeepSeek quietly routed Pro traffic to Flash — and that routing change says everything about its two-endpoint strategy. -
#5830: Serving Your Own Fine-Tuned Model in the CloudYou fine-tuned an open-weight model. Now how does anyone actually talk to it? Dedicated GPUs vs serverless inference, and the math that decides it. -
#5520: Keeping a Fine-Tune Alive Across Model ReleasesDaniel wants to fine-tune DeepSeek Flash 4.1 on edited podcast scripts. The hard part isn't training — it's surviving the next release.