topic · 2 notes
DeepSeek, as it ships.
Engineering notes on DeepSeek by Samir Sengupta - each one read from primary sources on the day it happened, with what it changes for people building on it.
DeepSeek v4.1 Flash pricing, benchmarks, and 890 bytes per token
DeepSeek V4.1-Flash is a 552B MoE with 8B active params at prefill, 16B at decode, 890 bytes per token KV cache, and off-peak rates at 50% of peak.
DeepSeek's agent harness is public, but the sampling config isn't in the README
DeepSeek open-sourced dsh (MIT, 12,293 commits) the same day V4 Pro 0813 hit OpenRouter at $0.435/$0.87 per 1M. Here is what the release specifies.
Hiring for AI or ML?
I am open to AI/ML Engineering, Data Science, and Python roles, plus research collaborations and consulting. New York based, shipping worldwide.