Deep Seek V3
product
3 mentions
3 recordings
first heard Apr 2025
last heard 4d ago
↑1 vs the 12 months before
Sparse AI model optimized for large batch inference
Trend
Peak in the week of 16 Sep — 1 mentions across 1 series.
Moments
newest first · ▶ plays the momentThomas SohmersThe Twenty Minute VC (20VC): Venture Capital | Startup Funding | The Pitch · 20VC: "Anti-Data Centres is a Chinese Psyop" | How Many Plan... · 57:12 · 4d ago
…And so, you know, Deep Seek, beginning of 2025 with uh Deep Seek V3 had, you know, made a lot…
Dwarkesh PatelDwarkesh Podcast · 8 Predictions for the Era of Continual Learning · 7:21 · 7 Aug
…Back of the envelope math suggests that the optimal inference batch size for a sparse model, like say Deep Seek V3, is more than 2400 concurrent sequences being generated at once.…
Jordi HaysTBPN · Meta AI Deep Dive, Jeff Huber, Sheel Mohnot, Leif Abraham, S... · 28:18 · Apr 2025
…And so uh this founder uh benchmarked Lama 4 against other models and found that Llama 4 came in uh maybe 8th below GPT-4.0 mini, below Cloud 3.5 Sonnet, and below Deep Seek V3.…
All 3 moments