Deep Seek V3

product
3 mentions 3 recordings first heard Apr 2025 last heard 4d ago ↑2 vs the 6 months before

Sparse AI model optimized for large batch inference

Trend

mentions per week · public audio 7d30d6M12M
1 · 16 Sep MarMayJunJulSepnow

Peak in the week of 16 Sep — 1 mentions across 1 series.

Moments

newest first · ▶ plays the moment
Thomas SohmersThe Twenty Minute VC (20VC): Venture Capital | Startup Funding | The Pitch · 20VC: "Anti-Data Centres is a Chinese Psyop" | How Many Plan... · 57:12 · 4d ago
…And so, you know, Deep Seek, beginning of 2025 with uh Deep Seek V3 had, you know, made a lot…
· Open transcript →
Dwarkesh PatelDwarkesh Podcast · 8 Predictions for the Era of Continual Learning · 7:21 · 7 Aug
…Back of the envelope math suggests that the optimal inference batch size for a sparse model, like say Deep Seek V3, is more than 2400 concurrent sequences being generated at once.…
· Open transcript →
Jordi HaysTBPN · Meta AI Deep Dive, Jeff Huber, Sheel Mohnot, Leif Abraham, S... · 28:18 · Apr 2025
…And so uh this founder uh benchmarked Lama 4 against other models and found that Llama 4 came in uh maybe 8th below GPT-4.0 mini, below Cloud 3.5 Sonnet, and below Deep Seek V3.…
· Open transcript →
All 3 moments