TensorRT
product
1 mentions
1 recordings
first heard Feb 2026
last heard 24 Feb
↑1 vs the 12 months before
NVIDIA runtime optimizer used for accelerating model inference
Trend
Peak in the week of 21 Feb — 1 mentions across 1 series.
Moments
newest first · ▶ plays the momentStefano ErmonThe Neuron: AI Explained · Diffusion for Text: Why Mercury Could Make LLMs 10x Faster · 15:17 · 24 Feb
…Things like VLLM, SGLang, TensorRT, like there is pretty mature serving stacks for ultra-aggressive models.…
All 1 moments