Vibhu (Vibhu?)

speaker
157 appearances 1 recordings 1 series first heard Jul 2026 last heard 23 Jul

Vibhu (Vibhu?)’s voice in public audio — every appearance, attributed to the second.

Trend

recordings per month · last 12 months
1 · Jul OctJan 26AprJulnow

Recordings per month over the last 12 months — 1 in all, peaking in Jul 2026 with 1.

Appearances

newest first · ▶ plays the moment
You can learn a lot more from the web.
Um, at the same time you took 30B and scaled it up to 120B, right?
Um, is there any gating on how small is too small?
So I'm, I'm just going to ramble for a bit.
I'll come to a question at the end, but you know, part of Karpathy's thesis was cognitive core, right?
We've seen vibe thinker, Nanbage, 3B, 4Bs that reason a lot.
And then, you know, the, the idea is,
you offload to a different model for the work.
These are small reasoning models.
So have you found anything interesting in model sizes, like 20, 30 Bs on device, 100 Bs on single GPU?
Can you squeeze out more there?
And that's on all axes of, there's like an axis of how long a model will reason.
So how long can it stay agentic?
And there's also efficiency, right?
You want to ideally push on both.
And the thing to clarify you guys aren't doing right now, which we do see at Frontier Labs is the distillation, right?
You have a big, big model that you don't really ship to users and what you put out for inference is typically distilled from that, which gets you quite a bit of gains, right?
Showing 141–157 of 157 · page 8 of 8 ← Previous