Emad Mostaque
speaker
1,351 appearances
1 recordings
1 series
first heard Dec 2024
last heard Dec 2024
Emad Mostaque’s voice in public audio — every appearance, attributed to the second.
Trend
recordings per month · last 12 monthsNo recordings in the last 12 months.Older appearances are listed below; set an alert to hear about the next one.
Appearances
Like we saw the Sana team, previously Pixart, create a stable diffusion level model with twenty five million images.
And we use two billion.
What is the lower bound of data that you need to have for a language model to achieve X to teach your kid?
Does it need to have seen Reddit?
No.
Does it need to have seen this?
If you want to build AGI, yeah, you need all that data.
Maybe.
But do you need it all in one place?
Like what is the limited data analysis?
And we've seen interesting things like I think Clear AI just did like an AI that was just trained on knowledge up to three hundred B C.
I haven't seen what it's like, but those are the types of things that very it really interest me.
Like, what does that look like?
So it feels like AI.
Right.
It feels like we're it feels like we have an order of magnitude improvement still to come, just from data, honestly.
And we're seeing this with fine web and kind of other things and the synthetic data sets and augmented data sets that we've had.
Yeah, they did thirteen point eight epochs on like four hundred billion tokens for the core.
But then they have to add in like a map of the internet because they're so boring.
And people don't want to interact with boring models.
Showing 661–680 of 1,351 · page 34 of 68
← Previous
Next →