Nikhila Ravi

speaker
227 appearances 1 recordings 1 series first heard Dec 2025 last heard 18 Dec

Nikhila Ravi’s voice in public audio — every appearance, attributed to the second.

Trend

recordings per month · last 12 months
1 · Dec OctJan 26AprJulnow

Recordings per month over the last 12 months — 1 in all, peaking in Dec 2025 with 1.

Appearances

newest first · ▶ plays the moment
One of the tables in the paper, which shows the training data set distribution.
I think it's table 24.
We have about 70, more than 70% of the annotations are these like negative phrases that are not present in the image.
So we have to really train the model to not detect stuff that is not in the image.
We basically add this presence token to the model, which explicitly separates the task of recognition and r and localization.
So Yeah we know.
Basically it it simplifies the task.
And so the model doesn't have to try to do everything with just the um with just the proposals in the
be able to l have this global like sort of learned token just for the recognition part.
Yeah, the architecture diagram.
Yeah, I mean th this this is nice, this diagram, I'm glad you brought it up because Sound3 isn't just a version bump.
It's, you know, an entirely new approach to do segmentation.
It's like this, you know, new interface.
For segmentation and it combines so many different tasks where previously you would have needed a task-specific model for each of these tasks.
You know, interactive segmentation, uh, text prompting, um, open vocabulary detection, tracking, like all of these tasks, you would have needed a separate model.
And so, you know, really had to do a lot of work to bring it together.
I think one of the things
we did was really decouple the
detection component and the tracking components.
So you can see, you know, we still preserve the tracking components from SAM2.
Showing 101–120 of 227 · page 6 of 12 ← Previous Next →