Menu
Sign In Search Podcasts Charts People & Topics Add Podcast API Pricing
Podcast Image

AI Podcast

InternVideo2.5:通过长文本和丰富上下文建模增强视频多模态大型语言模型

10 Feb 2025

Description

讨论 InternVideo2.5 如何通过长文本和丰富上下文建模来提升视频多模态大型语言模型(MLLM)的性能,包括其架构、训练方法以及在视频理解和特定视觉任务上的实验结果。

Audio
Featured in this Episode

No persons identified in this episode.

Transcription

This episode hasn't been transcribed yet

Help us prioritize this episode for transcription by upvoting it.

0 upvotes
🗳️ Sign in to Upvote

Popular episodes get transcribed faster

Comments

There are no comments yet.

Please log in to write the first comment.