30 Aug 2025
29m
【多模态RAG:让AI看视频并答题】视频问答整体框架!多模态大模型教程:qwen2.5-VL视频图像理解;Whisper音频提取;clip模型视频和文字的桥梁 卢菁博士 #人工智能 #rag #ai
Dr.LuAIclass 卢菁 北大博士后 AI 专家
Open in Podwise to generate AI notes
Sign in to process this episode and unlock summaries, transcripts, highlights and translations.
Shownotes are not generated by Podwise.

