Learning Video LLM with Streaming Speech Transcription at Scale (CVPR 2025)
Joya Chen PRO
chenjoya
AI & ML interests
Video LLM
Recent Activity
upvoted
a
paper
about 18 hours ago
Paper2Poster: Towards Multimodal Poster Automation from Scientific
Papers