Dataset and Models focus on long video-language understanding.
Jingyang Lin
jylins
AI & ML interests
vision-and-language
Recent Activity
upvoted a paper about 1 hour ago
GenFirst: Generation Before Reconstruction for Stable End-to-End Latent Generative Modeling updated a dataset 3 months ago
jylins/videomme-v2 published a dataset 3 months ago
jylins/videomme-v2