AI
AI Digest
EN RU
Home Archive About RSS

#video-understanding

3 items

  • 17 июл VideoChat3: Fully Open Video MLLM for Efficient and Generalist Video Understanding MCG-NJU research
  • 29 июл Mage-VL: An Efficient Codec-Native Streaming Multimodal Foundation Model Microsoft research
  • 10 авг GST-Bench: Can VLMs Develop Global Spatial Awareness from Video? ByteDance Seed research

ai-digest.kerby.pro

© 2026 Alexei Lukin · CC BY 4.0

RSS · JSON Feed · About