vietanh.dev
HomeAboutProjectsBlogCourses & Books

What is: VideoBERT?

SourceVideoBERT: A Joint Model for Video and Language Representation Learning
Year2000
Data SourceCC BY-SA - https://paperswithcode.com

VideoBERT adapts the powerful BERT model to learn a joint visual-linguistic representation for video. It is used in numerous tasks, including action classification and video captioning.

Collections

Transformers
Representation-Learning

Previous Term

H3DNet

Next Term

Support-set Based Cross-Supervision
← Back to the glossary list
Viet-Anh on Software Logo

Viet-Anh on Software

AI systems under real constraints

Measured studies, open-source implementations, and production lessons across edge AI, agent reliability, privacy, cost, and security.

Machine Learning Lead at Scopic SoftwareFounder and maintainer at Neural Research Lab

Explore

  • Home
  • About
  • Projects
  • Contact

Content

  • Blog
  • Notes
  • Courses & Books
  • Videos
  • Glossary

Tools & Resources

  • Source Code: Based On Tailblaze
  • PageSpeed Insights
© 2026 Viet-Anh Nguyen. All rights reserved.
Privacy PolicyAI TransparencyTerms of ServiceLoginPageSpeed