MSRVTT

doi:doi:10.57702/2sfaor1e

MSRVTT

Followers: 0

Organization

No Organization

There is no description for this organization

License

No License Provided

Export

DCAT(rdf/xml) DCAT(xml) DCAT(N3) DCAT(ttl) DCAT(jsonld) DataCite CSL DublinCore BibTex

MSRVTT

The MSRVTT is a large-scale dataset for video captioning. It contains 10k video clips and each video clip is accompanied with 20 human-edited English sentence descriptions, resulting in 200K video-caption pairs in total.

BibTex:

Before browse our site, please accept our cookies policy