Story

arxiv_cs_lg ยท Aug 25, 2026 ยท paper

Source brief

LAION-BVD: A 10-Million-Hour Open Video Dataset for Multimodal Pre-training

arxiv.orgAug 25, 2026
original source linked

In brief

We present LAION-BVD, a large-scale open video dataset for multimodal learning, which contains 1.3B platform-specific video URLs collected from CommonCrawl. From these, we download 80M videos with a total duration of...

Feed lens
eval

Continue reading

Read the original at arxiv.org โ†’Open in live feedRead that dayโ€™s brief

Earlier in this thread 4 items