Dataset · Text to video
MaBiao
This project is a massive collection of prompts used for Bytedance's Seedance 2.0 and the resulting generated videos. The entire dataset exceeds 50GB and contains 8100+ videos, all structured into a comprehensive dataset. Due to GitHub's limitations with large file storage, the full dataset is hosted on Hugging Face. The Hugging Face repository contains the generated videos (.mp4), cover images (.jpg), and a highly structured.jsonl file that holds all prompt metadata. No login required, lighting-fast response. Launched by GokuOpenLab, seedance-2-prompts-datasets is a prompt data infrastructure project created for developers and researchers. In the current AI ecosystem, prompts are the new…
Publicly accessible
cc-by-4.0
1K<n<10K
This dataset package defines the clean HF repository layout and contains the public metadata needed to run the PAWBench Full31 coverage and Table1 PAW-Cal PAWEval rerun inputs. In the local no-copy handoff, videos are not duplicated under this directory; they are uploaded into sharded videos/bymediaid/ subdirectories from the private upload plan after explicit human approval. - videos/bymediaid/ /.mp4: one physical video per mediaid after human upload from the private upload plan. - manifests/sharedmediaregistry.jsonl: public media identity registry. - manifests/full31pawevalmanifest.jsonl: Full31 logical PAWEval rows. - manifests/table1pawevalrerunrows.jsonl: Table1 PAW-Cal rerun rows.…
Access requested at publisher
cc-by-4.0