Skip to content
Advertisement
AudioImageMultimodalText

Omni-StoryBench

Omni-StoryBench

Omni-StoryBench Omni-StoryBench is a context-aware omnimodal story generation benchmark.Each sample provides a current story page and requires generating the next page’s image, narration text, and speech utterance. Dataset Structure The dataset contains: data/testset.jsonl: Main benchmark file. images/: Page images. texts/: Page text files. speech/: Generated speech audio files. instruction/: Source-level instruction metadata. Data Fields Each JSONL sample… See the full description on the dataset page:

Source: Hugging Face Hub (omnibench/anonymous-storybench). Metadata imported from the dataset’s Hub tags.

Advertisement