Benchmark

SummScreenFD

summarization long_context text

SummScreenFD is the ForeverDreaming subset of the SummScreen dataset for abstractive screenplay summarization, comprising pairs of TV series transcripts and human-written recaps from 88 different shows. The dataset provides a challenging testbed for abstractive summarization where plot details are often expressed indirectly in character dialogues and scattered across the entirety of the transcript, requiring models to find and integrate these details to form succinct plot descriptions.

语言EN
满分1
参评模型2

模型排名

名次 模型 机构 分数 来源
1 Phi-3.5-MoE-instruct Microsoft 16.9 来源 ↗
2 Phi-3.5-mini-instruct Microsoft 16.0 来源 ↗