Retrieved-Span Training for Efficient Query-Focused Meeting Summarization on QMSum
QMSum ships without a scorer, so we rescored 15 systems under one implementation. A 406M specialist retrained on the same retrieved transcript spans as our 1.2B system scored 36.33 ROUGE-1 against 35.41, a difference QMSum cannot statistically resolve, while using about a third of the parameters and under half the peak inference memory.
arXiv · GitHub · Hugging Face model