Skip to content

Research

Page 39 of 104

Automatic Labelling of Speech Translation Errors

research note

Automatic Labelling of Speech Translation Errors

·7 min read·Dominik Macháček, Maike Züfle, Ondrej Klejch

This paper addresses the challenge of evaluating quality and confidence in Speech Translation (ST) systems by proposing the Speech Translation Error Labelling (STEL) task

researchspeech-translationerror-labellingquality-estimationmultimodal-llm

Read note → Source paper ↗

Benchmark Everything Everywhere All at Once

research note

Benchmark Everything Everywhere All at Once

·8 min read·Shiyun Xiong, Dongming Wu, Peiwen Sun et al.

This paper addresses the challenges in constructing benchmarks for evaluating large language models (LLMs) and multimodal large language models (MLLMs), focusing on the problems of labor-intensive …

researchlarge-language-modelsbenchmark-constructionautonomous-agentmultimodal-evaluation

Read note → Source paper ↗

Articles are CC BY 4.0 — feel free to quote with attribution