LAMBADA (Language Modeling Benchmark)
LAMBADA is a benchmark designed to evaluate the ability of language models to predict the last word of sentences, emphasizing contextual understanding.
Free Information Center
LAMBADA is a benchmark designed to evaluate the ability of language models to predict the last word of sentences, emphasizing contextual understanding.
BIG-bench is a large-scale benchmark designed to evaluate the capabilities of language models across diverse and challenging tasks. It aims to provide a comprehensive assessment of model performance beyond conventional benchmarks.