Introducing the MLPerf End-to-End RAG Inference Benchmark
MLPerf Inference now benchmarks complete RAG pipelines — ingestion, retrieval, and multi-hop reasoning — rather than single models in isolation.
MLCommons · lori@mlcommons.org
Topics: Benchmarks
Entities: Benchmarks