MLPerf Inference now benchmarks complete RAG pipelines — ingestion, retrieval, and multi-hop reasoning — rather than single models in isolation.