open-llm-leaderboard-old/details_xformAI__facebook-opt-125m-qcqa-ub-6-best-for-KV-cache Updated Jan 23, 2024 • 21 • 1
open-llm-leaderboard-old/details_saarvajanik__facebook-opt-6.7b-gqa-ub-16-best-for-KV-cache Updated Jan 28, 2024 • 20 • 1
saarvajanik/facebook-opt-6.7b-gqa-ub-16-best-for-KV-cache Text Generation • Updated Jan 28, 2024 • 65 • 1
nintwentydo/pixtral-12b-FP8-dynamic-FP8-KV-cache Image-Text-to-Text • 13B • Updated Jan 6, 2025 • 52 • 2
open-llm-leaderboard-old/details_xformAI__opt-125m-gqa-ub-6-best-for-KV-cache Updated Jan 23, 2024 • 31 • 1
Boxoffice1280/Neurips2026_evaluating_accuracy_KV-cache_reuse_techniques Viewer • Updated May 6 • 5.68k • 128 • 1
open-llm-leaderboard-old/details_saarvajanik__facebook-opt-6.7b-qcqa-ub-16-best-for-KV-cache Updated Jan 28, 2024 • 28 • 1
position-specialist-speculative-decoding/Speed-E3-Llama3.1-8B-Instruct-vllm 0.9B • Updated Jan 28 • 3 • 1