Skip to main content

hybrid-benchmarking

Benchmark a hybrid-attention model across the LMCache performance ladder (vLLM no-hybrid-allocator → hybrid allocator + prefix caching → hybrid allocator + LMCache) and produce a decode-throughput / TTFT / cache-hit-rate comparison. Use when asked to benchmark, demo, or quantify the value of LMCache caching for a hybrid model.

الانتقال إلى التثبيت

معلومات المصدر

المستودع
LMCache/LMCache
آخر نشاط في المصدر
١٧ يونيو ٢٠٢٦ في ٠١:٠١
لغة SKILL.md المكتشفة
الإنجليزية
النجوم
١١٬٢١٥
التفرعات
١٬٧٣٠

خيارات التثبيت

يُحدَّد Prompt الذي يراجع المصدر أولًا بشكل افتراضي. يمكنك التبديل إلى أمر مباشر أو تنزيل نسخة محلية.

مراجعة ملفات المصدر

اقرأ SKILL.md وأي ملفات مرافقة يعرضها SkillsMP قبل أن تقرر التثبيت.