#001a3m-router1 skills101updated 2026-05-15100% of creatorskilloccupationdescriptionupdatedtmlpdsoftware-developersResearch-backed Multi-LLM Router with parallel execution, streaming, caching, token compression (ISON), local provider support (Ollama/vLLM/LM Studio), batch processing. Based on arXiv research: RouteLLM routing, RadixAttention prefix caching, Medusa/EAGLE speculative decoding. Python bindings for LangChain/LlamaIndex/AutoGen/CrewAI. 120+ keywords for LLM/ML discoverability. Use for multi-model comparison, cost optimization, batch processing, local privacy, context compression, adaptive routing.2026-05-15Showing top 1 of 1 collected skills in this repository.Load 0 more skillsLoading skills...