ai-llm-inference
Installation
SKILL.md
LLM Inference Production Skill Hub
Operational guidance for choosing and tuning modern inference stacks. Focus on runtime fit, routing, output guarantees, adapter loading, multimodal serving, and measured performance under load.
Use current primary sources for volatile facts such as versions, hardware support, benchmarks, pricing, and release status.