ai-llm-inference

Installation
SKILL.md

LLM Inference Production Skill Hub

Operational guidance for choosing and tuning modern inference stacks. Focus on runtime fit, routing, output guarantees, adapter loading, multimodal serving, and measured performance under load.

Use current primary sources for volatile facts such as versions, hardware support, benchmarks, pricing, and release status.

ASCII Flow

Installs
160
GitHub Stars
79
First Seen
Jan 23, 2026
ai-llm-inference — vasilyu1983/ai-agents-public