llm-pipeline
Installation
SKILL.md
llm-pipeline
Wire multiple LLM calls into a reliable, controllable production pipeline. You chain steps where one call's validated output feeds the next, put a router in front of providers so an outage fails over instead of taking you down, and engineer the cross-cutting concerns: timeouts, bounded retries, fallbacks, caching, and cost caps.
Treat the LLM as an unreliable network dependency, not a local function call. Every rule below follows from that: providers have outages, rate limits, and latency tails, so no single provider is a single point of failure and no call is allowed to run unbounded.
Do you even need a pipeline?
This skill is the orchestration around calls. If you only have one call, you are in the wrong place.
| Situation | Go to |
|---|---|
| Make one prompt better, few-shot, system-prompt design | ../prompt-engineering/SKILL.md |
| One call must return a typed object validated against a schema | ../structured-extraction/SKILL.md |
| The model decides its own next step / tool to call | ../building-agents/SKILL.md |
| Chunk/embed/retrieve context to stuff into a prompt | ../rag/SKILL.md |
| Pure spend ledger / attribution / dashboard | ../cost-tracking/SKILL.md |
| Fixed multi-step flow + reliability layer | here |