nebius-deploy-lora
Installation
SKILL.md
Nebius Post-Training: Deploy LoRA Adapter
Serve your fine-tuned LoRA adapter as a serverless endpoint with per-token billing. Two deployment paths: from a Nebius fine-tuning job, or from a local archive / HuggingFace link.
Prerequisites
pip install requests openai
export NEBIUS_API_KEY="your-key"
API: https://api.tokenfactory.nebius.com
Supported base models for serverless LoRA deployment
| Base model | Fine-tuning type |
|---|---|
meta-llama/Meta-Llama-3.1-8B-Instruct |
LoRA + full (LoRA deployable) |
meta-llama/Llama-3.3-70B-Instruct |
LoRA (deployable) |