Skip to main content

3 docs tagged with "dedicated-inference"

View all tags

Available models

Base models supported by Self-Serve Deployments, including LoRA-compatible checkpoints from Serverless Fine-Tuning

Deploy a self-serve deployment

The full reference for creating, running, and managing self-serve inference deployments—including fine-tuned models—in Python and curl.

Self-Serve Deployments

Reserved-capacity inference deployments on Crusoe's managed infrastructure with predictable performance and no shared rate limits