Skip to main content

4 docs tagged with "throughput"

View all tags

Available models

Base models supported by Self-Serve Deployments, including LoRA-compatible checkpoints from Serverless Fine-Tuning

Deploy a self-serve deployment

The full reference for creating, running, and managing self-serve inference deployments—including fine-tuned models—in Python and curl.

Self-Serve Deployments

Reserved-capacity inference deployments on Crusoe's managed infrastructure with predictable performance and no shared rate limits

Usage and billing

Usage and billing information for Serverless Inference, Serverless Fine-Tuning, and Self-Serve Deployments