Available models
Base models supported by Serverless Inference on Crusoe's Intelligence Foundry
Base models supported by Serverless Inference on Crusoe's Intelligence Foundry
Base models supported by Self-Serve Deployments, including LoRA-compatible checkpoints from Serverless Fine-Tuning
Base models supported by Serverless Fine-Tuning for LoRA adapter training
The full reference for dataset formats, hyperparameters, and every serverless fine-tuning operation in Python and curl.
Choose how to run inference on hosted models based on your traffic patterns and latency requirements
How Serverless Inference enforces tokens-per-minute and requests-per-minute limits, and how to interpret 429 and 503 responses
Usage and billing information for Serverless Inference, Serverless Fine-Tuning, and Self-Serve Deployments