Crusoe Serverless Fine-Tuning

Tune it.
Deploy it.
Own it.

Iterate faster with Serverless Fine-Tuning.
No infrastructure wrangling. No surprise bills. Just breakthroughs.

Fine-tuning, simplified

01
Select your base model
Select your base model
Choose from a curated library of top models. Find the perfect fit for your specific use case, whether you need lightweight agility or frontier reasoning power.
02
Load your dataset
Load your dataset
Upload your training data to your isolated environment in standard JSONL or Parquet format.
03
Configure your fine-tuning job
Configure your fine-tuning job
Tune the settings with pre-configured defaults built on best practices. We run the LoRA (Low-Rank Adaptation) workflow for fast, precise, and cost-effective fine-tuning. Submit your job via the UI, SDK, or API.
04
Deploy or download your model
Deploy or download your model
Go live in one click using Self-Serve Deployments for inference in Crusoe Intelligence Foundry, or download your weights to deploy anywhere.

Choose from top LLMs to fine-tune

Kimi
models
Model
Parameters
Context
Yutori
models
Model
Parameters
Context
DeepSeek
models
Model
Parameters
Context
Google
models
Model
Parameters
Context
OpenAI
models
Model
Parameters
Context
Meta
models
Model
Parameters
Context
Alibaba
models
Model
Parameters
Context
Z.ai
models
Model
Parameters
Context
NVIDIA
models
Model
Parameters
Context

DeepSeek V4 Flash

Parameters
158.1B
Context
1,048,576

Gemma 4 31B it

Parameters
32.7B
Context
262,144

GLM 5.2

Parameters
379.2B
753.3B
Context
262,144
1,048,576

GPT-OSS 120B

Parameters
120.4B
Context
131,072

GPT-OSS 20B

Parameters
21B
Context
128,000

Llama 3.3 70B Instruct

Parameters
70.6B
Context
131,072

Llama 3.1 8B Instruct

Parameters
8B
Context
128,000

Qwen3 235B A22B Instruct 2507

Parameters
235.1B
Context
262,144

Qwen3 8B

Parameters
8.2B
Context
32,000

Qwen3.5 9B

Parameters
9B
Context
256,000

Qwen3.5 2B

Parameters
2B
Context
256,000

Qwen3.6 35B A3B

Parameters
35B
Context
256,000

Nemotron 3.5 Lightning

Parameters
30B
Context
1,000,000

More discovery. Less drudgery.

Improve accuracy, fix outputs, and ship faster.

Features

AI-optimized infrastructure

Tune models reliably and cost-effectively on purpose-built AI infrastructure, without having to manage it. If a hardware blip occurs, the system automatically recovers and restarts, so downtime stays minimal.

Cost-effective fine-tuning

Only pay for what works. Token-based pricing tracks spend directly to model progress. Early stopping halts the job and the billing the moment your LLM stops improving.

No lock-in

Your model goes with you. Deploy in one click to Self-Serve Deployments for inference, or download raw weights in .safetensors format to use any other model deployment platform.

Full lineage,
zero guesswork

Fine-tuned models carry full lineage from Crusoe Object Store back to the exact dataset, config, and eval that produced it. Any run is reproducible and any outcome is auditable.

Observability
that travels

View detailed granular training metrics. Audit trails are included on every job.

Your workload data. Your models. Your keys.

Trust the infrastructure underneath your models.

Bring your own KMS keys to encrypt fine-tuning data and model artifacts with Crusoe Customer-Managed Keys (CCMK) — no exceptions, no shared custody.

  • Permission is yours to revoke at any moment. We never hold your key in plaintext; we simply ask for permission to use it.

  • One less blocker for security and compliance teams navigating SOC 2 or ISO 27001 requirements

Get started
Our early experience with Crusoe's Serverless Fine Tuning product was seamless, and it worked like a charm. We look forward to leveraging it to optimize the latency and cost of our AI agents as we scale our infrastructure.
Dr. Will Leeney
AI Researcher
Dr. Hiskias Dingeto
AI Researcher

Resources

Resources to help you configure, launch, and evaluate fine-tuning jobs without rebuilding your stack.

Frequently
asked questions

Are you ready to build something amazing?

A rural landscape showing hybrid generation, a large array of solar panels alongside a canal and a line of wind turbines.