Blogs›Crusoe›Serverless Fine-Tuning
Crusoe has launched Serverless Fine-Tuning, a managed service within Crusoe Intelligence Foundry that allows users to customize open-source models using their own data. This service eliminates the need for users to provision GPU clusters or manage infrastructure, offering a pay-as-you-go model. It supports a variety of base models (Qwen, DeepSeek, Llama, Gemma, gpt-oss, etc.) and data formats (JSONL, Parquet). The pipeline includes data pre-processing (cleaning, tokenization, de-duplication), au. This post details the integration of NVIDIA's Nemotron 3 Ultra model with LangChain Deep Agents, running on Crusoe Cloud. It highlights how Crusoe's Managed Inference and the `langchain-crusoe` integration enable cost-effective deployment of frontier-class open agents, emphasizing harness engineering over model fine-tuning for performance gains. The post also introduces Crusoe's MemoryAlloy KV cache fabric for improved agentic workload performance and discusses tiered model deployment strategies using Nemotron 3 Ultra for orchestration and Nemotron 3 Nano Omni for execution.