Our
Blog

Latest stories

Thank you! Your submission has been received!
Oops! Something went wrong while submitting the form.
All
Cloud
Engineering
Illustration of analytics dashboards with charts and percentage rings, representing tokenizer benchmarking.
July 31, 2026
fastokens v2: Accelerating inference and training time
Alon Kejzman
Omri Berkovitch
Omer Landau
All
Cloud
Engineering
Diagram showing serverless fine-tuning collapsing a multi-layer model stack into one deployed adapter, from dataset to deployment.
July 29, 2026
From dataset to deployment: A practical guide to serverless fine-tuning
Abhishek Srikanth
All
Cloud
Engineering
Server racks and cloud infrastructure icons representing automated GPU datacenter deployment
July 28, 2026
Before first boot: How Crusoe Pre-Deployment Automation readies every GPU server
Alex Akesson
Angela Zhang
Sylvia Cruz-Albrecht
All
No items found.
Isometric illustration of cluster nodes linked over a network, representing peer-to-peer image distribution
July 28, 2026
Faster image pulls at scale: How peer-to-peer image distribution cuts pull times on Crusoe Managed Kubernetes
Jason Lee
All
General
Cloud
Black-and-white iceberg on a black background, most of its mass hidden underwater, illustrating the hidden cost of AI inference. Crusoe blog header reading "The inference iceberg."
July 23, 2026
Tokenomics in the Age of Agentic Inference
Nikhil Kaul
Stuart Pitts
All
No items found.
Comparing the GLM-5.2 run and two Claude Opus 4.8 runs solving the Wiz Day-One CTF.
July 23, 2026
Solving the Wiz Day-One CTF with Crusoe Managed Inference
Elazar Leibovich
Nir Levy
All
Cloud
Engineering
Isometric cloud, server racks, and storage disks representing Crusoe's managed inference infrastructure.
July 20, 2026
The open agent stack: Running Nous Research’s Hermes Agent on Crusoe Managed Inference
Emmanuel Acheampong
All
Cloud
General
2026 Gartner Magic Quadrant for AI Infrastructure placing Crusoe in the Visionaries quadrant
July 15, 2026
Crusoe named a Visionary because of its ability to execute and completeness of vision
Erwan Menard
All
Cloud
Engineering
Deployment config panel connecting to two GPU server racks, showing the one-click path from settings to a live deployment
July 14, 2026
Self-Serve Deployments: Predictable performance, without the infrastructure lift
Sydnee Mayers
Nir Levy
Peleg Yair
All
Cloud
Dot cluster shifting from white to yellow, illustrating a base model transformed into a fine-tuned model
July 14, 2026
Tune it. Deploy it. Own it: Crusoe's next step in becoming the best cloud for open models
Piyush Kadam
Nir Levy
Janaki Ram Goteti
All
Cloud
Graphic for the Self-Serve Deployments launch: dedicated inference endpoints now live in Crusoe Intelligence Foundry
July 14, 2026
Introducing Self-Serve Deployments for production-scale inference
Sydnee Mayers
Nir Levy
Peleg Yair
All
Cloud
Engineering
Isometric cloud and server illustration for Crusoe Managed Inference hosting the Nemotron 3 model family
July 8, 2026
Frontier Agents at 10x lower cost: NVIDIA Nemotron 3 Ultra + LangChain Deep Agents on Crusoe Managed Inference
Emmanuel Acheampong
No results found

Scaling inference doesn’t need faster GPUs

Learn how to skip the inference bottleneck.

Are you ready to build something amazing?