Identify the right model for your use case by filtering available foundation models by capabilities and price.
Inference
Last verified 13 Jul 2026
Inference provides a single control plane for managing inference workflows. It includes a Model Catalog where you can view available foundation models, including both DigitalOcean-hosted and third-party commercial models, compare model capabilities and pricing, use routing to match inference requests to the best-fit model, and run inference using serverless or dedicated deployments.
Test and compare foundation models in the Model Playground.
Send API requests directly to foundation models without creating an AI agent or managing infrastructure.
Deploy open-source and commercial LLMs on dedicated GPUs as an inference endpoint.
Route serverless inference requests to foundation models using rules.
Determine which model best fits your specific use case.
Batch Inference lets you run large collections of LLM requests as a single asynchronous job.
Use to build fully-managed AI agents with knowledge bases for retrieval-augmented generation, multi-agent routing, guardrails, and more.
Latest Updates
1 October 2026
-
You can now prepay for Serverless Inference and Managed Agents with the Inference and Agents balance, which no other products draw from. Add $5 to $500 at a time or turn on auto-reload from the Billing page. Funds do not expire. You can also pay for eligible usage at a discount with a monthly plan, now in public preview. See Inference and Agents Balance and How to Pay for Harness Runtime.
-
The Agent Development Kit (ADK), including the
gradient-adkPython package and thegradientCLI, is deprecated. You cannot deploy any new agents with ADK. Agents already deployed with ADK continue to run normally and you are billed for model usage. If you are using a DigitalOcean-hosted model, you are charged for those model keys. Migration guidance and end-of-support timeline will be published in the future.
30 September 2026
-
The following OpenAI model is now available on DigitalOcean Inference for serverless inference and Agent Development Kit:
For more information, see the Available Models page.
29 September 2026
-
Released v1.175.0 of doctl, the official DigitalOcean CLI. This release deprecates the
doctl gradientcommands.doctl gradient agentcommands have no CLI replacement. To manage agents, use the control panel or API.doctl gradient knowledge-basecommands work but will be removed in a future release. Usedoctl knowledge-baseinstead.To list models, regions, and OpenAI keys, use the
doctl serverless-inferencecommands.
For more information, see the full release notes.