Release Note

Last verified 28 Apr 2026

You can now evaluate models available for serverless inference, inference routers, and dedicated inference deployments using a judge model. Scoring includes metrics such as correctness, completeness, ground truth faithfulness, and safety metrics. This features is in public preview. You can opt in from the Feature Preview page. For more information, see Evaluate Models.

We can't find any results for your search.

Try using different keywords or simplifying your search terms.