Schemas#
The ray.serve.schema module defines the models that describe Serve’s state and configuration. The Serve REST API accepts the config schemas and returns the response schemas.
Status and details#
Detailed info about a Ray Serve actor. |
|
Detailed info about a Ray Serve ProxyActor. |
|
Describes the status of an application and all its deployments. |
|
Describes the status of Serve. |
|
Describes the status of a deployment. |
|
Encoding type for the serve logs. |
|
PublicAPI (alpha): This API is in alpha and may change before becoming stable. |
|
PublicAPI (alpha): This API is in alpha and may change before becoming stable. |
|
One autoscaling decision with minimal provenance. |
|
Deployment-level autoscaler observability. |
|
Replica rank model. |
Abstract base class for task processing adapters. |
Config schemas#
Multi-application config for deploying a list of Serve applications to the Ray cluster. |
|
Options to start the gRPC Proxy with. |
|
Options to start the HTTP Proxy with. |
|
Describes one Serve application, and currently can also be used as a standalone config to deploy a single application to a Ray cluster. |
|
Options with which to start a replica actor. |
|
Celery adapter config. You can use it to configure the Celery task processor for your Serve application. |
|
Task processor config. You can use it to configure the task processor for your Serve application. |
|
Task result Model. |
|
Request schema for scaling a deployment's replicas. |
Response schemas#
Serve metadata with system-level info and details on all applications deployed to the Ray cluster. |
|
Detailed info about a Serve application. |
|
Detailed info about a deployment within a Serve application. |
|
Detailed info about a single deployment replica. |
|
PublicAPI (alpha): This API is in alpha and may change before becoming stable. |
|
PublicAPI (alpha): This API is in alpha and may change before becoming stable. |
|
Represents a node in the deployment topology. |
|
Represents the dependency graph of deployments in an application. |
|
Health metrics for the Ray Serve controller. |
|
Statistics for a collection of duration/latency measurements. |
Tracks the type of API that an application originates from. |
|
The current status of the application. |
|
The current status of the proxy. |