connections

GET /v1/inference/deployments

The organisation's model deployments and what is actually serving.

Tất cả các điểm cuối connections

Xác thực

Gửi khóa API dưới dạng mã thông báo bearer. Điểm cuối này không nêu rõ quyền cụ thể trong thông số kỹ thuật, vì vậy hãy cấp cho khóa của bạn quyền tối thiểu cần thiết và kiểm tra phản hồi thay vì phỏng đoán.

Endpoint này không nhận ID tổ chức. Khóa của bạn đã xác định tổ chức mà nó thuộc về và phản hồi được giới hạn trong phạm vi đó.

Dùng thử

Thay thế bất kỳ nội dung nào trong ngoعل (angle brackets) bằng giá trị của riêng bạn và trình giữ chỗ key bằng một key từ trang tổng quan của bạn.

curl -X GET https://api.zinndigital.com/v1/inference/deployments \
  -H "Authorization: Bearer zdk_live_…"

Đã đăng nhập? Bảng điều khiển API trong trang quản lý của bạn sẽ tự điền ID tổ chức thực tế và khóa của riêng bạn, sau đó chạy yêu cầu đối với API trực tiếp để bạn có thể xem phản hồi thực tế. Mở điểm cuối này trong bảng điều khiển API

Chi tiết

Model endpoints on the customer's own inference account. ⛔ `replicas` is what was **asked for** and `ready_replicas` is what is **serving**. A screen showing only the first reports a healthy deployment that is answering nothing. `ready_replicas` is `null` when the provider did not report readiness at all — which is not the same as zero. `state` includes `degraded` (some replicas up, some not) and `deleting` (being torn down, and possibly still holding GPUs). Neither is rounded onto a neighbour: a `degraded` deployment reported as `running` hides missing capacity, and a `deleting` one reported as `stopped` implies it has stopped costing money.

Phản hồi

TênLoạiBắt buộcNội dung này là gì
deploymentsInferenceDeployment[]

Các lỗi điểm cuối này có thể trả về

401 · 429