connections

GET /v1/inference/deployments

The organisation's model deployments and what is actually serving.

Gbogbo àwọn connections endpoints

Gbogbo àwọn ìwé àṣẹ olùgbékalẹ̀ àgbékalẹ̀

Ìfàṣẹ́pọ̀

Fi bọ́kì kọ́kọ́ (bearer token) ránṣẹ́ gẹ́gẹ́ bí àmì ìdánimọ̀ API. Ibùdó yìí kò sọ àṣẹ pàtó kan nínú àlàyé rẹ̀, nítorí náà fún kọ́kọ́ rẹ̀ ní ohun tó kéré jù lọ tí ó nílò kí o sì ṣàyẹ̀wò ìdáhùn náà dípò kí o kàn rò ó.

Ojú abánisọ̀rọ̀ yìí kò gba id ajọ kankan. Kọ́kọ́rọ́ rẹ ti mọ ajọ ti o jẹ ti e, a o si fèsì nipa rẹ̀.

Gbiyanju rẹ

Rọ́pọ̀ èyíkéyìí nínú àwọn àmì ìtọ́ka < > pẹ̀lú iye tirẹ̀, àti àmì ìdánimọ̀ bọ́tìnnì náà pẹ̀lú bọ́tìnnì kan láti inú dásibọ̀ọ̀dù rẹ.

curl -X GET https://api.zinndigital.com/v1/inference/deployments \
  -H "Authorization: Bearer zdk_live_…"

Ṣé o ti wọlé? Iwọ̀n api ní nú ìgbékalẹ̀ rẹ kún id àjọ gidi rẹ ati bọtini tirẹ, o si nṣiṣẹ ibeere na lòdì si api gidi ki o le rii esi gidi na. Ṣí ojú abáná yìí sílẹ̀ nínú kọnsólù API

Àwọn kúlẹ̀kúlẹ̀

Model endpoints on the customer's own inference account. ⛔ replicas is what was asked for and ready_replicas is what is serving. A screen showing only the first reports a healthy deployment that is answering nothing. ready_replicas is null when the provider did not report readiness at all — which is not the same as zero. state includes degraded (some replicas up, some not) and deleting (being torn down, and possibly still holding GPUs). Neither is rounded onto a neighbour: a degraded deployment reported as running hides missing capacity, and a deleting one reported as stopped implies it has stopped costing money.

Idahun

OrúkọIruTí a nílòKini o jẹ
deploymentsInferenceDeployment[]Bẹẹni

Awọn aṣiṣe ti ibudo ipari yii le da pada

401 · 429