Skip to content

GLM 5.3 Flash — Shield

glm.5.3-flash-shield

A confidential multimodal route for text, images, tools and structured outputs, with streaming.

CHF 0.365 per million input tokens and CHF 1.215 per million output tokens. Cache pricing does not apply.

GLM 5.3 Flash — Shield

A confidential multimodal route for text, images, tools and structured outputs, with streaming.

  • Shield combines a protected execution environment with a receipt signed by Sealarca. This receipt documents the exchange observed by our gateway.
  • Hashes cover JSON bodies and SSE streams between the Sealarca gateway and the encryption proxy, before response conversions. They do not identify the original client HTTP bytes. Manifest digests are deployment observations without an independent hardware binding to the response.
  • Supply the conversation in every request. store=true, background, previous_response_id, conversation and cache_salt are rejected. Server-executed tools are unavailable.
  • reasoning_effort accepts low, high or max. The default is low; reasoning cannot be disabled. Reasoning tokens are billed as output tokens, including when a stream is interrupted before an answer is displayed.
  • Model context: up to 1,048,576 tokens. Output cap configured by Sealarca: 131,072 tokens per request. Route quotas and effective limits may reduce this cap. Responses is bridged to Chat Completions. Preview route.
  • Signed receipts and metadata are retained for 90 days. Receipts and logs contain no prompts, responses or images. Shared caching is disabled.
POST /v1/responsesbash
curl https://api.sealarca.ch/v1/responses \
  -H "Authorization: Bearer $SEALARCA_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
  "model": "glm.5.3-flash-shield",
  "input": "Reply with OK.",
  "reasoning": {
    "effort": "low"
  },
  "store": false,
  "max_output_tokens": 256
}'
POST /v1/chat/completionsbash
curl https://api.sealarca.ch/v1/chat/completions \
  -H "Authorization: Bearer $SEALARCA_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
  "model": "glm.5.3-flash-shield",
  "messages": [
    {
      "role": "user",
      "content": "Reply with OK."
    }
  ],
  "reasoning_effort": "low",
  "max_tokens": 256,
  "stream": true
}'
POST /v1/messagesbash
curl https://api.sealarca.ch/v1/messages \
  -H "Authorization: Bearer $SEALARCA_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
  "model": "glm.5.3-flash-shield",
  "max_tokens": 256,
  "messages": [
    {
      "role": "user",
      "content": "Reply with OK."
    }
  ]
}'

Read X-Sealarca-Receipt-ID and X-Sealarca-Receipt-URL from the response. The receipt becomes available after finalization and delivery; a 404 may be temporary. complete, interrupted and failed describe the observed exchange.

Shield receiptbash
# X-Sealarca-Receipt-ID / X-Sealarca-Receipt-URL / X-Sealarca-Proof-Profile
curl https://api.sealarca.ch/v1/vault/proofs/$RECEIPT_ID/bundle \
  -H "Authorization: Bearer $SEALARCA_API_KEY"
# JWKS: https://api.sealarca.ch/v1/vault/proofs/signing-keys
View models and pricing