anoman

Zhipu (GLM)

GLM Flash Latest

modelglm-flash-latest
BudgetStableStreamingToolsBatchLong contextGlobal

Context

1.3M

Input / 1M tokens

$0.0750

Output / 1M tokens

$0.2500

About this model

GLM Flash Latest

GLM Flash Latest is a budget-tier model available through the Anoman gateway. Detailed performance benchmarks and use-case guidance for this model are coming soon. In the meantime, see the code example below to integrate it into your workflow, and the tier + region information in the hero for pricing and residency.

Code example

Chat completion with GLM Flash Latest

from openai import OpenAI

client = OpenAI(
    base_url="https://api.anoman.io/v1",
    api_key="anm-sk-..."
)

response = client.chat.completions.create(
    model="glm-flash-latest",
    messages=[{"role": "user", "content": "Hello!"}]
)
print(response.choices[0].message.content)

Use GLM Flash Latest through Anoman.