Beta

glm-4.5-flash

By Zhipu AI

Text only: this model does not read images.

Model name for the APIglm-4.5-flash
Uptime, last 3 days
1 call so far
Too few to give a percentage
Credits per reply
Free
Limited by number of calls

Your chats and training

Used to train and improve the provider’s models.

The company that answers for this model states it may use prompts and replies to improve or train its models. Do not send anything private to it.

Which company answers for a model is not published here, so this says what the policy is, not whose it is. Checked against their terms on Sep 28, 2026.

Model facts

Context length
128K tokens
Longest reply
96K tokens
Calls tools
Yes
Knowledge cutoff
Apr 2025
Released
28 Jul 2025
Open weights
Yes

Context, reply length, tool calling, knowledge cutoff, release date and open weights come from the models.dev open dataset, read once a day.

Call it through the API

Any OpenAI-compatible client works: use the address in the sample and this model name.

import requests

response = requests.post(
    "https://onomeo.com/v1/chat/completions",
    headers={"Authorization": "Bearer YOUR_API_KEY"},
    json={"model": "glm-4.5-flash", "messages": [
        {"role": "user", "content": "Hello!"}
    ]}
)
print(response.json())

Figures come from real calls made through this site. Models are served by third-party providers and may slow down or be withdrawn; this page follows those changes.