glm-4-7
neev_cloud/glm-4-7
Context
—
Max output
200K
Input / 1M
$57.12
Output / 1M
$209.45
Modality
Chat
Cutoff
Unknown
About this model
Flagship GLM model with stronger coding, reliable multi-step reasoning, improved agentic workflows, and enhanced front-end generation quality.
Best suited for
- Advanced coding assistants
- Tool-calling AI agents
- Front-end code generation
- Multi-step reasoning workflows
- Multilingual conversational AI
Capabilities
Tools
Native function calling, so agents can invoke your endpoints.
System prompt
Honours a dedicated system role, separate from the user turn.
Supported parameters
temperatureControls randomness and creativity of responses.
top_pControls diversity by limiting token probability sampling.
max_tokensMax Tokens LimitMaximum number of tokens generated in the response.
thinking_modeEnables enhanced reasoning and multi-step problem solving.
toolsDefines external tools available to the model.
tool_choiceDetermines whether the model can use tools.
response_typeSpecifies the output response format.
parallel_tool_callsAllows multiple tools to execute simultaneously.
streamStreams the response in real-time.