Neev Cloud logoNeev Cloud

    glm-4-7

    neev_cloud/glm-4-7

    ChatToolsSystem prompt
    Use this model

    Context

    Max output

    200K

    Input / 1M

    $57.12

    Output / 1M

    $209.45

    Modality

    Chat

    Cutoff

    Unknown

    About this model

    Flagship GLM model with stronger coding, reliable multi-step reasoning, improved agentic workflows, and enhanced front-end generation quality.

    Best suited for

    • Advanced coding assistants
    • Tool-calling AI agents
    • Front-end code generation
    • Multi-step reasoning workflows
    • Multilingual conversational AI

    Capabilities

    Tools

    Native function calling, so agents can invoke your endpoints.

    System prompt

    Honours a dedicated system role, separate from the user turn.

    Supported parameters

    temperature

    Controls randomness and creativity of responses.

    top_p

    Controls diversity by limiting token probability sampling.

    max_tokensMax Tokens Limit

    Maximum number of tokens generated in the response.

    thinking_mode

    Enables enhanced reasoning and multi-step problem solving.

    tools

    Defines external tools available to the model.

    tool_choice

    Determines whether the model can use tools.

    response_type

    Specifies the output response format.

    parallel_tool_calls

    Allows multiple tools to execute simultaneously.

    stream

    Streams the response in real-time.

    Use this model

    OpenAI-SDK compatible, with gateway fallback and routing across providers.

    Other Neev Cloud models

    ModelContextMax outputInputOutputCapabilities
    minimax-m2.7

    Chat

    66K$0.2852$1.1409Tools, System prompt
    minimax-m2.7-highspeed

    Chat

    66K$57.12$228.49Tools, System prompt
    llama-3.1-8b-instant

    Chat

    131K$4.76$7.62Tools, System prompt
    gpt-oss-20b

    Chat

    131K$7.14$28.56Tools, System prompt
    gpt-oss-120b

    Chat

    131K$9.52$47.60Tools, System prompt
    deepseek-v3-2

    Chat

    131K$26.66$39.99Tools, System prompt
    llama-3.3-70b-versatile

    Chat

    131K$56.17$75.21Tools, System prompt

    Pricing & provider details

    Neev Cloud logo

    Neev Cloud

    neev_cloud/glm-4-7

    Input · per 1M
    $57.12
    Output · per 1M
    $209.45
    Cached input · per 1M
    Context window
    Max output
    200,000 tokens
    Knowledge cutoff
    Unknown
    Auto-router
    Not supported

    Start building today

    Route glm-4-7 — and every other model in the catalogue — through one endpoint, with failover built in.