CONCENTRATE

One API for every major LLM provider — routing, spend, logs, and controls in one place.

New York

130 E 59th St, 17th floor

New York, NY 10022

Wilmington

1201 N. Market Street, Suite 200

Wilmington, DE 19801

LLM Gateway
  • LLM Gateway
  • Request Routing
  • Usage Monitoring
  • Spend Management
  • Data Security
  • Access Controls
Features
  • Universal API Keys
  • BYOK
  • Spend Tracking
  • Token Allocation
  • Usage Analytics
  • Request Logs
  • Alerts
  • Data Redaction
  • Zero Data Retention
  • Audit Logs
Teams
  • AI Engineering
  • Engineering Leadership
  • Finance & Operations
  • Security & Compliance
Integrations
  • All Integrations
  • Migration Guides
Platform
  • Pricing
  • Model Fortress
  • Enterprise
  • Documentation
  • Status
Legal
  • Privacy Policy
  • Terms of Service
  • Data Processing Addendum
  • Acceptable Use Policy
  • Trust Center

LLM Gateway

  • LLM Gateway
  • Request Routing
  • Usage Monitoring
  • Spend Management
  • Data Security
  • Access Controls

Features

  • Universal API Keys
  • BYOK
  • Spend Tracking
  • Token Allocation
  • Usage Analytics
  • Request Logs
  • Alerts
  • Data Redaction
  • Zero Data Retention
  • Audit Logs

Teams

  • AI Engineering
  • Engineering Leadership
  • Finance & Operations
  • Security & Compliance

Integrations

  • All Integrations
  • Migration Guides

Platform

  • Pricing
  • Model Fortress
  • Enterprise
  • Documentation
  • Status

Legal

  • Privacy Policy
  • Terms of Service
  • Data Processing Addendum
  • Acceptable Use Policy
  • Trust Center

Offices

New York

130 E 59th St, 17th floor

New York, NY 10022

Wilmington

1201 N. Market Street, Suite 200

Wilmington, DE 19801

AICPA SOC for Service OrganizationsSOC 2 Type II

© 2026 Concentrate AI. All rights reserved.

CONCENTRATE
PricingModelsDocs
Back to Model Fortress

GLM-5.3-Flash

z.ai
glm-5.3-flash
Released Aug 26, 2026
Chat
Providers
12
Context
1.0M
Input
$0.07/M
Output
$0.25/M
Aliases:glm5.3-flashglm-53-flashzhipu-glm-5.3-flash

Z-AI's first natively multimodal GLM-5 model: image and PDF input on a 1M-token context at flash pricing. Reasoning is always on (low/high/max), with function calling and implicit prompt caching.

Z-AI's first natively multimodal GLM-5 model: image and PDF input on a 1M-token context at flash pricing. Reasoning is always on (low/high/max), with function calling and implicit prompt caching.

z.ai

zai/glm-5.3-flash
1.0M context131K max outQuant: Unspecified
FlexPriority / Fast
Input Price
$0.15/M tokens
Output Price
$0.50/M tokens
Cache read: $0.03/M

Runware

runware/glm-5.3-flash
1.0M context131K max outQuant: Unspecified
FlexPriority / Fast
Input Price
$0.15/M tokens
Output Price
$0.50/M tokens
Cache read: $0.03/M

DeepInfra

deepinfra/glm-5.3-flash
1.0M context131K max outQuant: fp4
FlexPriority / Fast
Input Price
$0.07/M tokens
Output Price
$0.25/M tokens
Cache read: $0.01/M

Cloudflare

cloudflare/glm-5.3-flash
1.0M context131K max outQuant: Unspecified
FlexPriority / Fast
Input Price
$0.15/M tokens
Output Price
$0.50/M tokens
Cache read: $0.03/M

Fireworks AI

fireworks/glm-5.3-flash
1.0M context131K max outQuant: Unspecified
Flex
Input Price
$0.15/M tokens
Output Price
$0.50/M tokens
Cache read: $0.03/M

Novita

novita/glm-5.3-flash
1.0M context131K max outQuant: fp8
FlexPriority / Fast
Input Price
$0.15/M tokens
Output Price
$0.50/M tokens
Cache read: $0.03/M

Azure Fireworks

azure-fw/glm-5.3-flash
1.0M context131K max outQuant: Unspecified
FlexPriority / Fast
Input Price
$0.19/M tokens
Output Price
$0.63/M tokens
Cache read: $0.04/M

Crusoe

crusoe/glm-5.3-flash
1.0M context131K max outQuant: Unspecified
FlexPriority / Fast
Input Price
$0.15/M tokens
Output Price
$0.50/M tokens
Cache read: $0.03/M

Baseten

baseten/glm-5.3-flash
1.0M context262K max outQuant: Unspecified
FlexPriority / Fast
Input Price
$0.15/M tokens
Output Price
$0.50/M tokens
Cache read: $0.03/M

Modal

modal/glm-5.3-flash
1.0M context131K max outQuant: Unspecified
FlexPriority / Fast
Input Price
$0.45/M tokens
Output Price
$1.50/M tokens
Cache read: $0.09/M

Wafer

wafer/glm-5.3-flash
1.0M context131K max outQuant: Unspecified
FlexPriority / Fast
Input Price
$0.75/M tokens
Output Price
$0.50/M tokens
Cache read: $0.03/M

Nebius

nebius/glm-5.3-flash
1.0M context131K max outQuant: Unspecified
FlexPriority / Fast
Input Price
$0.15/M tokens
Output Price
$0.50/M tokens
CONCENTRATE
Blog
Request a DemoLog In
Log In
CONCENTRATE

One API for every major LLM provider — routing, spend, logs, and controls in one place.

New York

130 E 59th St, 17th floor

New York, NY 10022

Wilmington

1201 N. Market Street, Suite 200

Wilmington, DE 19801

LLM Gateway
  • LLM Gateway
  • Request Routing
  • Usage Monitoring
  • Spend Management
  • Data Security
  • Access Controls
Features
  • Universal API Keys
  • BYOK
  • Spend Tracking
  • Token Allocation
  • Usage Analytics
  • Request Logs
  • Alerts
  • Data Redaction
  • Zero Data Retention
  • Audit Logs
Teams
  • AI Engineering
  • Engineering Leadership
  • Finance & Operations
  • Security & Compliance
Integrations
  • All Integrations
  • Migration Guides
Platform
  • Pricing
  • Blog
  • Model Fortress
  • Enterprise
  • Documentation
  • Status
Legal
  • Privacy Policy
  • Terms of Service
  • Data Processing Addendum
  • Acceptable Use Policy
  • Trust Center

LLM Gateway

  • LLM Gateway
  • Request Routing
  • Usage Monitoring
  • Spend Management
  • Data Security
  • Access Controls

Features

  • Universal API Keys
  • BYOK
  • Spend Tracking
  • Token Allocation
  • Usage Analytics
  • Request Logs
  • Alerts
  • Data Redaction
  • Zero Data Retention
  • Audit Logs

Teams

  • AI Engineering
  • Engineering Leadership
  • Finance & Operations
  • Security & Compliance

Integrations

  • All Integrations
  • Migration Guides

Platform

  • Pricing
  • Blog
  • Model Fortress
  • Enterprise
  • Documentation
  • Status

Legal

  • Privacy Policy
  • Terms of Service
  • Data Processing Addendum
  • Acceptable Use Policy
  • Trust Center

Offices

New York

130 E 59th St, 17th floor

New York, NY 10022

Wilmington

1201 N. Market Street, Suite 200

Wilmington, DE 19801

AICPA SOC for Service OrganizationsSOC 2 Type II

© 2026 Concentrate AI. All rights reserved.