LLM Reference

GLM-Realtime Models by Zhipu AI

Zhipu AIProprietaryAudioMultimodal
1 model

Last refreshed 2026-07-11. Next refresh: weekly.

Details

ResearcherZhipu AI
LicenseProprietary
Commercial useCommercial use: conditional
Models1

Capabilities

VisionAll models
MultimodalAll models

Links

Website

About

GLM-Realtime is Zhipu AI's real-time voice and multimodal family for live audio, video, and text interaction over WebSocket and SDK APIs.

Decision facts

Best fit
audiomultimodalrealtime voice
Capability starting point
GLM-Realtime with multimodal inputs
Lowest tracked input
Not tracked
Closest related family
GLM-5

Current Variants

Use-when guidance is based on each model's tracked capabilities, context window, release date, and replacement status.

1 in view

Use when the workload needs realtime voice, multimodal inputs, and audio.

Unknown releaserealtime voicemultimodal inputsaudio

Release Timeline

1 release group
Unknown release
1 current
GLM-Realtime
realtime voicemultimodal inputsaudio
Current

Specifications(1 models)

GLM-Realtime model specifications comparison
ModelReleasedVisionMultimodal
GLM-RealtimeYesYes

Available From(1 provider)