Z.ai
Available in Shortly AI
GLM 4.7 Flash
Released January 19, 2026
ReasoningTool Use
GLM-4.7-Flash balances high performance with efficiency, making it the perfect lightweight deployment option. Beyond coding, it is also recommended for creative writing, translation, long-context tasks, and roleplay.
- Context window
- 200K tokens
- Maximum output
- 131K tokens
- Input pricing
- $0.07/M
- Output pricing
- $0.40/M
FAQ
Similar AI models
GPT 5.6 LunaOpenAIGPT-5.6 Luna is a fast, affordable GPT-5.6 model that brings strong capability at the lowest cost in the series.Hy3TencentTencent Hy3 is an open-source Mixture-of-Experts (MoE) large language model developed by Tencent's Hunyuan team. It has 295B total parameters, 21B active parameters, a 3.8B-parameter MTP layer, and a 256K context window. Hy3 is designed for agentic workflows, coding, reasoning, long-context understanding, tool-use scenarios, document analysis, and structured generation.
The model improves on Hy3 Preview with stronger tool-call reliability, better output-format stability, improved long-context retention, and lower hallucination rates in Tencent's internal evaluations.DeepSeek V4 FlashDeepSeekDeepseek V4 Flash is an AI model available through Shortly AI.
Available with one subscription
Start with GLM 4.7 Flash in Shortly AI
Keep your conversations, files and model choices together in one workspace.