OpenAI
Available in Shortly AI
GPT-4.1 nano
Released April 14, 2025
Vision (Image)Tool Use
GPT-4.1 nano is the fastest, most cost-effective GPT 4.1 model.
- Context window
- 1M tokens
- Maximum output
- 32.8K tokens
- Input pricing
- $0.10/M
- Output pricing
- $0.40/M
Independent benchmark
LiveBench historical snapshot
46.6overall
Historical benchmark data from LiveBench, release 2025-04-25. Scores are pinned and are not fetched on page load.
35.6
63.2
42.4
49.8
31.0
57.5
| Category | Score |
|---|---|
| Reasoning | 35.583 |
| Coding | 63.212 |
| Mathematics | 42.391 |
| Data Analysis | 49.820 |
| Language | 30.958 |
| Instruction Following | 57.537 |
Benchmarks measure selected evaluation tasks and do not guarantee performance for every prompt or workflow.
FAQ
Similar AI models
GLM 5.3 FlashZ.aiGLM-5.3-Flash is Z.ai’s native multimodal coding model, featuring 320B total parameters, 18B activated parameters, and a 1M-token context window. Its efficient hybrid attention architecture supports visual coding, tool use, and end-to-end professional workflows across code, browsers, documents, and graphical interfaces.DeepSeek V4 Flash Vision ExpDeepSeekDeepSeek-V4-Flash-Vision-Exp is the first experimental multimodal model in the DeepSeek-V4 family. It builds on DeepSeek-V4-Flash by adding visual modules and continued training for visual understanding, with substantially improved multimodal agent capabilities while remaining comparable on text-only agent tasks.Nemotron 3.5 Lightning 30BNVIDIANVIDIA-Nemotron-3.5-Lightning-30B-A3B-NVFP4 is a large language model (LLM) trained by NVIDIA.
The model employs a hybrid Mixture-of-Experts architecture, utilizing interleaved Mamba-2 and MoE layers, along with select Attention layers. The Lightning 3.5 model is released alongside a number of speculative decoding methods for faster text generation. The model has 3B active parameters and 30B parameters in total.
Available with one subscription
Start with GPT-4.1 nano in Shortly AI
Keep your conversations, files and model choices together in one workspace.