Meta
Available in Shortly AI
Llama 3.3 70B Instruct
Released December 6, 2024
Tool Use
Where performance meets efficiency. This model supports high-performance conversational AI designed for content creation, enterprise applications, and research, offering advanced language understanding capabilities, including text summarization, classification, sentiment analysis, and code generation.
- Context window
- 128K tokens
- Maximum output
- 8.2K tokens
- Input pricing
- $0.72/M
- Output pricing
- $0.72/M
FAQ
Similar AI models
Inkling SmallThinkingmachinesInkling-Small is a lighter-weight model with 12B active parameters, trained with a similar recipe, to Inkling that achieves strong performance with even lower cost and latency.Gemini 3.5 Flash LiteGoogleGemini 3.5 Flash Lite features upgraded agentic capabilities, making the model ideal for subagents in complex workflows.Gemini 3.6 FlashGoogleGemini 3.6 Flash delivers higher quality across coding, agentic workflows, and web development with reduced token consumption and fewer model calls compared to previous model iterations.
Available with one subscription
Start with Llama 3.3 70B Instruct in Shortly AI
Keep your conversations, files and model choices together in one workspace.