Best Input Price
$1.05 / 1M
Fast, low-cost multimodal model for understanding text, images, audio, video, and PDFs, with tool calling and a 1M-token context window.
Context Window
1,000,000 tokens
Reasoning
Supported
Tool Calling
Supported
Released
2024-05-15