DeepSeek-V4.1-Flash is a fast and efficient multimodal AI model with advanced reasoning, coding, and agent capabilities. Supporting up to 1M tokens of context and native vision understanding, it delivers frontier-level intelligence with lower inference costs, making it ideal for AI agents, automation, coding assistants, and scalable applications.
Input$0.2857 / M tokens
Output$1.1429 / M tokens
Cache Read$0.0057 / M tokens
GPT-6 Astra is OpenAI’s next-generation flagship model, designed for complex reasoning, software engineering, AI agents, research, and demanding end-to-end professional workflows. It excels at multi-step tasks across code, browsers, and professional software, delivering stronger performance in problem solving, tool use, computer use, and long-context understanding.
GPT-6 Astra supports a context window of up to 1.05M tokens and a maximum output of 128K tokens, with multiple reasoning effort levels including Low, Medium, High, XHigh, and Max, allowing developers to balance reasoning capability, efficiency, and cost based on task complexity.
It is well suited for advanced AI agents, complex software development, automated workflows, deep research, long-document processing, data analysis, and enterprise AI applications.
Input$10 / M tokens
Output$50 / M tokens
Cache Read$1 / M tokens
Claude Fable 5 is Anthropic’s Mythos-class model designed for demanding reasoning, software engineering, knowledge work, visual analysis, and scientific research. Built for long-horizon agentic tasks, it can sustain planning and execution across complex workflows, analyze large codebases and documents, and refine its work over multiple steps. Claude Fable 5 is well suited to production applications that require advanced coding, research, professional analysis, and reliable autonomous task completion.
Input$10 / M tokens
Output$50 / M tokens
Cache Read$0.25 / M tokens