
AUTO
CometAPI Auto API é um recurso inteligente de roteamento de modelos que permite aos desenvolvedores acessar o modelo de IA adequado sem precisar especificar um ID de modelo para cada solicitação.
Pesquise e compare modelos de texto, imagem, vídeo e áudio dos principais provedores. Veja recursos e preços, depois integre com uma única chave de API.

CometAPI Auto API é um recurso inteligente de roteamento de modelos que permite aos desenvolvedores acessar o modelo de IA adequado sem precisar especificar um ID de modelo para cada solicitação.

Explore a API de GPT-6 Sol.

Claude Opus 5.2 may be undergoing limited testing in Claude Code. Here is what the leaks claim, what remains unverified, and how CometAPI access would work after release.

GPT-6 Astra, the flagship model for complex reasoning and coding. Choose GPT-5.6 Terra to balance intelligence and cost,

Explore a API de Gemini 4.

Gemini 3.8 Flash is a new-generation lightweight Gemini model, with the goal of achieving a better balance among speed, cost, and coding capability.

claude-fable-5-1 Inteligência de próxima geração para agentes de longa duração

From fast verification to cinematic output, billing is based on the generated video duration, selecting the appropriate resolution, and turning every generation into a planned creative investment activity rule

O modelo de geração de imagens mais avançado da OpenAI, com renderização de texto quase perfeita em vários idiomas, resolução de até 4K e o Thinking Mode potencializado por raciocínio. Projetado para fluxos de trabalho de produção que exigem precisão, velocidade e resultados visuais alinhados à marca.

GLM-5.3-Flash is the first native multimodal model in the GLM-5 series, delivering stronger intelligence than GLM-5.2 while maintaining an exceptionally cost-efficient architecture.

FLUX 3 - Real World Models: Towards Multimodal Flow Models as the Backbone of Visual Intelligence.

Explore a API de gpt-oss-20b(free).

Access to Claude Fable 5 has been restored. It brings 5th-generation intelligence to your most ambitious coding and professional work.

O modelo Gemini 3.1 Flash Lite Image é um especialista em eficiência na família de geração de imagens, projetado para gerar e modificar imagens com latência ultrabaixa e excelente relação custo-benefício.

HappyHorse 1.1 é um modelo multimodal de geração de vídeo projetado para criação de conteúdo profissional, publicidade, curtas-metragens, produção para redes sociais e narrativa. Ele amplia as capacidades do HappyHorse 1.0 — que recebeu grande atenção após ficar bem classificado em avaliações independentes de geração de vídeo — com maior coerência entre cenas e fidelidade visual aprimorada.

Claude Opus 4.8 is a premium AI model designed for advanced reasoning, deep analysis, and high-quality content generation. It excels at handling complex instructions, long-context understanding, and sophisticated problem-solving across professional and technical domains.

GPT-5.4 Mini is a lightweight and efficient AI model optimized for speed and everyday productivity. It provides reliable conversational capabilities, content generation, and task assistance while maintaining low latency and resource usage.

Explore a API de GPT Image 2.5(sunburst).

Claude Sonnet 5 API is live on CometAPI at $1.6 per million input tokens and $8 per million output tokens, 20 percent below Anthropic list price. One key gives you Claude Sonnet 5 plus 500+ models from OpenAI, Google, and ByteDance under all in one pricing and a single invoice. No seat fees. No monthly minimum. Pay only for what you use. Start free and make your first call in under five minutes.

Gemini 3.5 Flash is a high-speed AI model designed for fast response and efficient coding performance. It delivers significantly improved generation speed while maintaining strong reasoning ability, making it suitable for real-time applications and developer workflows.

Seedance 2.0 é o modelo base multimodal de vídeo de próxima geração da ByteDance, focado na geração de vídeos narrativos cinematográficos de múltiplos planos. Diferentemente de demonstrações de texto para vídeo de plano único, Seedance 2.0 enfatiza o controle baseado em referências (imagens, clipes curtos, áudio), a consistência de personagens/estilo entre os planos e a sincronização nativa de áudio/vídeo — visando tornar o vídeo com IA útil para fluxos de trabalho criativos profissionais e de pré-visualização.

GLM-5.3 is a next-generation open-source large language model optimized for complex coding, long-horizon tasks, and cybersecurity scenarios. It delivers significantly improved coding capabilities and agent performance compared with GLM-5.2.
GPT-Realtime-2 is our most capable realtime voice model. It supports speech-to-speech interactions with configurable reasoning effort, stronger instruction following, and more reliable tool use for complex voice-agent workflows

GPT-5.4 Nano is an ultra-lightweight AI model built for maximum speed and efficiency. It is optimized for simple tasks, real-time interactions, and large-scale deployments where low latency and minimal resource consumption are essential.

seedance-2-0

3.7 Flash delivers substantial improvements across software engineering, knowledge work, and web development workflows — with an introductory price of half the original 3.6 Flash cost per million tokens.

The best voice model for audio in, audio out.

O modelo mais capaz da OpenAI, projetado para as tarefas mais difíceis e para fluxos de trabalho orientados por agentes de longa duração. O GPT-5.5 Pro se destaca em programação complexa, uso do computador, pesquisa aprofundada, análise de dados e raciocínio científico — oferecendo inteligência de ponta com latência de GPT-5.4 e maior eficiência de tokens. Ideal para casos de uso empresariais e profissionais que exigem o mais alto padrão de precisão e execução autônoma de tarefas.

MiniMax-M2.7 offers the same top-tier intelligence as the standard version—including recursive self-evolution and expert-level office productivity—but is designed for applications requiring sub-second latency and high-speed token generation. Leveraging an enhanced inference backbone architecture, its output speed is 66% faster than the standard model (reaching 100 tps). It is the preferred choice for interactive programming assistants, real-time agent loop execution, and high-throughput enterprise pipelines with stringent completion time requirements.

GPT-5.3-Codex is optimized for agentic coding tasks in Codex or similar environments. GPT-5.3-Codex supports low, medium, high, and xhigh reasoning effort settings.

DeepSeek V4 Flash is an efficiency-optimized Mixture-of-Experts model from DeepSeek with 284B total parameters and 13B activated parameters, supporting a 1M-token context window. It is designed for fast inference and high-throughput workloads, while maintaining strong reasoning and coding performance.

GPT Image 2 ALL is a comprehensive image generation model designed to handle a wide range of creative and professional visual tasks. It combines high-quality image creation, advanced prompt understanding, and versatile style support to deliver exceptional results across diverse use cases.