
AUTO
CometAPI Auto API ist eine intelligente Modell-Routing-Funktion, die es Entwicklern ermöglicht, auf das passende KI-Modell zuzugreifen, ohne für jede Anfrage eine spezifische Modell-ID anzugeben.
Durchsuchen und vergleichen Sie Text-, Bild-, Video- und Audiomodelle führender Anbieter. Funktionen und Preise prüfen und mit einem einzigen API-Schlüssel integrieren.

CometAPI Auto API ist eine intelligente Modell-Routing-Funktion, die es Entwicklern ermöglicht, auf das passende KI-Modell zuzugreifen, ohne für jede Anfrage eine spezifische Modell-ID anzugeben.

GPT-6 Sol-API entdecken.

Claude Opus 5.2 may be undergoing limited testing in Claude Code. Here is what the leaks claim, what remains unverified, and how CometAPI access would work after release.

GPT-6 Astra, the flagship model for complex reasoning and coding. Choose GPT-5.6 Terra to balance intelligence and cost,

Gemini 4-API entdecken.

Gemini 3.8 Flash is a new-generation lightweight Gemini model, with the goal of achieving a better balance among speed, cost, and coding capability.

claude-fable-5-1 Intelligenz der nächsten Generation für dauerhaft laufende Agenten

From fast verification to cinematic output, billing is based on the generated video duration, selecting the appropriate resolution, and turning every generation into a planned creative investment activity rule

OpenAIs leistungsfähigstes Bildgenerierungsmodell mit nahezu perfekter Textdarstellung in mehreren Sprachen, bis zu 4K-Auflösung und Reasoning-gestütztem Thinking Mode. Entwickelt für Produktions-Workflows, die Genauigkeit, Geschwindigkeit und markenkonforme visuelle Ergebnisse verlangen.

GLM-5.3-Flash is the first native multimodal model in the GLM-5 series, delivering stronger intelligence than GLM-5.2 while maintaining an exceptionally cost-efficient architecture.

FLUX 3 - Real World Models: Towards Multimodal Flow Models as the Backbone of Visual Intelligence.

gpt-oss-20b(free)-API entdecken.

Access to Claude Fable 5 has been restored. It brings 5th-generation intelligence to your most ambitious coding and professional work.

Das Modell Gemini 3.1 Flash Lite Image ist ein Effizienzspezialist in der Modellfamilie für Bildgenerierung und wurde für ultraniedrige Latenz sowie kosteneffiziente Bildgenerierung und -bearbeitung entwickelt.

HappyHorse 1.1 ist ein multimodales Modell zur Videogenerierung, das für professionelle Inhaltserstellung, Werbung, Kurzfilme, Social-Media-Produktion und Storytelling konzipiert ist. Es erweitert die Fähigkeiten von HappyHorse 1.0 — das nach hohen Platzierungen in unabhängigen Evaluierungen zur Videogenerierung große Aufmerksamkeit erlangte — um eine stärkere Szenenkohärenz und eine verbesserte Bildtreue.

Claude Opus 4.8 is a premium AI model designed for advanced reasoning, deep analysis, and high-quality content generation. It excels at handling complex instructions, long-context understanding, and sophisticated problem-solving across professional and technical domains.

GPT-5.4 Mini is a lightweight and efficient AI model optimized for speed and everyday productivity. It provides reliable conversational capabilities, content generation, and task assistance while maintaining low latency and resource usage.

GPT Image 2.5(sunburst)-API entdecken.

Claude Sonnet 5 API is live on CometAPI at $1.6 per million input tokens and $8 per million output tokens, 20 percent below Anthropic list price. One key gives you Claude Sonnet 5 plus 500+ models from OpenAI, Google, and ByteDance under all in one pricing and a single invoice. No seat fees. No monthly minimum. Pay only for what you use. Start free and make your first call in under five minutes.

Gemini 3.5 Flash is a high-speed AI model designed for fast response and efficient coding performance. It delivers significantly improved generation speed while maintaining strong reasoning ability, making it suitable for real-time applications and developer workflows.

Seedance 2.0 ist das multimodale Video-Grundlagenmodell der nächsten Generation von ByteDance, das auf filmische, narrative Videogenerierung mit mehreren Einstellungen ausgerichtet ist. Im Gegensatz zu Text-zu-Video-Demos mit nur einer Einstellung legt Seedance 2.0 den Schwerpunkt auf referenzbasierte Steuerung (Bilder, kurze Clips, Audio), durchgängige Konsistenz von Charakteren und Stil über mehrere Einstellungen hinweg sowie native Audio-/Video-Synchronisierung — mit dem Ziel, KI-Video für professionelle Kreativ- und Previsualisierungs-Workflows nutzbar zu machen.

GLM-5.3 is a next-generation open-source large language model optimized for complex coding, long-horizon tasks, and cybersecurity scenarios. It delivers significantly improved coding capabilities and agent performance compared with GLM-5.2.
GPT-Realtime-2 is our most capable realtime voice model. It supports speech-to-speech interactions with configurable reasoning effort, stronger instruction following, and more reliable tool use for complex voice-agent workflows

GPT-5.4 Nano is an ultra-lightweight AI model built for maximum speed and efficiency. It is optimized for simple tasks, real-time interactions, and large-scale deployments where low latency and minimal resource consumption are essential.

seedance-2-0

3.7 Flash delivers substantial improvements across software engineering, knowledge work, and web development workflows — with an introductory price of half the original 3.6 Flash cost per million tokens.

The best voice model for audio in, audio out.

OpenAIs leistungsfähigstes Modell, entwickelt für die anspruchsvollsten Aufgaben und langandauernde agentische Workflows. GPT-5.5 Pro überzeugt bei komplexem Programmieren, der Computernutzung, tiefgehender Recherche, der Datenanalyse und wissenschaftlichem Schlussfolgern — und liefert Intelligenz auf Spitzenniveau bei GPT-5.4-Latenz mit höherer Token-Effizienz. Ideal für Unternehmens- und professionelle Anwendungsfälle, die höchste Präzision und autonome Aufgabenausführung verlangen.

MiniMax-M2.7 offers the same top-tier intelligence as the standard version—including recursive self-evolution and expert-level office productivity—but is designed for applications requiring sub-second latency and high-speed token generation. Leveraging an enhanced inference backbone architecture, its output speed is 66% faster than the standard model (reaching 100 tps). It is the preferred choice for interactive programming assistants, real-time agent loop execution, and high-throughput enterprise pipelines with stringent completion time requirements.

GPT-5.3-Codex is optimized for agentic coding tasks in Codex or similar environments. GPT-5.3-Codex supports low, medium, high, and xhigh reasoning effort settings.

DeepSeek V4 Flash is an efficiency-optimized Mixture-of-Experts model from DeepSeek with 284B total parameters and 13B activated parameters, supporting a 1M-token context window. It is designed for fast inference and high-throughput workloads, while maintaining strong reasoning and coding performance.

GPT Image 2 ALL is a comprehensive image generation model designed to handle a wide range of creative and professional visual tasks. It combines high-quality image creation, advanced prompt understanding, and versatile style support to deliver exceptional results across diverse use cases.