
AUTO
CometAPI Auto API là một tính năng định tuyến mô hình thông minh cho phép các nhà phát triển truy cập mô hình AI phù hợp mà không cần chỉ định ID mô hình cụ thể cho từng yêu cầu.
Tìm kiếm và so sánh các mô hình văn bản, hình ảnh, video và âm thanh từ các nhà cung cấp hàng đầu. Xem tính năng và giá cả, sau đó tích hợp chỉ với một API key.

CometAPI Auto API là một tính năng định tuyến mô hình thông minh cho phép các nhà phát triển truy cập mô hình AI phù hợp mà không cần chỉ định ID mô hình cụ thể cho từng yêu cầu.

Khám phá API của GPT-6 Sol.

Claude Opus 5.2 may be undergoing limited testing in Claude Code. Here is what the leaks claim, what remains unverified, and how CometAPI access would work after release.

GPT-6 Astra, the flagship model for complex reasoning and coding. Choose GPT-5.6 Terra to balance intelligence and cost,

Khám phá API của Gemini 4.

Gemini 3.8 Flash is a new-generation lightweight Gemini model, with the goal of achieving a better balance among speed, cost, and coding capability.

claude-fable-5-1 Trí tuệ thế hệ tiếp theo cho các tác nhân vận hành trong thời gian dài

From fast verification to cinematic output, billing is based on the generated video duration, selecting the appropriate resolution, and turning every generation into a planned creative investment activity rule

Mô hình tạo hình ảnh mạnh mẽ nhất của OpenAI, với khả năng kết xuất văn bản gần như hoàn hảo trong nhiều ngôn ngữ, độ phân giải lên đến 4K và Thinking Mode được hỗ trợ bởi khả năng suy luận. Được xây dựng cho các quy trình sản xuất đòi hỏi độ chính xác, tốc độ và đầu ra hình ảnh đúng nhận diện thương hiệu.

GLM-5.3-Flash is the first native multimodal model in the GLM-5 series, delivering stronger intelligence than GLM-5.2 while maintaining an exceptionally cost-efficient architecture.

FLUX 3 - Real World Models: Towards Multimodal Flow Models as the Backbone of Visual Intelligence.

Khám phá API của gpt-oss-20b(free).

Access to Claude Fable 5 has been restored. It brings 5th-generation intelligence to your most ambitious coding and professional work.

Mô hình Gemini 3.1 Flash Lite Image là một chuyên gia về hiệu suất trong dòng mô hình tạo sinh hình ảnh, được thiết kế cho độ trễ siêu thấp và cho việc tạo sinh, chỉnh sửa hình ảnh hiệu quả về chi phí.

HappyHorse 1.1 là một mô hình tạo video đa mô thức, được thiết kế cho sáng tạo nội dung chuyên nghiệp, quảng cáo, phim ngắn, sản xuất nội dung trên mạng xã hội và kể chuyện. Mô hình này mở rộng các khả năng của HappyHorse 1.0 — vốn đã thu hút sự chú ý đáng kể sau khi đạt thứ hạng cao trong các đánh giá độc lập về tạo video — với tính liền mạch giữa các cảnh tốt hơn và độ trung thực hình ảnh được cải thiện.

Claude Opus 4.8 is a premium AI model designed for advanced reasoning, deep analysis, and high-quality content generation. It excels at handling complex instructions, long-context understanding, and sophisticated problem-solving across professional and technical domains.

GPT-5.4 Mini is a lightweight and efficient AI model optimized for speed and everyday productivity. It provides reliable conversational capabilities, content generation, and task assistance while maintaining low latency and resource usage.

Khám phá API của GPT Image 2.5(sunburst).

Claude Sonnet 5 API is live on CometAPI at $1.6 per million input tokens and $8 per million output tokens, 20 percent below Anthropic list price. One key gives you Claude Sonnet 5 plus 500+ models from OpenAI, Google, and ByteDance under all in one pricing and a single invoice. No seat fees. No monthly minimum. Pay only for what you use. Start free and make your first call in under five minutes.

Gemini 3.5 Flash is a high-speed AI model designed for fast response and efficient coding performance. It delivers significantly improved generation speed while maintaining strong reasoning ability, making it suitable for real-time applications and developer workflows.

Seedance 2.0 là mô hình nền tảng video đa phương thức thế hệ mới của ByteDance, tập trung vào việc tạo sinh video mang tính điện ảnh, theo mạch truyện nhiều cảnh. Khác với các bản demo chuyển văn bản thành video một cảnh, Seedance 2.0 nhấn mạnh điều khiển dựa trên tham chiếu (hình ảnh, clip ngắn, âm thanh), duy trì nhân vật và phong cách nhất quán, liền mạch giữa các cảnh, cùng đồng bộ âm thanh–hình ảnh tích hợp sẵn — nhằm biến video AI trở nên hữu ích cho các quy trình sáng tạo chuyên nghiệp và tiền trực quan hóa.

GLM-5.3 is a next-generation open-source large language model optimized for complex coding, long-horizon tasks, and cybersecurity scenarios. It delivers significantly improved coding capabilities and agent performance compared with GLM-5.2.
GPT-Realtime-2 is our most capable realtime voice model. It supports speech-to-speech interactions with configurable reasoning effort, stronger instruction following, and more reliable tool use for complex voice-agent workflows

GPT-5.4 Nano is an ultra-lightweight AI model built for maximum speed and efficiency. It is optimized for simple tasks, real-time interactions, and large-scale deployments where low latency and minimal resource consumption are essential.

seedance-2-0

3.7 Flash delivers substantial improvements across software engineering, knowledge work, and web development workflows — with an introductory price of half the original 3.6 Flash cost per million tokens.

The best voice model for audio in, audio out.

Mô hình mạnh nhất của OpenAI, được thiết kế cho các tác vụ khó nhất và các quy trình tác nhân tự chủ kéo dài. GPT-5.5 Pro xuất sắc trong lập trình phức tạp, sử dụng máy tính, nghiên cứu chuyên sâu, phân tích dữ liệu và suy luận khoa học — mang lại trí tuệ đẳng cấp tiên phong với độ trễ tương đương GPT-5.4 cùng hiệu quả sử dụng token cao hơn. Lý tưởng cho các trường hợp sử dụng cấp doanh nghiệp và chuyên nghiệp đòi hỏi tiêu chuẩn cao nhất về độ chính xác và khả năng thực thi tác vụ tự chủ.

MiniMax-M2.7 offers the same top-tier intelligence as the standard version—including recursive self-evolution and expert-level office productivity—but is designed for applications requiring sub-second latency and high-speed token generation. Leveraging an enhanced inference backbone architecture, its output speed is 66% faster than the standard model (reaching 100 tps). It is the preferred choice for interactive programming assistants, real-time agent loop execution, and high-throughput enterprise pipelines with stringent completion time requirements.

GPT-5.3-Codex is optimized for agentic coding tasks in Codex or similar environments. GPT-5.3-Codex supports low, medium, high, and xhigh reasoning effort settings.

DeepSeek V4 Flash is an efficiency-optimized Mixture-of-Experts model from DeepSeek with 284B total parameters and 13B activated parameters, supporting a 1M-token context window. It is designed for fast inference and high-throughput workloads, while maintaining strong reasoning and coding performance.

GPT Image 2 ALL is a comprehensive image generation model designed to handle a wide range of creative and professional visual tasks. It combines high-quality image creation, advanced prompt understanding, and versatile style support to deliver exceptional results across diverse use cases.