kimi-k2.7-code
- kimi-k2.7-code: Kimi's most intelligent coding model to date, reliably follows instructions in long contexts and completes programming tasks with a higher success rate.
Developer Documentation
🌟 2026-06-10
🔹 claude-fable-5
- claude-fable-5:Anthropic's most capable, widely released model, for the most demanding reasoning and long-horizon agentic work
📚 Developer Documentation
🔹 qwen3.7-plus
- qwen3.7-plus: Alibaba Cloud's high-performance large language model, supporting up to 128K-token long-context understanding, function calling, multilingual tasks, complex reasoning, coding, and instruction-following scenarios.
📚 Developer Documentation
⚠️ Model Deprecation Notice
Impact Time: 2026-09-08, subject to the actual change time.
To continuously improve service quality and optimize underlying model resources, CometAPI will delist the following Qwen models on 2026-09-08. Please switch to the recommended replacement models in advance.
|
| qwen3.6-max-preview | qwen3.7-max |
| qwen3-max-preview | qwen3.7-max |
| qwen3-max | qwen3.7-max |
| qwen3-coder-plus | qwen3.7-plus |
✨ Feature Update
CometAPI now supports several new video generation models, including wan2.6, wan2.7, happyhorse-1.0, viduq3-turbo, and viduq3. These models support the /v1/videos API format. Pricing varies by model and resolution and is billed by generated video duration, in USD / second.
|
wan2.6 / wan2.7 series | 720p, 1080p | Billed per second |
happyhorse-1.0 | 720p, 1080p | Billed per second |
viduq3 / viduq3-turbo series | 360p, 540p, 720p, 1080p | Billed per second |
✨ Model Update
CometAPI now supports the minimax-m3 model.
📚 Developer Documentation
✨ Adjustment Details
Due to ongoing resource stability challenges with the veo3 series, Comet will adjust the resource channels and billing rules for the veo3 series.
Previously, the veo3 series was billed per generation. It will now be adjusted to official same-price per-second billing, with the website’s default 20% discount applied.
|
veo3.1 | 720p | $0.4 / second |
veo3.1 | 1080p | $0.4 / second |
veo3.1 | 4K | $0.6 / second |
veo3.1-fast | 720p | $0.1 / second |
veo3.1-fast | 1080p | $0.12 / second |
veo3.1-fast | 4K | $0.3 / second |
For example, when generating a 1080p video with veo3.1, the price after the 20% discount is:
$0.4 × 0.8 = $0.32 / second
For detailed pricing and available groups, please refer to the model details page: https://www.cometapi.com/models/
🌟 2026-05-29
🔹 claude-opus-4-8
- claude-opus-4-8: Most intelligent model for agents and coding
📚 Developer Documentation
CometAPI now supports Midjourney V8. When using the Midjourney image generation API, you can enable V8 by adding the --v 8 parameter directly in the prompt.
For more details about the Midjourney image generation API, please refer to the documentation:
https://apidoc.cometapi.com/api/image/midjourney/imagine
Following official model updates and iterations, Comet will retire the following models. Please switch to the latest available models as soon as possible.
Details page: https://www.cometapi.com/models/
May 29, 2026 — 12:00 PM UTC
|
glm-4.x series |
kimi-k2 series |
minimax-m2.1 |
minimax-m2 |
o1-mini |
gpt-3.5-turbo |
gemini-3.1-flash-lite-preview |
gemini-3-pro-preview |
gemini-2.5-flash-preview-09-25 |
gemini-2.5-flash-lite-preview-09-2025 |
gemini-2.5-flash-image-preview |
June 15, 2026
|
claude-opus-4 |
claude-sonnet-4-20250514 |
July 23, 2026
|
gpt-5-codex |
gpt-5.1-codex |
gpt-5.2-codex |
gpt-5-chat-latest |
gpt-5.1-chat-latest |
gpt-5.2-chat-latest |
gpt-4o-realtime |
gpt-realtime-mini |
gpt-audio-mini |
gpt-4o-mini-search-preview |
October 23, 2026
|
o4-mini |
gpt-4.1-nano |
o1-pro |
o3-mini |
🌟 2026-05-22
🔹 qwen3.7-max
- qwen3.7-max: Qwen3.7-Max's core strength lies in the breadth and depth of its agentic capabilities, excelling at tool use, task planning, and complex instruction execution.
📚 Supported Endpoints
🌟 2026-05-20
🔹 gemini-3.5-flash
- gemini-3.5-flash: Google's most intelligent model, built for speed, combining frontier intelligence with outstanding search and factual grounding.
📚 Supported Endpoints
📅 2026-05-06
✨ Change Details
- New model:
grok-4.3 is now live — excels at autonomous reasoning, knowledge work, and tool use. Ideal for complex agent workflows and deep analysis tasks.
🍡 Recommended Models
📚 Developer Documentation
✨ Pricing Changes
- The
grok-4.2 series pricing has been reduced to match grok-4.3 pricing.
- Changes take effect immediately. No action required on your end.
✨ Change Details
- Deprecating soon: The following models will be retired on May 15, 2026 at 12:00 PM Pacific Time (PT). Please refer to the migration table below:
|
grok-4-1-fast-reasoning | grok-4.3 |
grok-4-1-fast-non-reasoning | grok-4.20-non-reasoning |
grok-4-fast-reasoning | grok-4.3 |
grok-4-fast-non-reasoning | grok-4.20-non-reasoning |
grok-4-0709 | grok-4.3 |
grok-code-fast-1 | grok-4.3 |
- Migration guide: Reasoning models →
grok-4.3; Non-reasoning models → grok-4.20-non-reasoning.
📅 2026-04-29
✨ Change Details
- Deprecating soon: Per the official announcement, the
deepseek-chat and deepseek-reasoner model families are being phased out and will cease API service.
- Migration guide: Please migrate to the new DeepSeek V4 series as soon as possible.
- General / high-throughput workloads →
deepseek-v4-flash
- Advanced reasoning / coding / long-horizon agent workflows →
deepseek-v4-pro
🍡 Recommended Models
- Model names:
deepseek-v4-flash, deepseek-v4-pro
📚 Developer Documentation
✨ Pricing Changes
|
doubao-seedream-4-5-251128 | $0.04 / image | $0.04 / image (unchanged) | $0.032 / image |
doubao-seedream-4-0-250828 | $0.03 / image | $0.04 / image | $0.032 / image |
doubao-seedream-5-0-260128 | $0.035 / image | $0.04 / image | $0.032 / image |
- List prices are now unified at $0.04 / image; discounted price unified at $0.032 / image (20% off).
- Pricing changes take effect immediately. No action required on your end.
🌟 2026-04-25
🔹 GPT-5.5
- Model:
gpt-5.5
- Details: A next-generation multimodal flagship model balancing exceptional performance with efficient response, dedicated to providing comprehensive and stable general-purpose AI services.
- ⚠️ Note: This model supports the standard Chat interaction format.
🔹 GPT-5.5-Pro
- Model:
gpt-5.5-pro
- Details: An advanced model engineered for extremely complex logic and professional demands, representing the highest standard of deep reasoning and precise analytical capabilities.
- ⚠️ Note: This model supports the Response interaction format only.
🌟 2026-04-24
🔹 GPT-5.5 Series
- Models:
gpt-5.5-all / gpt-5.5-medium-all / gpt-5.5-high-all / gpt-5.5-xhigh-all / gpt-5.5-low-all
- Details: Designed to cover varying levels of task complexity.
- ⚠️ Note: All models above support standard Chat and Response interaction formats.
🔹 GPT-Image-2 Series
- Model:
gpt-image-2-all (Multimodal and image generation model)
- Endpoints:
- 💬 Chat Completions (
/v1/chat/completions): Supports multimodal inputs (image-to-image) and complex instructions. ⚠️ Note: supports stream: true only. Reference Guide
- 🎨 Image Generations (
/v1/images/generations): Supports standard text-to-image generation. Reference Guide
🔹 DeepSeek V4
- deepseek-v4-pro: A 1.6T parameter MoE model supporting a 1M-token context. Designed for advanced reasoning, coding, and long-horizon agent workflows.
- deepseek-v4-flash: A 284B parameter efficiency-optimized MoE model supporting a 1M-token context. Designed for fast inference and high-throughput workloads.
📚 Developer Documentation
🌟 2026-04-21
🔹 gpt-image-2
- gpt-image-2: Adopts a new autoregressive multimodal architecture with a core breakthrough in near-perfect text rendering capabilities. It promises generation via natural language while preserving character, lighting, and scene context, capable of directly outputting 4K resolution commercial design materials.
📚 Supported Endpoints
This model supports following standard OpenAI format:
- Image Generations (
/v1/images/generations): Supports standard text-to-image generation.Reference Guide
🌟 2026-04-20
🔹 doubao-seedance-2-0 / doubao-seedance-2-0-fast
- doubao-seedance-2-0: ByteDance's latest high-quality video generation model, supporting both text-to-video & image-to-video.
- doubao-seedance-2-0-fast: The accelerated version of Seedance 2.0 — faster generation, same powerful quality.
⚠️ Important Notice
The legacy Seedance-series official API format is deprecated. All Seedance models (including doubao-seedance-1-5-pro and doubao-seedance-1-0-pro) should now use the unified v1/videos endpoint going forward.
Note: doubao-seedance-1-5-pro and doubao-seedance-1-0-pro do not support image-to-video (input_reference is not available for these models).
📋 Parameters
|
| Duration (seconds) | 4–15 seconds, default 5 |
| Aspect Ratio (size) | 21:9 / 16:9 / 4:3 / 1:1 / 3:4 / 9:16, default 16:9 |
| Resolution (resolution) | 480p / 720p / 1080p*, default 720p |
*1080p only available for doubao-seedance-1-5-pro and doubao-seedance-1-0-pro
🚀 Usage Example
curl --location --request POST 'https://api.cometapi.com/v1/videos' \
--header 'Authorization: sk-your-key' \
--header 'Content-Type: multipart/form-data' \
--form 'prompt="a cat running on the beach"' \
--form 'model="doubao-seedance-2-0"' \
--form 'seconds="5"' \
--form 'size="16:9"' \
--form 'resolution="720p"' \
--form 'input_reference=@"your_image.png"'
💡 input_reference is optional — include a reference image for image-to-video, or omit it for text-to-video. Only supported by doubao-seedance-2-0 and doubao-seedance-2-0-fast.
📖 Documentation
For full API details, please refer to: https://apidoc.cometapi.com/api/video/seedance/create
🌟 2026-04-16
🔹 claude-opus-4-7
- claude-opus-4-7: The smartest agentic and coding model.
📚 Developer Documentation
🔹 kimi-k2.6
- kimi-k2.6: Kimi K2.6 preview version is now open for integration testing.
🔹 qwen3.6-plus
- qwen3.6-plus: Newly launched, featuring comprehensively enhanced code development capabilities and synchronized improvements in multimodal recognition and reasoning efficiency, delivering an outstanding Vibe Coding experience.
🔹 glm-5.1
- glm-5.1: Zhipu's latest flagship model, with greatly enhanced coding capabilities and significantly improved performance in long-horizon tasks.
🌟 2026-03-27
- Superior Audio Quality: Significantly enhanced audio clarity, vocal performance, and mixing precision.
- Immersive Experience: Delivers lifelike vocals and powerful creative control.
- Professional Creation: Generates emotionally rich, genre-accurate, high-quality songs.
🛠 Usage
Set the request parameter mv to chirp-fenix.
{
"mv": "chirp-fenix",
"gpt_description_prompt": "cat"
}
🌟 2026-03-25
🔹 mimo-v2-flash , mimo-v2-omni , mimo-v2-pro
- Xiaomi MiMo-V2 Model Series: Towards the Agentic Era. Integrating trillion-scale parameters, omni-modal perception, and human-like interaction—unifying understanding and action, from the present into the future.
🌟 2026-03-18
🔹 gpt-5.4-mini
🔹 gpt-5.4-mini-2026-03-17
🔹 gpt-5.4-nano
🔹 gpt-5.4-nano-2026-03-17
-
gpt-5.4-mini, gpt-5.4-mini-2026-03-17: OpenAI's most powerful small model to date. It brings the capabilities of GPT-5.4 to a faster, more efficient architecture, designed specifically for coding, computer operations, and high-volume workloads.
-
gpt-5.4-nano, gpt-5.4-nano-2026-03-17: The most affordable GPT-5.4 class model, designed specifically for simple, massive-scale tasks where speed and cost are prioritized (such as classification, data extraction, and sorting).
🔹 glm-5-turbo
- glm-5-turbo: GLM-5-Turbo is a base model deeply optimized for OpenClaw Lobster scenarios, delivering excellent performance and precision in domain-specific tasks.
🔹 qwen3.5-122b-a10b
🔹 qwen3.5-27b
🔹 qwen3.5-35b-a3b
🔹 qwen3.5-flash
- qwen3.5 Series: The latest generation of model families from Alibaba Cloud Qwen (Tongyi Qianwen). The entire series features significant improvements in coding, mathematics, and logical reasoning capabilities. The Flash version offers ultimate inference speed, while the MoE architecture versions (122B-A10B/35B-A3B) significantly reduce computational overhead while maintaining flagship-level performance, perfectly balancing performance and cost.
📣 Log Retention Notice
To ensure service stability and manage storage costs, logs on this site will only be retained for 3 months. Expired logs will be automatically deleted and cannot be recovered!
If you need long-term retention, please download or back up your logs within 3 months!
- ⏰ Effective Date: Immediately
Thank you for your understanding and support!
🌟 2026-03-12
🔹 grok-4.20-multi-agent-beta-0309, grok-4.20-beta-0309-reasoning, grok-4.20-beta-0309-non-reasoning
-
grok-4.20-beta series: Grok 4.20 Beta is X.ai's newest flagship model with industry-leading speed and agentic tool calling capabilities. It combines the lowest hallucination rate on the market with strict prompt adherence, delivering consistently precise and truthful responses.
-
grok-4.20-multi-agent-beta-0309:
- grok-4.20-beta-0309-reasoning, grok-4.20-beta-0309-non-reasoning:
🌟 2026-03-06
🔹 gpt-5.4-pro-2026-03-05, gpt-5.4-2026-03-05, gpt-5.4-pro, gpt-5.4
- gpt-5.4-pro, gpt-5.4-pro-2026-03-05: GPT-5.4 Pro utilizes more powerful computing capabilities for deeper thinking to consistently deliver superior answers, designed to solve complex problems. GPT-5.4 Pro supports
reasoning.effort: medium, high, xhigh.
- gpt-5.4-2026-03-05, gpt-5.4: GPT-5.4 is a frontier model for handling complex professional tasks.
reasoning.effort supports the following options: none (default), low, medium, high, and xhigh.
🌟 2026-03-04
🔹 gpt-5.3-chat-latest
- gpt-5.3-chat-latest: This model not only provides more accurate answers but also delivers richer, more contextually relevant results. It focuses on the experience details users perceive most in daily use: tone, response relevance, and conversational flow.
📚 Developer Documentation
🌟 2026-03-04
🔹 gemini-3.1-flash-lite-preview
🔹 gemini-3.1-flash-lite
gemini-3.1-flash-lite is the most cost-effective model in the Gemini series, optimized for high-volume agent tasks, translation, and simple data processing.
📚 Developer Documentation
🌟 2026-02-28
- grok-imagine-video: The latest video generation model from xAI, supporting Text-to-Video and Video Editing tasks.
- Configurable: Supports custom
duration (e.g., 5s, 10s), aspect_ratio (e.g., 16:9, 1:1), and resolution (e.g., 720p) via simple parameters.
- Async API: The endpoint operates asynchronously. It returns a
request_id immediately upon submission; use the GET endpoint to check status and retrieve the generated video.
🌟 2026-02-27
🚀 Nano Banana 2 (Gemini 3.1) Flagship Image Model Released!
Designed for professional workflows, integrating reasoning with high-fidelity illustration.
💎 Available Models:
gemini-3.1-flash-image-preview
gemini-3.1-flash-image
🔥 Core Evolutions:
- Ultimate Quality: Native 4K (4096px) + Supports extreme 1:8 / 8:1 aspect ratios.
- Superior Consistency: Supports 14 Reference Images (10 objects + 4 characters), perfectly replicating styles and character consistency.
- Intelligent Reasoning: Built-in Thinking Process (Chain of Thought) to understand complex Prompts.
- All-around Enhancements: Advanced text rendering (poster-ready) + Google Web Search verification.
🔗 Integration Docs:
https://apidoc.cometapi.com/gemini-image-generation
📘 Usage Guide:
https://apidoc.cometapi.com/guide-nanobanana
⚠️ Important Migration Notice: Gemini 3 Pro Deprecation
Affected Models: gemini-3-pro-preview, gemini-3-pro-preview-thinking
Current Status: ❌ Deprecated
Shutdown Date: March 9, 2026
🚨 Action Required: To avoid service interruption, please migrate to Gemini 3.1 Pro Preview as soon as possible.
🌟 2026-02-25
🔹 gpt-5.3-codex,gpt-audio-1.5,gpt-realtime-1.5
- gpt-5.3-codex: GPT-5.3-Codex is the most powerful agent programming model to date. Optimized for agent programming tasks in Codex or similar environments. GPT-5.3-Codex supports reasoning parameters set to Low, Medium, High, and Ultra-High.
- gpt-audio-1.5: The best voice model for audio input and audio output in Chat Completions.
- gpt-realtime-1.5: The best voice model for audio input and audio output.
- This model follows the OpenAI Realtime API format.
🌟 2026-02-24
-
doubao-seedream-5-0-260128 - Doubao-Seedream-5.0-lite is ByteDance's latest image generation model. This model is the first to feature web retrieval capabilities, integrating real-time online information to enhance the timeliness of generated images. Additionally, the model's intelligence has been upgraded, enabling it to accurately parse complex prompts and visual content. Furthermore, it boasts enhancements in the breadth of world knowledge, reference consistency, and generation quality for professional scenarios, better satisfying enterprise-level visual creation needs.
- Model ID:
doubao-seedream-5-0-260128
📚 Developer Documentation
🌟 2026-02-19
🔹 Gemini 3.1 Series
- gemini-3.1-pro-preview, gemini-3.1-pro-preview-thinking: Gemini 3.1 Pro is the next generation in the Gemini series of models, a suite of highly-capable, natively multimodal, reasoning models. Gemini 3 Pro is now Google’s most advanced model for complex tasks, and can comprehend vast datasets, challenging problems from different information sources, including text, audio, images, video, and entire code repositories
📚 Developer Documentation
🌟 2026-02-18
✨ Core Features
- Most Powerful All-Round Model: Claude Sonnet 4.6 delivers a world-class experience in coding and logical reasoning.
- Dual-Protocol Support: Seamlessly compatible with both the OpenAI Standard Format and the Anthropic Native Format.
🔌 Call Parameters
- Model Names:
claude-sonnet-4-6, cometapi-sonnet-4-6
📚 Developer Documentation
🌟 2026-02-17
🔹 Qwen3.5 Series
- qwen3.5-397b-a17b: The Qwen3.5 series 397B-A17B native vision-language model is based on a hybrid architecture design, fusing linear attention mechanisms with sparse Mixture-of-Experts (MoE) models to achieve higher inference efficiency. In various tasks such as language understanding, logical reasoning, code generation, agent tasks, image understanding, video understanding, and graphical user interfaces (GUI), it demonstrates excellent performance comparable to current top-tier frontier models. It possesses powerful code generation and agent capabilities, with good generalization for various agent scenarios.
- qwen3.5-plus, qwen3.5-plus-2026-02-15, qwen3.5-plus-thinking: The Qwen3.5 native vision-language series Plus model is based on a hybrid architecture design, fusing linear attention mechanisms with sparse Mixture-of-Experts (MoE) models to achieve higher inference efficiency. In multiple task evaluations, the 3.5 series demonstrates excellent performance comparable to current top-tier frontier models, achieving a leap in performance in both pure text and multimodal aspects compared to the 3 series. This version is a snapshot from February 15, 2026.
📚 Developer Documentation
🌟 2026-02-14
🔹 Doubao Seed 2.0 Series
- doubao-seed-2-0-code-preview-260215
Focuses on long-chain reasoning capabilities and complex task stability, adapted for complex scenarios in real business environments. As the coding-enhanced version of Seed 2.0, it is better suited for Agentic Coding.
- doubao-seed-2-0-lite-260215
Balances generation quality with response speed, making it suitable as a general-purpose production-grade model.
- doubao-seed-2-0-mini-260215
Designed for low-latency, high-concurrency, and cost-sensitive scenarios. It emphasizes rapid response and flexible inference deployment, supporting four-level thinking and multimodal understanding capabilities.
📚 Developer Documentation
🌟 2026-02-13
🔹 minimax-m2.5
The world's first production-grade model natively designed for Agents. Its Coding & Agentic performance benchmarks directly against Claude Opus 4.6.
- Full-Stack Coding: Supports PC, App, and cross-platform application development.
- Office SOTA: Leads the industry in core productivity scenarios such as advanced Excel processing, in-depth research, and PPT generation.
📚 Developer Documentation
🌟 2026-02-12
🔹 glm-5
Zhipu's new generation flagship base model, built for Agentic Engineering. It provides reliable productivity in complex system engineering and long-horizon Agent tasks; the usage experience in real-world coding scenarios approaches Claude Opus 4.5.
📚 Developer Documentation
🌟 2026-02-06
✨ Core Features
- Ultimate Intelligence Model: Claude Opus 4.6 delivers world-class programming and logical reasoning experience.
- Dual Protocol Support: Perfectly compatible with OpenAI Standard Format and Anthropic Native Format.
🔌 Call Parameters
- Model Names:
claude-opus-4-6 ,cometapi-opus-4-6
📚 Developer Documentation
⚠️ 2026-02-05
✨ Change Details
- Upcoming Shutdown: In accordance with the official schedule,
chatgpt-4o-latest will be discontinued on Feb 17, 2026.
- Migration: Please migrate to the latest flagship GPT-5.2 Series. We recommend
gpt-5.2 for most use cases or gpt-5.2-chat-latest for the newest chat improvements.
🔌 Recommended Models
- Model Names:
gpt-5.2 , gpt-5.2-chat-latest
📚 Developer Documentation
⚠️ 2026-02-04
🔄 CometAPI: Doubao Model Update Notice
✨ Change Details
Legacy Deprecation: In compliance with official policy, the Doubao 1.5 / 1.6 Series have been discontinued.
Migration: Please switch to doubao-seed-1.8.
🔌 Recommended Model
Model Name: doubao-seed-1.8
📚 Developer Documentation
🌟 2026-01-28
🦌 Comet Update: Qwen3 Flagship / Kimi Long Context / OCR v2
🚀 New Models
-
qwen3-max-2026-01-23 (General Flagship)
-
The strongest snapshot of the Qwen3 series, introducing a Deep Reasoning module. Improves complex logic deduction and code refactoring capabilities by 40%. Ideal for research assistance and system-level instructions.
-
kimi-k2.5 (Long Context)
-
Kimi's smartest model to date. Built on a native multimodal architecture, supporting both vision and text inputs simultaneously.
-
deepseek-ocr-2 (Visual Extraction)
-
Specialized in handwriting and complex table restoration. Eliminates hallucinations in dense formulas and supports direct Markdown/JSON structured output.
-
👉 API Docs
🌟 2026-01-19
🚀 Available Models & Usage Guide
🔹 gpt-5.2-codex (For Professional Code Tasks)
- Model ID:
gpt-5.2-codex
- Description: Optimized for coding tasks like code generation, completion, and analysis to leverage its best-in-class coding capabilities.
- Required Endpoint:
/v1/responses (Note: This endpoint must be used for this model.)
- Documentation: 👉 Check out the Responses API documentation