mô hình
mô hình
wan2.6
wan2.6
veo3.1-lite
veo3.1-lite
Lightweight base variant of Veo3.1 with lower computing overhead, meets basic video generation demands for low-cost bulk preview scenarios.
doubao-seedance-2-0-fast
doubao-seedance-2-0-fast
Accelerated variant of Seedance 2.0. Balances generation speed and visual quality, reduces inference latency for high-concurrency testing and mass rapid content production.
doubao-seedance-2-0-mini
doubao-seedance-2-0-mini
Doubao Seedance 2.0 Mini is ByteDance’s lightweight text-to-video model under the Seed series. Optimized for fast inference and low GPU overhead, it generates smooth, style-consistent short videos from text prompts. Designed for API integration, test environments and low-cost batch video production.

GLM 5.1
TEST GLM-5.1 (released April 2026), purpose-built for long-horizon autonomous tasks. Unlike traditional models optimized for short interactions, GLM-5.1 excels at maintaining goal alignment, reducing strategy drift, and delivering production-grade results over extended periods — up to 8 hours of continuous autonomous work on a single complex task. It represents a major leap in agentic engineering, shifting evaluation from single-turn intelligence to real-world sustained execution.
gpt-5.6
gpt-5.6
Nano Banana 2 lite
Nano Banana 2 lite
Mô hình Gemini 3.1 Flash Lite Image là một chuyên gia về hiệu suất trong dòng mô hình tạo sinh hình ảnh, được thiết kế cho độ trễ siêu thấp và cho việc tạo sinh, chỉnh sửa hình ảnh hiệu quả về chi phí.
GPT-4.1 nano
GPT-4.1 nano
GPT-4.1 nano là một mô hình trí tuệ nhân tạo do OpenAI cung cấp. gpt-4.1-nano: Có cửa sổ ngữ cảnh lớn hơn—hỗ trợ tới 1 triệu token ngữ cảnh và tận dụng ngữ cảnh đó tốt hơn nhờ khả năng hiểu ngữ cảnh dài được cải thiện. Có mốc kiến thức được cập nhật là tháng 6 năm 2024. Mô hình này hỗ trợ độ dài ngữ cảnh tối đa là 1,047,576 token.

MiMo-V2.5
MiMo-V2.5 is Xiaomi's native full-modal model. It achieves professional-grade agent performance at about half the cost of inference, while outperforming MiMo-V2-Omni in multimodal perception in image and video understanding tasks.

GPT-5.4 pro
seedance-2-0
Happy Horse 1.1
Happy Horse 1.1
HappyHorse 1.1 là một mô hình tạo video đa mô thức, được thiết kế cho sáng tạo nội dung chuyên nghiệp, quảng cáo, phim ngắn, sản xuất nội dung trên mạng xã hội và kể chuyện. Mô hình này mở rộng các khả năng của HappyHorse 1.0 — vốn đã thu hút sự chú ý đáng kể sau khi đạt thứ hạng cao trong các đánh giá độc lập về tạo video — với tính liền mạch giữa các cảnh tốt hơn và độ trung thực hình ảnh được cải thiện.

GPT Image 2
Mô hình tạo hình ảnh mạnh mẽ nhất của OpenAI, với khả năng kết xuất văn bản gần như hoàn hảo trong nhiều ngôn ngữ, độ phân giải lên đến 4K và Thinking Mode được hỗ trợ bởi khả năng suy luận. Được xây dựng cho các quy trình sản xuất đòi hỏi độ chính xác, tốc độ và đầu ra hình ảnh đúng nhận diện thương hiệu.

DeepSeek V4 Flash
DeepSeek V4 Flash is an efficiency-optimized Mixture-of-Experts model from DeepSeek with 284B total parameters and 13B activated parameters, supporting a 1M-token context window. It is designed for fast inference and high-throughput workloads, while maintaining strong reasoning and coding performance.
.jpeg&w=3840&q=75)
GPT 5.5
Mô hình chủ lực thông minh và trực quan nhất của OpenAI, được thiết kế cho lập trình phức tạp, luồng công việc dựa trên tác nhân, sử dụng máy tính, phân tích dữ liệu và nghiên cứu chuyên sâu. Mang lại trí tuệ ở đẳng cấp tiên phong với độ trễ thấp tương đương GPT-5.4, đồng thời hoàn thành tác vụ với hiệu quả sử dụng token cao hơn. Là lựa chọn hàng đầu cho các khối lượng công việc chuyên nghiệp và doanh nghiệp đòi hỏi cao.

GPT 5.5 Pro
Mô hình mạnh nhất của OpenAI, được thiết kế cho các tác vụ khó nhất và các quy trình tác nhân tự chủ kéo dài. GPT-5.5 Pro xuất sắc trong lập trình phức tạp, sử dụng máy tính, nghiên cứu chuyên sâu, phân tích dữ liệu và suy luận khoa học — mang lại trí tuệ đẳng cấp tiên phong với độ trễ tương đương GPT-5.4 cùng hiệu quả sử dụng token cao hơn. Lý tưởng cho các trường hợp sử dụng cấp doanh nghiệp và chuyên nghiệp đòi hỏi tiêu chuẩn cao nhất về độ chính xác và khả năng thực thi tác vụ tự chủ.

Claude Sonnet 4.6
Claude Sonnet 4.6 is our most capable Sonnet model yet. It’s a full upgrade of the model’s skills across coding, computer use, long-context reasoning, agent planning, knowledge work, and design. Sonnet 4.6 also features a 1M token context window in beta.
Grok Imagine Video 1.5
Grok Imagine Video 1.5
Mô hình tạo video từ ảnh mới nhất của xAI tích hợp khả năng tạo âm thanh đồng bộ — video và âm thanh được tạo ra trong một lượt suy luận duy nhất. Hỗ trợ xuất 480p/720p, clip tối đa 15 giây, và được xếp hạng #1 trên bảng xếp hạng Image-to-Video Arena.

Claude Opus 4.7
Claude Opus 4.7 is a hybrid reasoning model designed specifically for frontier-level coding, AI agents, and complex multi-step professional work. Unlike lighter models (e.g., Sonnet or Haiku variants), Opus 4.7 prioritizes depth, consistency, and autonomy on the hardest tasks.
Happy Horse 1.0
Happy Horse 1.0
Happy Horse 1.0 — A high-quality audio-video generation model that supports text-to-video and image-to-video creation. It can generate synchronized visuals, audio, and lip movements, making it suitable for short films, advertising creatives, and product showcases.

Grok 4.3
Excels at agentic reasoning, knowledge work, and tool use.

GPT Image 2 ALL
GPT Image 2 is openai state-of-the-art image generation model for fast, high-quality image generation and editing. It supports flexible image sizes and high-fidelity image inputs.

Doubao-Seedance-2-0
Seedance 2.0 là mô hình nền tảng video đa phương thức thế hệ mới của ByteDance, tập trung vào việc tạo sinh video mang tính điện ảnh, theo mạch truyện nhiều cảnh. Khác với các bản demo chuyển văn bản thành video một cảnh, Seedance 2.0 nhấn mạnh điều khiển dựa trên tham chiếu (hình ảnh, clip ngắn, âm thanh), duy trì nhân vật và phong cách nhất quán, liền mạch giữa các cảnh, cùng đồng bộ âm thanh–hình ảnh tích hợp sẵn — nhằm biến video AI trở nên hữu ích cho các quy trình sáng tạo chuyên nghiệp và tiền trực quan hóa.

Sora 2 Pro
Sora 2 Pro là mô hình tạo sinh đa phương tiện tiên tiến và mạnh mẽ nhất của chúng tôi, có khả năng tạo video với âm thanh được đồng bộ hóa. Nó có thể tạo các đoạn video chi tiết, sinh động từ ngôn ngữ tự nhiên hoặc hình ảnh.

Nano Banana 2
Core Capabilities Overview: Resolution: Up to 4K (4096×4096), on par with Pro. Reference Image Consistency: Up to 14 reference images (10 objects + 4 characters), maintaining style/character consistency. Extreme Aspect Ratios: New 1:4, 4:1, 1:8, 8:1 ratios added, suitable for long images, posters, and banners. Text Rendering: Advanced text generation, suitable for infographics and marketing poster layouts. Search Enhancement: Integrated Google Search + Image Search. Grounding: Built-in thinking process; complex prompts are reasoned before generation.

GPT-5.2 Pro
gpt-5.2-pro is the highest-capability, production-oriented member of OpenAI’s GPT-5.2 family, exposed through the Responses API for workloads that demand maximal fidelity, multi-step reasoning, extensive tool use and the largest context/throughput budgets OpenAI offers.

DeepSeek V4 Pro
DeepSeek V4 Pro is a large-scale Mixture-of-Experts model from DeepSeek with 1.6T total parameters and 49B activated parameters, supporting a 1M-token context window. It is designed for advanced reasoning, coding, and long-horizon agent workflows, with strong performance across knowledge, math, and software engineering benchmarks.

MiniMax-M2.7
MiniMax-M2.7 offers the same top-tier intelligence as the standard version—including recursive self-evolution and expert-level office productivity—but is designed for applications requiring sub-second latency and high-speed token generation. Leveraging an enhanced inference backbone architecture, its output speed is 66% faster than the standard model (reaching 100 tps). It is the preferred choice for interactive programming assistants, real-time agent loop execution, and high-throughput enterprise pipelines with stringent completion time requirements.

GPT-5.4 nano
GPT-5.4 nano is designed for tasks where speed and cost matter most like classification, data extraction, ranking, and sub-agents.

GPT-5.4 mini
GPT-5.4 mini brings the strengths of GPT-5.4 to a faster, more efficient model designed for high-volume workloads.
Gemini omni fast
Gemini omni fast
Omni is the new model that can create anything from any input — starting with video. With Omni, you can combine images, audio, video and text as input and generate high-quality videos grounded in Gemini's real-world knowledge. You can also easily edit your videos through conversation.