
Gemini 3.7 Flash
3.7 Flash delivers substantial improvements across software engineering, knowledge work, and web development workflows — with an introductory price of half the original 3.6 Flash cost per million tokens.

3.7 Flash delivers substantial improvements across software engineering, knowledge work, and web development workflows — with an introductory price of half the original 3.6 Flash cost per million tokens.

Gemini 3.6 Flash provides sustained frontier-level intelligence optimized for real-world tasks at a higher speed and lower cost. Designed for the agentic era, it excels at code generation, agentic execution, and spatial reasoning. This model is particularly effective for rapid agentic loops involving complex coding cycles and iterations.

Gemini 3.1 Flash Lite Image model은 이미지 생성 제품군에서 효율에 특화된 모델로, 초저지연과 비용 효율적인 이미지 생성 및 수정을 위해 설계되었습니다.

Gemini 3.5 Flash is a high-speed AI model designed for fast response and efficient coding performance. It delivers significantly improved generation speed while maintaining strong reasoning ability, making it suitable for real-time applications and developer workflows.

Gemini 3.1 Pro is the next generation in the Gemini series of models, a suite of highly-capable, natively multimodal, reasoning models. Gemini 3 Pro is now Google’s most advanced model for complex tasks, and can comprehend vast datasets, challenging problems from different information sources, including text, audio, images, video, and entire code repositories

Core Capabilities Overview: Resolution: Up to 4K (4096×4096), on par with Pro. Reference Image Consistency: Up to 14 reference images (10 objects + 4 characters), maintaining style/character consistency. Extreme Aspect Ratios: New 1:4, 4:1, 1:8, 8:1 ratios added, suitable for long images, posters, and banners. Text Rendering: Advanced text generation, suitable for infographics and marketing poster layouts. Search Enhancement: Integrated Google Search + Image Search. Grounding: Built-in thinking process; complex prompts are reasoned before generation.

Gemini Omni Fast is a lightweight multimodal video generation model designed for fast and flexible content creation. It enables efficient video generation with support for multiple input types, making it suitable for interactive and iterative workflows.
Veo 3.1-Fast란 무엇인가 Veo 3.1-Fast는 Google의 Veo 3.1 생성형 비디오 모델 제품군에서 속도에 최適화된 변형 모델입니다. 이는 Veo 3.1에서 도입된 향상된 시청각 충실도를 유지하면서, 짧은 소셜 미디어용 영상 생성을 위해 지연 시간과 비용을 줄이도록 특별히 최적화되었습니다. 이전 Veo 릴리스와 비교해 Veo 3.1과 Veo 3.1-Fast는 더 풍부한 네이티브 오디오 생성, 더 강력한 프롬프트 준수, 그리고 새로운 편집 플로우(예: 첫/마지막 프레임 보간, “ingredients to video”, 및 장면 확장)를 추가합니다。

Veo 3.1은 Google의 Veo 텍스트·이미지→비디오 제품군에 대한 점진적이지만 중요한 업데이트로, 더 풍부한 네이티브 오디오, 더 길고 더 세밀하게 제어 가능한 비디오 출력, 그리고 더 정교한 편집 및 장면 수준 제어를 추가합니다.

Gemini 3.1 Flash-Lite is a highly cost-efficient and low-latency Tier-3 model in Google’s Gemini 3 series, designed for high-volume production AI workflow where throughput and speed matter more than maximal reasoning depth. It combines a large multimodal context window with efficient inference performance at a lower cost than most flagship counterparts.

Nano Banana Pro는 텍스트 중심 워크플로에서 범용 지원을 제공하는 AI 모델이다. 구조를 제어할 수 있는 형태로 콘텐츠를 생성·변환·분석하기 위한 지시문 기반 프롬프팅에 적합하다. 주요 활용 사례로는 채팅 어시스턴트, 문서 요약, 지식 질의응답, 워크플로 자동화가 있다. 공개된 기술 세부 정보는 제한적이며; 통합 방식은 구조화된 출력, 검색 증강 프롬프트, 도구 또는 함수 호출 등 일반적인 AI 어시스턴트 패턴과 부합한다.
Veo 3 Fast는 Google의 생성형 비디오 모델인 Veo 제품군(Veo 3 / Veo 3.1 등)의 속도 최적화 변형입니다. 처리량과 초당 비용을 우선시하면서 자체적으로 생성된 오디오를 포함한 짧고 고품질 비디오 클립을 생성하도록 설계되었으며—최상급 시각적 충실도 및/또는 더 긴 단일 샷 지속 시간의 일부를 포기하는 대신 훨씬 더 빠른 생성 속도와 더 낮은 가격을 제공합니다. Veo 3 Fast란 무엇인가 — 간단한 소개
Google DeepMind의 Veo 3는 텍스트-투-비디오 생성의 최첨단을 대표하며, 대규모 생성형 AI 모델이 고충실도 비디오를 대사, 효과음, 환경음 등 동반 오디오와 끊김 없이 동기화한 것은 이번이 처음이다.

Gemini 3 Flash is a lightweight, efficient multimodal large-scale model from Google tailored for real-world scenarios that require fast responses and low latency.
gemini-3.1-flash coming soon
coming soon
Gemini 3.5 Flash-Lite is a low-latency, cost-effective multimodal model optimized for high-throughput, low-cost execution for subagent tasks and document parsing. The model supports text, image, video, audio, and PDF inputs, and is designed for high-volume agentic workflows, simple data extraction, and applications where latency and API cost are the primary constraints
coming soon