برای برنامههای موبایل و وب، کیتهای توسعه نرمافزار Firebase AI Logic به شما امکان میدهند مستقیماً از طریق برنامه خود با مدلهای پشتیبانیشده Gemini تعامل داشته باشید.
مدلهای Gemini چندوجهی در نظر گرفته میشوند زیرا قادر به پردازش و حتی تولید چندین حالت، از جمله متن، کد، PDF، تصاویر، ویدیو و صدا هستند.
همچنین، سوالات متداول ما را در مورد تمام مدلهایی که Firebase AI Logic پشتیبانی میکند و پشتیبانی نمیکند، بررسی کنید.
مدلهای ویژه
فلش جمینی ۳.۷
gemini-3.7-flash
عملکردی در سطح کلاس Frontier که با کسری از قیمت، مدلهای بزرگتر را به رقابت میخواند.
جمینی ۳.۵ فلش-لایت
gemini-3.5-flash-lite
مدلی کارآمد، مقرونبهصرفه و با حجم تولید بالا، با عملکرد و کیفیت سری Gemini 3.
ایمیج فلش Gemini 3.1 ( نانو موز ۲ )
gemini-3.1-flash-image
مدل قدرتمند و با راندمان بالا برای تولید و ویرایش تصویر، بهینه شده برای سرعت و موارد استفاده با حجم بالا.
مدلهای عمومی
جمینی ۳.۱ پرو
gemini-3.1-pro-preview
هوش پیشرفته، مهارتهای حل مسئله پیچیده و قابلیتهای قدرتمند کدنویسی عاملی و ارتعاشی.
فلش جمینی ۳.۷
gemini-3.7-flash
عملکردی در سطح کلاس Frontier که با کسری از قیمت، مدلهای بزرگتر را به رقابت میخواند.
جمینی ۳.۵ فلش-لایت
gemini-3.5-flash-lite
مدلی کارآمد، مقرونبهصرفه و با حجم تولید بالا، با عملکرد و کیفیت سری Gemini 3.
مدلهای قدیمیتر و پایدار برای استفاده عمومی
Gemini 3.6 Flash (
gemini-3.6-flash): مدل قبلی Gemini 3.x Flash برای عملکرد در سطح بالا که با کسری از قیمت، مدلهای بزرگتر را رقابت میکند. (نیازی به پرداخت صورتحساب نیست )Gemini 3.5 Flash (
gemini-3.5-flash): مدل قبلی Gemini 3.x Flash برای عملکرد در سطح بالا که با کسری از قیمت، مدلهای بزرگتر را رقابت میکند. (نیازی به پرداخت صورتحساب نیست )Gemini 3.1 Flash‑Lite (
gemini-3.1-flash-lite): مدل قبلی Gemini 3.x Flash‑Lite برای کارهای سنگین و حساس به هزینه. (نیازی به پرداخت صورتحساب نیست )
مدلهای تولید تصویر
تصویر Gemini 3 Pro ( نانو موز پرو )
gemini-3-pro-image
مدل پیشرفته تولید و ویرایش تصویر برای خلق تصاویر بومی با بافت بسیار بالا.
ایمیج فلش Gemini 3.1 ( نانو موز ۲ )
gemini-3.1-flash-image
مدل قدرتمند و با راندمان بالا برای تولید و ویرایش تصویر، بهینه شده برای سرعت و موارد استفاده با حجم بالا.
ایمیج Gemini 3.1 Flash-Lite ( نانو موز ۲ لایت )
gemini-3.1-flash-lite-image
مدل تولید و ویرایش تصویر با تأخیر بسیار کم و مقرونبهصرفه، طراحیشده برای موارد استفاده تعاملی با حجم بالا.
مدلهای تولید صدا
مدلهای تبدیل متن به گفتار (TTS)
شما میتوانید با مدلهای Gemini TTS از ورودی متن، گفتار تولید کنید.
جمینی ۳.x فلش TTS
gemini-3.1-flash-tts-preview
تولید گفتار قدرتمند و با تأخیر کم از ورودی متن.
مدلهای Live API
شما میتوانید با مدلهایی که از Gemini Live API پشتیبانی میکنند، صدای استریمشدهی دوطرفه تولید کنید.
فلش Gemini 3.x با صدای بومی Gemini Live API
رابط برنامهنویسی کاربردی توسعهدهندگان جمینی:
gemini-3.1-flash-live-preview
API پلتفرم عامل Gemini:
پشتیبانی نمیشود
تعاملات صوتی و تصویری با تأخیر کم و بلادرنگ را با مدل Gemini که دو طرفه است، امکانپذیر میکند.
فلش Gemini 2.5 با صدای بومی Gemini Live API
رابط برنامهنویسی کاربردی توسعهدهندگان جمینی:
gemini-2.5-flash-native-audio-preview-12-2025
API پلتفرم عامل Gemini:
gemini-live-2.5-flash-native-audio
تعاملات صوتی و تصویری با تأخیر کم و بلادرنگ را با مدل Gemini که دو طرفه است، امکانپذیر میکند.
ادامهی این صفحه اطلاعات دقیقی در مورد مدلهای پشتیبانیشده توسط Firebase AI Logic ارائه میدهد.
- ورودی و خروجی پشتیبانی شده
- مقایسه سطح بالا از قابلیتهای پشتیبانیشده
- مشخصات و محدودیتها، برای مثال حداکثر توکنهای ورودی یا حداکثر طول ویدیوی ورودی
شرح نحوهی نسخهبندی مدلها ، به ویژه نسخههای پایدار ، پیشنمایش و آزمایشی آنها
فهرست نام مدلهای موجود برای گنجاندن در کد شما در هنگام مقداردهی اولیه
لیست زبانهای پشتیبانیشده برای مدلها
در پایین این صفحه، میتوانید اطلاعات دقیقی در مورد مدلهای نسل قبلی مشاهده کنید .
مقایسه مدلها
هر مدل قابلیتهای متفاوتی برای پشتیبانی از موارد استفاده مختلف دارد. توجه داشته باشید که هر یک از جداول این بخش، هر مدل را هنگام استفاده با Firebase AI Logic شرح میدهند. هر مدل ممکن است قابلیتهای اضافی داشته باشد که هنگام استفاده از SDK های ما در دسترس نیستند.
اگر اطلاعات مورد نظر خود را در زیربخشهای زیر پیدا نکردید، میتوانید اطلاعات بیشتری را در مستندات ارائهدهنده API انتخابی خود بیابید: Gemini Developer API یاAgent Platform Gemini API (که قبلاً Vertex AI نام داشت) .
ورودی و خروجی پشتیبانی شده
جدول زیر انواع ورودی و خروجی پشتیبانی شده هنگام استفاده از هر مدل با Firebase AI Logic را فهرست میکند.
برای آشنایی با انواع فایلهای پشتیبانیشده، به بخش فایلهای ورودی پشتیبانیشده و الزامات مراجعه کنید.
| Gemini 3.x Pro, Flash, Flash‑Lite | Gemini 3.x Pro Image | Gemini 3.x Flash Image | Gemini 3.x Flash‑Lite Image | Gemini 3.x Flash TTS | Gemini Live | |
|---|---|---|---|---|---|---|
| انواع ورودی | ||||||
| متن | ||||||
| کد | ||||||
| اسناد (پیدیاف یا متن ساده) | ||||||
| تصاویر | ||||||
| ویدئو | ||||||
| صوتی | ||||||
| انواع خروجی | ||||||
| متن | ||||||
| متن (پخش) | ||||||
| کد | ||||||
| خروجی ساختاریافته (مثل جیسون) | ||||||
| تصاویر | ||||||
| ویدئو | ||||||
| صوتی | ||||||
قابلیتها و ویژگیهای پشتیبانیشده
جدول زیر قابلیتها و ویژگیهای پشتیبانیشده هنگام استفاده از هر مدل با Firebase AI Logic را فهرست میکند.
| Gemini 3.x Pro, Flash, Flash‑Lite | Gemini 3.x Pro Image | Gemini 3.x Flash Image | Gemini 3.x Flash‑Lite Image | Gemini 3.x Flash TTS | Gemini Live | |
|---|---|---|---|---|---|---|
| تفکر | ||||||
| تولید متن از ورودیهای فقط متنی یا چندوجهی | به صورت لایه لایه یا به عنوان بخشی از تصویر | به صورت لایه لایه یا به عنوان بخشی از تصویر | به صورت لایه لایه یا به عنوان بخشی از تصویر | رونویسی صدا | ||
| تولید تصاویر | ||||||
| ویرایش تصاویر | ||||||
| تولید صدا | گفتار | صدای استریم شده | ||||
| تولید خروجی ساختاریافته (مثل جیسون) | ||||||
| اسناد را تجزیه و تحلیل کنید (پیدیاف یا متن ساده) ( خروجی متن | خروجی تصویر ) | ||||||
| تصاویر را تجزیه و تحلیل کنید ( خروجی متن | خروجی تصویر ) | ||||||
| تجزیه و تحلیل ویدیو ( خروجی متن | خروجی تصویر ) | ویدیوی استریم شده | |||||
| تجزیه و تحلیل صدا | صدای استریم شده | |||||
| چت چند نوبتی | ||||||
| جریانسازی چندوجهی دوطرفه | ||||||
| ابزارهای پشتیبانی شده | ||||||
| فراخوانی تابع | ||||||
| اجرای کد | ||||||
| زمینه URL | ||||||
| اتصال به زمین با | ||||||
| اتصال به زمین با | ||||||
مشخصات و محدودیتها
جدول زیر مشخصات و محدودیتهای استفاده از هر مدل با Firebase AI Logic را فهرست میکند.
| ملک | Gemini 3.x Pro, Flash, Flash‑Lite | Gemini 3.x Pro Image | Gemini 3.x Flash Image | Gemini 3.x Flash‑Lite Image | Gemini 3.x Flash TTS | Gemini Live |
|---|---|---|---|---|---|---|
| محدودیت توکن ورودی * | ۱,۰۴۸,۵۷۶ توکن | ۶۵,۵۳۶ توکن | ۱۳۱,۰۷۲ توکن | ۶۵,۵۳۶ توکن | ۸,۱۹۲ توکن | اسناد را ببینید |
| محدودیت توکن خروجی * | ۶۵,۵۳۶ توکن | ۳۲۷۶۸ توکن | ۳۲۷۶۸ توکن | ۴,۰۹۶ توکن | ۱۶۳۸۴ توکن | |
| فایلهای PDF (بنا به درخواست) | ||||||
| حداکثر تعداد از فایلهای PDF ورودی ** | ۹۰۰ فایل | ۱۴ فایل | ۱۴ فایل | ۱۴ فایل | --- | اسناد را ببینید |
| حداکثر تعداد از صفحات به ازای هر فایل PDF ورودی ** | ۹۰۰ صفحه | ۱۴ صفحه | ۱۴ صفحه | ۱۴ صفحه | --- | |
| حداکثر اندازه به ازای هر فایل PDF ورودی | ۵۰ مگابایت | ۵۰ مگابایت | ۵۰ مگابایت | ۵۰ مگابایت | --- | |
| تصاویر (بنا به درخواست) | ||||||
| حداکثر تعداد از تصاویر ورودی | ۱۰۰۰ تصویر | ۱۴ تصویر | ۱۴ تصویر | ۱۴ تصویر | --- | اسناد را ببینید |
| حداکثر اندازه تصویر کدگذاری شده با base64 به ازای هر ورودی | ۷ مگابایت | ۷ مگابایت | ۷ مگابایت | ۷ مگابایت | --- | |
| حداکثر تعداد از تصاویر خروجی | --- | حداکثر محدودیت توکن خروجی | حداکثر محدودیت توکن خروجی | حداکثر محدودیت توکن خروجی | --- | |
| ویدئو (به درخواست) | ||||||
| حداکثر تعداد از فایلهای ویدیویی ورودی | ۱۰ فایل | --- | حداکثر محدودیت توکن ورودی | حداکثر محدودیت توکن ورودی | --- | اسناد را ببینید |
| حداکثر طول از تمام ویدیوهای ورودی (فقط قابها) | حدود ۶۰ دقیقه | --- | حدود ۲۵ دقیقه | حدود ۱۲ دقیقه | --- | |
| حداکثر طول از تمام ویدیوهای ورودی (فریمها + صدا) | حدود ۴۵ دقیقه | --- | --- | --- | --- | |
| صدا (به درخواست) | ||||||
| حداکثر تعداد از فایلهای صوتی ورودی | ۱ فایل | --- | --- | --- | --- | اسناد را ببینید |
| حداکثر طول از تمام صداهای ورودی | حدود ۸.۴ ساعت | --- | --- | --- | --- | |
* برای همه مدلهای Gemini ، یک توکن معادل حدود ۴ کاراکتر است، بنابراین ۱۰۰ توکن حدود ۶۰ تا ۸۰ کلمه انگلیسی است. برای مدلهای Gemini ، میتوانید تعداد کل توکنها را در درخواستهای خود با استفاده countTokens تعیین کنید.
** فایلهای PDF به عنوان تصویر در نظر گرفته میشوند، بنابراین یک صفحه از PDF به عنوان یک تصویر در نظر گرفته میشود. تعداد صفحات مجاز در یک درخواست محدود به تعداد تصاویری است که مدل میتواند پشتیبانی کند.
اطلاعات دقیق اضافی را پیدا کنید
سهمیهها و قیمتگذاری برای هر مدل متفاوت است. قیمتگذاری همچنین به ورودی و خروجی بستگی دارد.
در مورد انواع فایلهای ورودی پشتیبانیشده، نحوه تعیین نوع MIME و نحوه اطمینان از اینکه فایلهای ورودی و درخواستهای چندوجهی شما الزامات را برآورده میکنند و از بهترین شیوهها در فایلها و الزامات ورودی پشتیبانیشده پیروی میکنند، اطلاعات کسب کنید.
الگوهای نسخهبندی و نامگذاری مدل
مدلها در نسخههای پایدار ، پیشنمایش و آزمایشی ارائه میشوند. برای راحتی،-latest نامهای مستعار پشتیبانی میشوند، اما به دلیل تناقض در اینکه به کدام نسخه مدل اشاره میکنند، توصیه نمیشوند.
حتماً بهترین شیوههای ما را برای استفاده از نسخههای مدل بررسی کنید.
برای یافتن نامهای مدل خاص برای استفاده در کد خود، به بخش «نامهای مدل موجود» در ادامه همین صفحه مراجعه کنید.
| نوع نسخه / مرحله انتشار | توضیحات | الگوی نام مدل | |
|---|---|---|---|
| پایدار | نسخههای پایدار از تاریخ انتشار برای استفاده در محیط عملیاتی در دسترس و پشتیبانی میشوند.
| نام مدلهای نسخههای پایدار پسوند ندارند مثال: | |
| پیشنمایش | نسخههای پیشنمایش دارای قابلیتهای جدیدی هستند و پایدار محسوب نمیشوند .
| نام مدلهای نسخههای پیشنمایش به همراه ... پیوست شدهاند. مثال: | |
| تجربی | نسخههای آزمایشی قابلیتهای جدیدی دارند و پایدار تلقی نمیشوند .
| نام مدلهای نسخههای آزمایشی به همراه ... پیوست شده است. مثال: | |
| خاموشی (بازنشستگی) | نسخههای خاموش (بازنشسته) از تاریخ خاموشی (بازنشستگی) خود گذشتهاند و بهطور دائم غیرفعال شدهاند.
| --- | |
بهترین شیوهها برای استفاده از نسخههای مدل
در برنامههای کاربردی خود، از نام مدل صریح برای جدیدترین نسخه پایدار استفاده کنید.
برای API پلتفرم عامل Gemini (که قبلاً Vertex AI نام داشت) ، اگر تصمیم دارید از یک مدل دسترسی کوتاهمدت در برنامه تولیدی خود استفاده کنید، استفاده از Firebase Remote Config یا قالبهای اعلان سرور برای کنترل نام مدل مورد استفاده برای ویژگی هوش مصنوعی شما، بسیار مهمتر است.
فقط در طول نمونهسازی اولیه از نسخههای پیشنمایش و آزمایشی استفاده کنید . توصیه میکنیم هنگام شروع توسعه و آزمایش برای یک مورد استفاده در مرحله تولید، از نسخه پایدار استفاده کنید.
ما استفاده از آن را توصیه نمیکنیم
-latestنامهای مستعار (حتی در حین توسعه). این نام مستعار به آخرین نسخه برای یک مدل خاص اشاره میکند (که میتواند یک نسخه پایدار، پیشنمایش یا آزمایشی باشد). این نام مستعار با هر نسخه جدید از یک مدل خاص، به صورت خودکار تغییر میکند و فقط برای تغییرات جزئی، یک ایمیل اعلان ۲ هفته قبل ارسال میشود. این عدم ثبات در مدلی که واقعاً از آن استفاده میکنید، میتواند منجر به تغییرات رفتاری غیرمنتظره برای ویژگی هوش مصنوعی شما شود.
نام مدلهای موجود
نامهای مدل، مقادیر صریحی هستند که شما در هنگام مقداردهی اولیه مدل، در کد خود قرار میدهید.
مدلهای عمومی (مانند
gemini-3.7-flash)مدلهای تولیدکننده تصویر (مانند
gemini-3.1-flash-image، که با نام مدلهای "نانو موز" نیز شناخته میشوند)مدلهای تولید صدا :
- مدلهای تبدیل متن به گفتار (TTS) (مانند
gemini-3.1-flash-tts-preview) - مدلهای Live API (مانند
gemini-live-2.5-flash-native-audio)
- مدلهای تبدیل متن به گفتار (TTS) (مانند
برای مثالهای مقداردهی اولیه برای پلتفرم خود، به راهنمای شروع به کار مراجعه کنید.
برای جزئیات بیشتر در مورد مراحل انتشار (به ویژه برای موارد استفاده، صدور صورتحساب و خاموش کردن)، به الگوهای نسخهبندی و نامگذاری مدل مراجعه کنید.
لیست کردن تمام مدلهای موجود به صورت برنامهنویسی شده
شما میتوانید با استفاده از REST API، نام تمام مدلهای موجود را فهرست کنید:
رابط برنامهنویسی کاربردی توسعهدهندگان Gemini : فراخوانی نقطه پایانی
models.listAPI پلتفرم عامل Gemini (که قبلاً Vertex AI نام داشت) : با نقطه پایانی
publishers.models.listتماس بگیرید
توجه داشته باشید که این لیست برگشتی شامل تمام مدلهای پشتیبانیشده توسط ارائهدهندگان API خواهد بود، اما Firebase AI Logic فقط از مدلهای Gemini که در این صفحه توضیح داده شدهاند، پشتیبانی میکند.
مدلهای عمومی
نام مدلهای Gemini 3.x Pro
صرف نظر از ارائه دهنده API Gemini شما ، به طرح قیمت گذاری Blaze با پرداخت در محل نیاز دارد.
| نام مدل | توضیحات | مرحله انتشار | تاریخ انتشار | تاریخ خاموش شدن |
|---|---|---|---|---|
gemini-3.1-pro-preview | آخرین نسخه پیشنمایش Gemini 3.x Pro | پیشنمایش | ۲۰۲۶-۰۲-۱۹ | تعیین خواهد شد |
نام مدلهای Gemini 3.x Flash
اگر از رابط برنامهنویسی کاربردی (API) توسعهدهندگان Gemini استفاده میکنید، نیازی به طرح قیمتگذاری Blaze با پرداخت در محل ندارید .
| نام مدل | توضیحات | مرحله انتشار | تاریخ انتشار | تاریخ خاموش شدن |
|---|---|---|---|---|
gemini-3.7-flash | آخرین نسخه پایدار Gemini 3.x Flash این یک مدل دسترسی کوتاهمدت است (به الگوهای نسخهبندی و نامگذاری مدل مراجعه کنید). | پایدار | ۲۰۲۶-۰۸-۱۳ | تعیین خواهد شد |
gemini-3.6-flash | نسخه پایدار Gemini 3.x Flash این یک مدل دسترسی کوتاهمدت است (به الگوهای نسخهبندی و نامگذاری مدل مراجعه کنید). | پایدار | ۲۰۲۶-۰۷-۲۱ | تعیین خواهد شد |
gemini-3.5-flash | نسخه پایدار Gemini 3.x Flash | پایدار | ۲۰۲۶-۰۵-۱۹ | نه زودتر از ۲۰۲۷-۰۵-۱۹ |
gemini-3-flash-preview | نسخه پیشنمایش Gemini 3.x Flash | پیشنمایش | ۲۰۲۵-۱۲-۱۷ | تعیین خواهد شد |
نام مدلهای Gemini 3.x Flash‑Lite
اگر از رابط برنامهنویسی کاربردی (API) توسعهدهندگان Gemini استفاده میکنید، نیازی به طرح قیمتگذاری Blaze با پرداخت در محل ندارید .
| نام مدل | توضیحات | مرحله انتشار | تاریخ انتشار | تاریخ خاموش شدن |
|---|---|---|---|---|
gemini-3.5-flash-lite | آخرین نسخه پایدار Gemini 3.x Flash‑Lite | پایدار | ۲۰۲۶-۰۷-۲۱ | نه زودتر از ۲۰۲۷-۰۷-۲۱ |
gemini-3.1-flash-lite | نسخه پایدار Gemini 3.x Flash‑Lite | پایدار | ۲۰۲۶-۰۵-۰۷ | نه زودتر از ۲۰۲۷-۰۵-۰۷ |
مدلهای تولید تصویر
نام مدلهای Gemini 3.x Pro Image (معروف به "Nano Banana Pro")
صرف نظر از ارائه دهنده API Gemini شما ، به طرح قیمت گذاری Blaze با پرداخت در محل نیاز دارد.
| نام مدل | توضیحات | مرحله انتشار | تاریخ انتشار | تاریخ خاموش شدن |
|---|---|---|---|---|
gemini-3-pro-image | نسخه پایدار Gemini 3.x Pro Image (معروف به "نانو موز پرو") | پایدار | ۲۰۲۶-۰۵-۲۸ | نه زودتر از ۲۰۲۷-۰۵-۲۸ |
gemini-3-pro-image-preview | نسخه پیشنمایش Gemini 3.x Pro Image (معروف به "نانو موز پرو") | پیشنمایش | ۲۰۲۵-۱۱-۲۰ | همان اوایل که ۲۰۲۶-۰۶-۲۵ |
نامهای مدل تصویر فلش Gemini 3.x (معروف به "نانو موز ۲")
صرف نظر از ارائه دهنده API Gemini شما ، به طرح قیمت گذاری Blaze با پرداخت در محل نیاز دارد.
| نام مدل | توضیحات | مرحله انتشار | تاریخ انتشار | تاریخ خاموش شدن |
|---|---|---|---|---|
gemini-3.1-flash-image | نسخه پایدار Gemini 3.x Flash Image (معروف به "نانو موز ۲") | پایدار | ۲۰۲۶-۰۵-۲۸ | نه زودتر از ۲۰۲۷-۰۵-۲۸ |
gemini-3.1-flash-image-preview | نسخه پیشنمایش تصویر فلش Gemini 3.x (معروف به "نانو موز ۲") | پیشنمایش | ۲۶-۰۲-۲۰۲۶ | همان اوایل که ۲۰۲۶-۰۶-۲۵ |
نامهای مدل تصویر Gemini 3.x Flash‑Lite (معروف به "Nano Banana 2 Lite")
صرف نظر از ارائه دهنده API Gemini شما ، به طرح قیمت گذاری Blaze با پرداخت در محل نیاز دارد.
| نام مدل | توضیحات | مرحله انتشار | تاریخ انتشار | تاریخ خاموش شدن |
|---|---|---|---|---|
gemini-3.1-flash-lite-image | نسخه پایدار Gemini 3.x Flash‑Lite Image (معروف به "نانو موز ۲ لایت") این یک مدل دسترسی کوتاهمدت است (به الگوهای نسخهبندی و نامگذاری مدل مراجعه کنید). | پایدار | ۲۰۲۶-۰۶-۳۰ | تعیین خواهد شد |
مدلهای صوتی - تبدیل متن به گفتار (TTS)
نام مدلهای Gemini 3.x Flash TTS
اگر از رابط برنامهنویسی کاربردی (API) توسعهدهندگان Gemini استفاده میکنید، نیازی به طرح قیمتگذاری Blaze با پرداخت در محل ندارید .
| نام مدل | توضیحات | مرحله انتشار | تاریخ انتشار | تاریخ خاموش شدن |
|---|---|---|---|---|
gemini-3.1-flash-tts-preview | نسخه پیشنمایش برای Gemini 3.x Flash TTS | پیشنمایش | ۲۰۲۶-۰۶-۱۷ | تعیین خواهد شد |
صدا - مدلهای Live API
نام مدلهای Gemini 3.x Flash Live
| رابط برنامهنویسی کاربردی (API) توسعهدهندگان جمینی نام مدل | توضیحات | مرحله انتشار | تاریخ انتشار | تاریخ خاموش شدن |
|---|---|---|---|---|
gemini-3.1-flash-live-preview ۱ | نسخه پیشنمایش برای Live API در رابط برنامهنویسی توسعهدهنده Gemini | پیشنمایش | ۲۰۲۶-۰۳-۲۶ | تعیین خواهد شد |
| رابط برنامهنویسی کاربردی پلتفرم عامل Gemini (که قبلاً Vertex AI نام داشت) نام مدل | توضیحات | مرحله انتشار | تاریخ انتشار | تاریخ خاموش شدن |
|---|---|---|---|---|
| رابط برنامهنویسی کاربردی (API) پلتفرم عامل Gemini (که قبلاً Vertex AI نام داشت) از هیچ یک از مدلهای Gemini Live 3.x پشتیبانی نمیکند. | ||||
۱ فقط توسط رابط برنامهنویسی نرمافزار توسعهدهندگان Gemini پشتیبانی میشود. همچنین، اگرچه این یک مدل پیشنمایش است، اما در «ردیف رایگان» رابط برنامهنویسی نرمافزار توسعهدهندگان Gemini موجود است.
نام مدلهای Gemini 2.5 Flash Live
اگر از رابط برنامهنویسی Gemini Developer API استفاده میکنید، نیازی به طرح قیمتگذاری Blaze که در صورت استفاده پرداخت میشود، ندارد (معمولاً مدلهای پیشنمایش به طرح پولی نیاز دارند).
اگرچه مدلهای زیر بسته به ارائهدهنده API Gemini نامهای مختلفی دارند، اما ویژگیهای مدل یکسان است.
| رابط برنامهنویسی کاربردی (API) توسعهدهندگان جمینی نام مدل | توضیحات | مرحله انتشار | تاریخ انتشار | تاریخ خاموش شدن |
|---|---|---|---|---|
gemini-2.5-flash-native-audio-preview-12-2025 ۱ | نسخه پیشنمایش برای Live API در رابط برنامهنویسی توسعهدهنده Gemini | پیشنمایش | ۲۰۲۵-۱۲-۱۲ | تعیین خواهد شد |
gemini-2.5-flash-native-audio-preview-09-2025 ۱ | نسخه پیشنمایش اولیه برای Live API در Gemini Developer API | پیشنمایش | ۲۰۲۵-۰۹-۱۸ | تعیین خواهد شد |
| رابط برنامهنویسی کاربردی پلتفرم عامل Gemini (که قبلاً Vertex AI نام داشت) نام مدل | توضیحات | مرحله انتشار | تاریخ انتشار | تاریخ خاموش شدن |
|---|---|---|---|---|
gemini-live-2.5-flash-native-audio شدهاند | نسخه پایدار برای Live API در پلتفرم Agent API Gemini (که قبلاً Vertex AI نام داشت) | پایدار | ۲۰۲۵-۱۲-۱۲ | نه زودتر از ۲۰۲۶-۱۲-۱۲ |
gemini-live-2.5-flash-preview-native-audio-09-2025 2 | نسخه پیشنمایش برای Live API در پلتفرم عامل Gemini API (که قبلاً Vertex AI نام داشت) | پیشنمایش | ۲۰۲۵-۰۹-۱۸ | تعیین خواهد شد |
۱ فقط توسط رابط برنامهنویسی نرمافزار Gemini Developer پشتیبانی میشود. همچنین، اگرچه اینها مدلهای پیشنمایش هستند، اما در «ردیف رایگان» رابط برنامهنویسی نرمافزار Gemini Developer در دسترس هستند.
۲ فقط توسط API پلتفرم عامل Gemini (که قبلاً Vertex AI نام داشت) پشتیبانی میشود. همچنین، این مدلها در موقعیت global در دسترس نیستند .
زبانهای پشتیبانیشده
تمام مدلهای Gemini میتوانند زبانهای زیر را درک کرده و به آنها پاسخ دهند:
عربی (ar)، بنگالی (bn)، بلغاری (bg)، چینی سادهشده و سنتی (zh)، کرواتی (hr)، چکی (cs)، دانمارکی (da)، هلندی (nl)، انگلیسی (en)، استونیایی (et)، فنلاندی (fi)، فرانسوی (fr)، آلمانی (de)، یونانی (el)، عبری (iw)، هندی (hi)، مجارستانی (hu)، اندونزیایی (id)، ایتالیایی (it)، ژاپنی (ja)، کرهای (ko)، لتونیایی (lv)، لیتوانیایی (lt)، نروژی (no)، لهستانی (pl)، پرتغالی (pt)، رومانیایی (ro)، روسی (ru)، صربی (sr)، اسلواکی (sk)، اسلوونیایی (sl)، اسپانیایی (es)، سواحیلی (sw)، سوئدی (sv)، تایلندی (th)، ترکی (tr)، اوکراینی (uk)، ویتنامی (vi)
مدلهای Gemini 2.0 Flash ، Gemini 1.5 Pro و Gemini 1.5 Flash میتوانند زبانهای اضافی زیر را درک کرده و به آنها پاسخ دهند:
آفریکانس (af)، آمهری (am)، آسامی (ع)، آذربایجانی (az)، بلاروسی (be)، بوسنیایی (bs)، کاتالان (ca)، سبوانو (ceb)، کورسی (co)، ولزی (cy)، Dhivehi (dv)، اسپرانتو (eo)، باسک (eu)، فارسی (fa)، فیلیپینی (تاگالوگ) (fil)، (fy)، ایرلندی (ga)، اسکاتلندی (ga)، اسکاتلندی (ga) (ha)، هاوایی (haw)، همونگ (hmn)، کریول هائیتی (ht)، ارمنی (hy)، ایگبو (ig)، ایسلندی (is)، جاوه ای (jv)، گرجی (ka)، قزاقستان (kk)، خمر (km)، کانادا (kn)، کریو (kri)، کردی (ku)، قرقیز (ky)، لاتین (la)، لوکزامبورگ (lb)، لائو (lom)، مقدونیه، مالاگاسی (mk)، مالایالام (ml)، مغولی (mn)، Meiteilon (Manipuri) (mni-Mtei)، مراتی (mr)، مالایی (ms)، مالتی (mt)، میانمار (برمه) (my)، نپالی (ne)، Nyanja (Chichewa) (ny)، Odia (Oriya) (یا)، پنجابی (pa)، Pashinhales (Pashto) (si)، ساموآیی (sm)، شونا (sn)، سومالیایی (so)، آلبانیایی (sq)، سسوتو (st)، سوندانی (su)، تامیلی (ta)، تلوگو (te)، تاجیکی (tg)، اویغور (ug)، اردو (ur)، ازبکی (uz)، Xhosa (xh)، ییدیش (yi)، یروبا (yo)، زولو (zu)
اطلاعات مربوط به مدلهای قبلی
مدلهای زیر فعال هستند، اما از نسل قبلی میباشند. توصیه میکنیم در صورت امکان از جدیدترین مدلها استفاده کنید.
اگر اطلاعات مورد نظر خود را در زیربخشهای زیر پیدا نکردید، میتوانید اطلاعات بیشتری را در مستندات ارائهدهنده API انتخابی خود بیابید: Gemini Developer API یاAgent Platform Gemini API (قبلاً Vertex AI)
مدلهای قدیمیتر جمینی
-
gemini-2.5-pro -
gemini-2.5-flash -
gemini-2.5-flash-lite -
gemini-2.5-flash-image(معروف به "نانو موز") -
gemini-2.0-flash-001(و نام مستعار بهروزرسانیشده خودکار آنgemini-2.0-flash) -
gemini-2.0-flash-lite-001(و نام مستعار بهروزرسانیشده خودکار آنgemini-2.0-flash-lite)
برای اطلاعات بیشتر در مورد مدلهای قدیمیتر Gemini Live API ، به مستندات ارائهدهنده Gemini API مراجعه کنید:
مدلهای قدیمیتر ایمیجن
-
imagen-4.0-ultra-generate-001 -
imagen-4.0-generate-001 -
imagen-4.0-fast-generate-001 -
imagen-3.0-capability-001 -
imagen-3.0-generate-002 -
imagen-3.0-generate-001 -
imagen-3.0-fast-generate-001
مشاهده جزئیات درباره مدلهای قبلی
اینها انواع ورودی و خروجی هنگام استفاده از هر مدل با Firebase AI Logic هستند:
| Gemini 2.5 Pro, Flash, Flash‑Lite | Gemini 2.5 Flash Image | Gemini 2.0 Flash | Gemini 2.0 Flash‑Lite | ایمیجِن (تولید کردن) | ایمیجِن (قابلیت) | |
|---|---|---|---|---|---|---|
| انواع ورودی | ||||||
| متن | ||||||
| کد | ||||||
| اسناد (پیدیاف یا متن ساده) | ||||||
| تصاویر | ||||||
| ویدئو | ||||||
| صوتی | ||||||
| صدا (پخش جریانی) | ||||||
| انواع خروجی | ||||||
| متن | ||||||
| متن (پخش) | ||||||
| کد | ||||||
| خروجی ساختاریافته (مثل جیسون) | ||||||
| تصاویر | ||||||
| ویدئو | ||||||
| صوتی | ||||||
| صدا (پخش جریانی) | ||||||
این قابلیتها و ویژگیها هنگام استفاده از هر مدل با Firebase AI Logic وجود دارد:
| Gemini 2.5 Pro, Flash, Flash‑Lite | Gemini 2.5 Flash Image | Gemini 2.0 Flash | Gemini 2.0 Flash‑Lite | ایمیجِن (تولید کردن) | ایمیجِن (قابلیت) | |
|---|---|---|---|---|---|---|
| تفکر | ||||||
| تولید متن از ورودیهای فقط متنی یا چندوجهی | به صورت لایه لایه یا به عنوان بخشی از تصویر | |||||
| تولید تصاویر ( جوزا یا ایمیجن ) | ||||||
| ویرایش تصاویر ( جوزا یا ایمیجن ) | ||||||
| تولید صدا | ||||||
| تولید خروجی ساختاریافته (مثل جیسون) | ||||||
| اسناد را تجزیه و تحلیل کنید (پیدیاف یا متن ساده) ( خروجی متن | خروجی تصویر ) | ||||||
| تصاویر را تجزیه و تحلیل کنید ( خروجی متن | خروجی تصویر ) | ||||||
| تجزیه و تحلیل ویدیو ( خروجی متن | خروجی تصویر ) | ||||||
| تجزیه و تحلیل صدا | ||||||
| چت چند نوبتی | ||||||
| جریانسازی چندوجهی دوطرفه | ||||||
| ابزارهای پشتیبانی شده | ||||||
| فراخوانی تابع | ||||||
| اجرای کد | ||||||
| زمینه URL | ||||||
| اتصال به زمین با | ||||||
| اتصال به زمین با | ||||||
| دستورالعملهای سیستم | ||||||
| تعداد توکنها | ||||||
مشخصات و محدودیتهای استفاده از هر مدل با Firebase AI Logic به شرح زیر است:
| ملک | Gemini 2.5 Pro, Flash, Flash‑Lite | Gemini 2.5 Flash Image | Gemini 2.0 Flash | Gemini 2.0 Flash‑Lite | ایمیجِن (تولید کردن) | ایمیجِن (قابلیت) |
|---|---|---|---|---|---|---|
| محدودیت توکن ورودی * | ۱,۰۴۸,۵۷۶ توکن | ۳۲۷۶۸ توکن | ۱,۰۴۸,۵۷۶ توکن | ۱,۰۴۸,۵۷۶ توکن | ۴۸۰ توکن | ۴۸۰ توکن |
| محدودیت توکن خروجی * | ۶۵,۵۳۶ توکن | ۸,۱۹۲ توکن | ۸,۱۹۲ توکن | ۸,۱۹۲ توکن | --- | --- |
| تاریخ پایان دانش | ژانویه ۲۰۲۵ | --- | ژوئن ۲۰۲۴ | ژوئن ۲۰۲۴ | --- | --- |
| فایلهای PDF (بنا به درخواست) | ||||||
| حداکثر تعداد از فایلهای PDF ورودی ** | ۳۰۰۰ فایل | ۳ فایل | ۳۰۰۰ فایل | ۳۰۰۰ فایل | --- | --- |
| حداکثر تعداد از صفحات به ازای هر فایل PDF ورودی ** | ۱۰۰۰ صفحه | ۳ صفحه | ۱۰۰۰ صفحه | ۱۰۰۰ صفحه | --- | --- |
| حداکثر اندازه به ازای هر فایل PDF ورودی | ۵۰ مگابایت | ۵۰ مگابایت | ۵۰ مگابایت | ۵۰ مگابایت | --- | --- |
| تصاویر (بنا به درخواست) | ||||||
| حداکثر تعداد از تصاویر ورودی | ۳۰۰۰ تصویر | ۳ تصویر | ۳۰۰۰ تصویر | ۳۰۰۰ تصویر | --- | ۴ تصویر |
| حداکثر تعداد از تصاویر خروجی | --- | حداکثر محدودیت توکن خروجی | --- | --- | ۴ تصویر | ۴ تصویر |
| حداکثر اندازه تصویر کدگذاری شده با base64 به ازای هر ورودی | ۷ مگابایت | ۷ مگابایت | ۷ مگابایت | ۷ مگابایت | --- | --- |
| ویدئو (به درخواست) | ||||||
| حداکثر تعداد از فایلهای ویدیویی ورودی | ۱۰ فایل | --- | ۱۰ فایل | ۱۰ فایل | --- | --- |
| حداکثر طول از تمام ویدیوهای ورودی (فقط قابها) | حدود ۶۰ دقیقه | --- | حدود ۶۰ دقیقه | حدود ۶۰ دقیقه | --- | --- |
| حداکثر طول از تمام ویدیوهای ورودی (فریمها + صدا) | حدود ۴۵ دقیقه | --- | حدود ۴۵ دقیقه | حدود ۴۵ دقیقه | --- | --- |
| صدا (به درخواست) | ||||||
| حداکثر تعداد از فایلهای صوتی ورودی | ۱ فایل | --- | ۱ فایل | ۱ فایل | --- | --- |
| حداکثر تعداد از فایلهای صوتی خروجی | --- | --- | --- | --- | --- | --- |
| حداکثر طول از تمام صداهای ورودی | حدود ۸.۴ ساعت | --- | حدود ۸.۴ ساعت | حدود ۸.۴ ساعت | --- | --- |
| حداکثر طول از تمام صداهای خروجی | --- | --- | --- | --- | --- | --- |
* برای همه مدلهای Gemini ، یک توکن معادل حدود ۴ کاراکتر است، بنابراین ۱۰۰ توکن حدود ۶۰ تا ۸۰ کلمه انگلیسی است. برای مدلهای Gemini ، میتوانید تعداد کل توکنها را در درخواستهای خود با استفاده countTokens تعیین کنید.
** فایلهای PDF به عنوان تصویر در نظر گرفته میشوند، بنابراین یک صفحه از PDF به عنوان یک تصویر در نظر گرفته میشود. تعداد صفحات مجاز در یک درخواست محدود به تعداد تصاویری است که مدل میتواند پشتیبانی کند.
نامهای مدل، مقادیر صریحی هستند که شما در هنگام مقداردهی اولیه مدل، در کد خود قرار میدهید.
مدلهای جمینی
نام مدلهای Gemini 2.5 Pro
اگر از رابط برنامهنویسی کاربردی (API) توسعهدهندگان Gemini استفاده میکنید، نیازی به طرح قیمتگذاری Blaze با پرداخت در محل ندارید .
| نام مدل | توضیحات | مرحله انتشار | تاریخ انتشار | تاریخ خاموش شدن |
|---|---|---|---|---|
gemini-2.5-pro | نسخه پایدار Gemini 2.5 Pro | پایدار | ۲۰۲۵-۰۶-۱۷ | همان اوایل که ۲۰۲۶-۱۰-۱۶ |
نام مدلهای Gemini 2.5 Flash
اگر از رابط برنامهنویسی کاربردی (API) توسعهدهندگان Gemini استفاده میکنید، نیازی به طرح قیمتگذاری Blaze با پرداخت در محل ندارید .
| نام مدل | توضیحات | مرحله انتشار | تاریخ انتشار | تاریخ خاموش شدن |
|---|---|---|---|---|
gemini-2.5-flash | نسخه پایدار Gemini 2.5 Flash | پایدار | ۲۰۲۵-۰۶-۱۷ | همان اوایل که ۲۰۲۶-۱۰-۱۶ |
نام مدلهای Gemini 2.5 Flash‑Lite
اگر از رابط برنامهنویسی کاربردی (API) توسعهدهندگان Gemini استفاده میکنید، نیازی به طرح قیمتگذاری Blaze با پرداخت در محل ندارید .
| نام مدل | توضیحات | مرحله انتشار | تاریخ انتشار | تاریخ خاموش شدن |
|---|---|---|---|---|
gemini-2.5-flash-lite | نسخه پایدار Gemini 2.5 Flash‑Lite | پایدار | ۲۰۲۵-۰۷-۲۲ | همان اوایل که ۲۰۲۶-۱۰-۱۶ |
نام مدلهای Gemini 2.5 Flash Image (معروف به "نانو موز")
صرف نظر از ارائه دهنده API Gemini شما ، به طرح قیمت گذاری Blaze با پرداخت در محل نیاز دارد.
| نام مدل | توضیحات | مرحله انتشار | تاریخ انتشار | تاریخ خاموش شدن |
|---|---|---|---|---|
gemini-2.5-flash-image | نسخه پایدار Gemini 2.5 Flash Image (معروف به "نانو موز") | پایدار | ۲۰۲۵-۱۰-۰۲ | ۲۰۲۶-۱۰-۰۲ |
نام مدلهای Gemini 2.0 Flash
| نام مدل | توضیحات | مرحله انتشار | تاریخ انتشار | تاریخ خاموش شدن |
|---|---|---|---|---|
gemini-2.0-flash-001 | آخرین نسخه پایدار Gemini 2.0 Flash | پایدار | ۲۰۲۵-۰۲-۰۵ | ۲۰۲۶-۰۶-۰۱ |
gemini-2.0-flash | نام مستعار بهروزرسانیشده خودکار که به آخرین نسخه پایدار Gemini 2.0 Flash اشاره دارد (در حال حاضر gemini-2.0-flash-001 ) | پایدار | ۲۰۲۵-۰۲-۱۰ | ۲۰۲۶-۰۶-۰۱ |
نام مدلهای Gemini 2.0 Flash‑Lite
| نام مدل | توضیحات | مرحله انتشار | تاریخ انتشار | تاریخ خاموش شدن |
|---|---|---|---|---|
gemini-2.0-flash-lite-001 | آخرین نسخه پایدار Gemini 2.0 Flash‑Lite | پایدار | 2025-02-25 | ۲۰۲۶-۰۶-۰۱ |
gemini-2.0-flash-lite | نام مستعار بهروزرسانیشده خودکار که به آخرین نسخه پایدار Gemini 2.0 Flash‑Lite اشاره دارد (در حال حاضر gemini-2.0-flash-lite-001 ) | پایدار | 2025-02-25 | ۲۰۲۶-۰۶-۰۱ |
مدلهای ایمیجن
نام مدلهای ایمیجن ۴
| نام مدل | توضیحات | مرحله انتشار | تاریخ انتشار | تاریخ خاموش شدن |
|---|---|---|---|---|
imagen-4.0-generate-001 | نسخه پایدار Imagen 4 | پایدار | ۱۴-۰۸-۲۰۲۵ | ۲۰۲۶-۰۶-۳۰ |
نام مدلهای Imagen 4 Fast
| نام مدل | توضیحات | مرحله انتشار | تاریخ انتشار | تاریخ خاموش شدن |
|---|---|---|---|---|
imagen-4.0-fast-generate-001 | نسخه پایدار Imagen 4 Fast | پایدار | ۱۴-۰۸-۲۰۲۵ | ۲۰۲۶-۰۶-۳۰ |
نام مدلهای ایمیجن ۴ اولترا
| نام مدل | توضیحات | مرحله انتشار | تاریخ انتشار | تاریخ خاموش شدن |
|---|---|---|---|---|
imagen-4.0-ultra-generate-001 | نسخه پایدار Imagen 4 Ultra | پایدار | ۱۴-۰۸-۲۰۲۵ | ۲۰۲۶-۰۶-۳۰ |
نام مدلهای قابلیت Imagen 3
| نام مدل | توضیحات | مرحله انتشار | تاریخ انتشار | تاریخ خاموش شدن |
|---|---|---|---|---|
imagen-3.0-capability-001 | نسخه پایدار اولیه قابلیت Imagen 3 | پایدار | ۲۰۲۴-۱۲-۱۰ | ۲۰۲۶-۰۶-۳۰ |
نام مدلهای ایمیجن ۳
| نام مدل | توضیحات | مرحله انتشار | تاریخ انتشار | تاریخ خاموش شدن |
|---|---|---|---|---|
imagen-3.0-generate-002 | آخرین نسخه پایدار Imagen 3 | پایدار | ۲۰۲۵-۰۱-۲۳ | ۲۰۲۶-۰۶-۳۰ |
imagen-3.0-generate-001 | نسخه پایدار اولیه Imagen 3 | پایدار | ۲۰۲۴-۰۷-۳۱ | ۲۰۲۶-۰۶-۳۰ |
نام مدلهای سریع Imagen 3
| نام مدل | توضیحات | مرحله انتشار | تاریخ انتشار | تاریخ خاموش شدن |
|---|---|---|---|---|
imagen-3.0-fast-generate-001 | نسخه پایدار اولیه Imagen 3 Fast | پایدار | ۲۰۲۴-۰۷-۳۱ | ۲۰۲۶-۰۶-۳۰ |
مراحل بعدی
قابلیتهای رابط برنامهنویسی Gemini را امتحان کنید
- مکالمات چند نوبتی (چت) بسازید.
- تولید متن از درخواستهای فقط متنی .
- با استفاده از انواع فایلهای مختلف، مانند تصاویر ، فایلهای PDF ، ویدیو و صدا ، متن را با پرسوجو تولید کنید.
- خروجی ساختاریافته (مانند JSON) را از هر دو حالت متنی و چندوجهی تولید کنید .
- تصاویر را از متن و فرمهای چندوجهی تولید و ویرایش کنید .
- تولید گفتار (هم تکگوینده و هم چندگوینده) با استفاده از مدلهای تبدیل متن به گفتار (TTS) جمینی .
- ورودی و خروجی (از جمله صدا) را با استفاده از Gemini Live API استریم کنید.
- از ابزارهایی (مانند فراخوانی تابع و اتصال به زمین با
Google Search یاGoogle Maps ) برای اتصال یک مدل Gemini به سایر بخشهای برنامه و سیستمها و اطلاعات خارجی خود استفاده کنید.
برای برنامههای موبایل و وب، کیتهای توسعه نرمافزار Firebase AI Logic به شما امکان میدهند مستقیماً از طریق برنامه خود با مدلهای پشتیبانیشده Gemini تعامل داشته باشید.
مدلهای Gemini چندوجهی در نظر گرفته میشوند زیرا قادر به پردازش و حتی تولید چندین حالت، از جمله متن، کد، PDF، تصاویر، ویدیو و صدا هستند.
همچنین، سوالات متداول ما را در مورد تمام مدلهایی که Firebase AI Logic پشتیبانی میکند و پشتیبانی نمیکند، بررسی کنید.
مدلهای ویژه
فلش جمینی ۳.۷
gemini-3.7-flash
عملکردی در سطح کلاس Frontier که با کسری از قیمت، مدلهای بزرگتر را به رقابت میخواند.
جمینی ۳.۵ فلش-لایت
gemini-3.5-flash-lite
مدلی کارآمد، مقرونبهصرفه و با حجم تولید بالا، با عملکرد و کیفیت سری Gemini 3.
ایمیج فلش Gemini 3.1 ( نانو موز ۲ )
gemini-3.1-flash-image
مدل قدرتمند و با راندمان بالا برای تولید و ویرایش تصویر، بهینه شده برای سرعت و موارد استفاده با حجم بالا.
مدلهای عمومی
جمینی ۳.۱ پرو
gemini-3.1-pro-preview
هوش پیشرفته، مهارتهای حل مسئله پیچیده و قابلیتهای قدرتمند کدنویسی عاملی و ارتعاشی.
فلش جمینی ۳.۷
gemini-3.7-flash
عملکردی در سطح کلاس Frontier که با کسری از قیمت، مدلهای بزرگتر را به رقابت میخواند.
جمینی ۳.۵ فلش-لایت
gemini-3.5-flash-lite
مدلی کارآمد، مقرونبهصرفه و با حجم تولید بالا، با عملکرد و کیفیت سری Gemini 3.
مدلهای قدیمیتر و پایدار برای استفاده عمومی
Gemini 3.6 Flash (
gemini-3.6-flash): مدل قبلی Gemini 3.x Flash برای عملکرد در سطح بالا که با کسری از قیمت، مدلهای بزرگتر را رقابت میکند. (نیازی به پرداخت صورتحساب نیست )Gemini 3.5 Flash (
gemini-3.5-flash): مدل قبلی Gemini 3.x Flash برای عملکرد در سطح بالا که با کسری از قیمت، مدلهای بزرگتر را رقابت میکند. (نیازی به پرداخت صورتحساب نیست )Gemini 3.1 Flash‑Lite (
gemini-3.1-flash-lite): مدل قبلی Gemini 3.x Flash‑Lite برای کارهای سنگین و حساس به هزینه. (نیازی به پرداخت صورتحساب نیست )
مدلهای تولید تصویر
تصویر Gemini 3 Pro ( نانو موز پرو )
gemini-3-pro-image
مدل پیشرفته تولید و ویرایش تصویر برای خلق تصاویر بومی با بافت بسیار بالا.
ایمیج فلش Gemini 3.1 ( نانو موز ۲ )
gemini-3.1-flash-image
مدل قدرتمند و با راندمان بالا برای تولید و ویرایش تصویر، بهینه شده برای سرعت و موارد استفاده با حجم بالا.
ایمیج Gemini 3.1 Flash-Lite ( نانو موز ۲ لایت )
gemini-3.1-flash-lite-image
مدل تولید و ویرایش تصویر با تأخیر بسیار کم و مقرونبهصرفه، طراحیشده برای موارد استفاده تعاملی با حجم بالا.
مدلهای تولید صدا
مدلهای تبدیل متن به گفتار (TTS)
شما میتوانید با مدلهای Gemini TTS از ورودی متن، گفتار تولید کنید.
جمینی ۳.x فلش TTS
gemini-3.1-flash-tts-preview
تولید گفتار قدرتمند و با تأخیر کم از ورودی متن.
مدلهای Live API
شما میتوانید با مدلهایی که از Gemini Live API پشتیبانی میکنند، صدای استریمشدهی دوطرفه تولید کنید.
فلش Gemini 3.x با صدای بومی Gemini Live API
رابط برنامهنویسی کاربردی توسعهدهندگان جمینی:
gemini-3.1-flash-live-preview
API پلتفرم عامل Gemini:
پشتیبانی نمیشود
تعاملات صوتی و تصویری با تأخیر کم و بلادرنگ را با مدل Gemini که دو طرفه است، امکانپذیر میکند.
فلش Gemini 2.5 با صدای بومی Gemini Live API
رابط برنامهنویسی کاربردی توسعهدهندگان جمینی:
gemini-2.5-flash-native-audio-preview-12-2025
API پلتفرم عامل Gemini:
gemini-live-2.5-flash-native-audio
تعاملات صوتی و تصویری با تأخیر کم و بلادرنگ را با مدل Gemini که دو طرفه است، امکانپذیر میکند.
ادامهی این صفحه اطلاعات دقیقی در مورد مدلهای پشتیبانیشده توسط Firebase AI Logic ارائه میدهد.
- ورودی و خروجی پشتیبانی شده
- مقایسه سطح بالا از قابلیتهای پشتیبانیشده
- مشخصات و محدودیتها، برای مثال حداکثر توکنهای ورودی یا حداکثر طول ویدیوی ورودی
شرح نحوهی نسخهبندی مدلها ، به ویژه نسخههای پایدار ، پیشنمایش و آزمایشی آنها
فهرست نام مدلهای موجود برای گنجاندن در کد شما در هنگام مقداردهی اولیه
لیست زبانهای پشتیبانیشده برای مدلها
در پایین این صفحه، میتوانید اطلاعات دقیقی در مورد مدلهای نسل قبلی مشاهده کنید .
مقایسه مدلها
هر مدل قابلیتهای متفاوتی برای پشتیبانی از موارد استفاده مختلف دارد. توجه داشته باشید که هر یک از جداول این بخش، هر مدل را هنگام استفاده با Firebase AI Logic شرح میدهند. هر مدل ممکن است قابلیتهای اضافی داشته باشد که هنگام استفاده از SDK های ما در دسترس نیستند.
اگر اطلاعات مورد نظر خود را در زیربخشهای زیر پیدا نکردید، میتوانید اطلاعات بیشتری را در مستندات ارائهدهنده API انتخابی خود بیابید: Gemini Developer API یاAgent Platform Gemini API (که قبلاً Vertex AI نام داشت) .
Supported input and output
The following table lists the supported input and output types when using each model with Firebase AI Logic .
To learn about supported file types, see Supported input files and requirements .
| Gemini 3.x Pro, Flash, Flash‑Lite | Gemini 3.x Pro Image | Gemini 3.x Flash Image | Gemini 3.x Flash‑Lite Image | Gemini 3.x Flash TTS | Gemini Live | |
|---|---|---|---|---|---|---|
| Input types | ||||||
| متن | ||||||
| کد | ||||||
| اسناد (PDFs or plain-text) | ||||||
| تصاویر | ||||||
| ویدئو | ||||||
| صوتی | ||||||
| انواع خروجی | ||||||
| متن | ||||||
| Text (streaming) | ||||||
| کد | ||||||
| خروجی ساختاریافته (like JSON) | ||||||
| تصاویر | ||||||
| ویدئو | ||||||
| صوتی | ||||||
Supported capabilities and features
The following table lists the supported capabilities and features when using each model with Firebase AI Logic .
| Gemini 3.x Pro, Flash, Flash‑Lite | Gemini 3.x Pro Image | Gemini 3.x Flash Image | Gemini 3.x Flash‑Lite Image | Gemini 3.x Flash TTS | Gemini Live | |
|---|---|---|---|---|---|---|
| تفکر | ||||||
| Generate text from text-only or multimodal inputs | interleaved or as part of image | interleaved or as part of image | interleaved or as part of image | transcription of audio | ||
| Generate images | ||||||
| Edit images | ||||||
| Generate audio | گفتار | streamed audio | ||||
| Generate structured output (like JSON) | ||||||
| Analyze documents (PDFs or plain-text) ( text-output | image-output ) | ||||||
| Analyze images ( text-output | image-output ) | ||||||
| Analyze video ( text-output | image-output ) | streamed video | |||||
| Analyze audio | streamed audio | |||||
| Multi-turn chat | ||||||
| Bidirectional multimodal streaming | ||||||
| ابزارهای پشتیبانی شده | ||||||
| فراخوانی تابع | ||||||
| اجرای کد | ||||||
| زمینه URL | ||||||
| اتصال به زمین با | ||||||
| اتصال به زمین با | ||||||
Specifications and limitations
The following table lists the specifications and limitations when using each model with Firebase AI Logic .
| ملک | Gemini 3.x Pro, Flash, Flash‑Lite | Gemini 3.x Pro Image | Gemini 3.x Flash Image | Gemini 3.x Flash‑Lite Image | Gemini 3.x Flash TTS | Gemini Live |
|---|---|---|---|---|---|---|
| Input token limit * | 1,048,576 tokens | 65,536 tokens | 131,072 tokens | 65,536 tokens | 8,192 tokens | see docs |
| Output token limit * | 65,536 tokens | ۳۲۷۶۸ توکن | ۳۲۷۶۸ توکن | 4,096 tokens | 16,384 tokens | |
| PDFs (per request) | ||||||
| Max number of input PDF files ** | 900 files | 14 files | 14 files | 14 files | --- | see docs |
| Max number of pages per input PDF file ** | 900 pages | ۱۴ صفحه | ۱۴ صفحه | ۱۴ صفحه | --- | |
| حداکثر اندازه per input PDF file | ۵۰ مگابایت | ۵۰ مگابایت | ۵۰ مگابایت | ۵۰ مگابایت | --- | |
| Images (per request) | ||||||
| Max number of input images | 1,000 images | 14 images | 14 images | 14 images | --- | see docs |
| حداکثر اندازه per input base64-encoded image | 7 MB | 7 MB | 7 MB | 7 MB | --- | |
| Max number of output images | --- | Up to output token limit | Up to output token limit | Up to output token limit | --- | |
| Video (per request) | ||||||
| Max number of input video files | ۱۰ فایل | --- | Up to input token limit | Up to input token limit | --- | see docs |
| Max length of all input video (frames only) | ~60 minutes | --- | ~25 minutes | ~12 minutes | --- | |
| Max length of all input video (frames+audio) | ~45 minutes | --- | --- | --- | --- | |
| Audio (per request) | ||||||
| Max number of input audio files | 1 file | --- | --- | --- | --- | see docs |
| Max length of all input audio | ~8.4 hours | --- | --- | --- | --- | |
* For all Gemini models, a token is equivalent to about 4 characters, so 100 tokens are about 60-80 English words. For Gemini models, you can determine the total count of tokens in your requests using countTokens .
** PDFs are treated as images, so a single page of a PDF is treated as one image. The number of pages allowed in a request is limited to the number of images the model can support.
Find additional detailed information
Quotas and pricing are different for each model. Pricing also depends on input and output.
Learn about supported input file types, how to specify MIME type, and how to make sure that your input files and multimodal requests meet the requirements and follow best practices in Supported input files and requirements .
Model versioning and naming patterns
Models are offered in stable , preview , and experimental versions. For convenience, the-latest aliases are supported, but they're not recommended due to the inconsistency in which model version they point to.
Make sure to review our best practices for using model versions .
To find specific model names to use in your code, see the "available model names" section later on this page.
| Version type / Release stage | توضیحات | Model name pattern | |
|---|---|---|---|
| پایدار | Stable versions are available and supported for production use starting on the release date.
| Model names of stable versions have no suffix مثال: | |
| پیشنمایش | Preview versions have new capabilities and are considered not stable .
| Model names of preview versions are appended with مثال: | |
| تجربی | Experimental versions have new capabilities and are considered not stable .
| Model names of experimental versions are appended with مثال: | |
| Shutdown (retired) | Shutdown (retired) versions are past their shutdown (retirement) date and have been permanently deactivated.
| --- | |
Best practices for using model versions
In your production apps , use the explicit model name for the most recent stable version.
For Agent Platform Gemini API (formerly Vertex AI) , if you choose to use a short-term availability model in your production app, it's even more critical that you use Firebase Remote Config or server prompt templates to control the model name used for your AI feature.
Use preview and experimental versions only during prototyping . We recommend using a stable version when you start developing and testing for a production use case.
We do not recommend using the
-latestaliases (even during development). This alias points to the latest release for a specific model variation (which could be a stable, preview, or experimental version). This alias will get hot-swapped with every new release of a specific model variation, and only for breaking changes will a 2-week-prior notification email be sent. This instability of which model you're actually using can lead to unexpected behavior changes for your AI feature.
Available model names
Model names are the explicit values that you include in your code during initialization of the model.
General-use models (like
gemini-3.7-flash)Image-generating models (like
gemini-3.1-flash-image, aka the "Nano Banana" models)Audio-generating models :
- Text-to-speech (TTS) models (like
gemini-3.1-flash-tts-preview) - Live API models (like
gemini-live-2.5-flash-native-audio)
- Text-to-speech (TTS) models (like
For initialization examples for your platform, see the getting started guide .
For details about the release stages (especially for use cases, billing, and shutdown), see model versioning and naming patterns .
Programmatically list all available models
You can list all available models names using the REST API:
Gemini Developer API : Call the
models.listendpointAgent Platform Gemini API (formerly Vertex AI) : Call the
publishers.models.listendpoint
Note that this returned list will include all models supported by the API providers, but Firebase AI Logic only supports the Gemini models described on this page.
General-use models
Gemini 3.x Pro model names
Requires the pay-as-you-go Blaze pricing plan regardless of your Gemini API provider.
| نام مدل | توضیحات | Release stage | تاریخ انتشار | تاریخ خاموش شدن |
|---|---|---|---|---|
gemini-3.1-pro-preview | Latest preview version of Gemini 3.x Pro | پیشنمایش | ۲۰۲۶-۰۲-۱۹ | To be determined |
Gemini 3.x Flash model names
Does not require the pay-as-you-go Blaze pricing plan if you're using the Gemini Developer API .
| نام مدل | توضیحات | Release stage | تاریخ انتشار | تاریخ خاموش شدن |
|---|---|---|---|---|
gemini-3.7-flash | Latest stable version of Gemini 3.x Flash This is a short-term availability model (see Model versioning and naming patterns ). | پایدار | ۲۰۲۶-۰۸-۱۳ | To be determined |
gemini-3.6-flash | Stable version of Gemini 3.x Flash This is a short-term availability model (see Model versioning and naming patterns ). | پایدار | ۲۰۲۶-۰۷-۲۱ | To be determined |
gemini-3.5-flash | Stable version of Gemini 3.x Flash | پایدار | ۲۰۲۶-۰۵-۱۹ | No earlier than 2027-05-19 |
gemini-3-flash-preview | Preview version of Gemini 3.x Flash | پیشنمایش | ۲۰۲۵-۱۲-۱۷ | To be determined |
Gemini 3.x Flash‑Lite model names
Does not require the pay-as-you-go Blaze pricing plan if you're using the Gemini Developer API .
| نام مدل | توضیحات | Release stage | تاریخ انتشار | تاریخ خاموش شدن |
|---|---|---|---|---|
gemini-3.5-flash-lite | Latest stable version of Gemini 3.x Flash‑Lite | پایدار | ۲۰۲۶-۰۷-۲۱ | No earlier than 2027-07-21 |
gemini-3.1-flash-lite | Stable version of Gemini 3.x Flash‑Lite | پایدار | ۲۰۲۶-۰۵-۰۷ | No earlier than 2027-05-07 |
Image-generating models
Gemini 3.x Pro Image model names (aka "Nano Banana Pro")
Requires the pay-as-you-go Blaze pricing plan regardless of your Gemini API provider.
| نام مدل | توضیحات | Release stage | تاریخ انتشار | تاریخ خاموش شدن |
|---|---|---|---|---|
gemini-3-pro-image | Stable version of Gemini 3.x Pro Image (aka "Nano Banana Pro") | پایدار | ۲۰۲۶-۰۵-۲۸ | No earlier than 2027-05-28 |
gemini-3-pro-image-preview | Preview version of Gemini 3.x Pro Image (aka "Nano Banana Pro") | پیشنمایش | ۲۰۲۵-۱۱-۲۰ | As early as ۲۰۲۶-۰۶-۲۵ |
Gemini 3.x Flash Image model names (aka "Nano Banana 2")
Requires the pay-as-you-go Blaze pricing plan regardless of your Gemini API provider.
| نام مدل | توضیحات | Release stage | تاریخ انتشار | تاریخ خاموش شدن |
|---|---|---|---|---|
gemini-3.1-flash-image | Stable version of Gemini 3.x Flash Image (aka "Nano Banana 2") | پایدار | ۲۰۲۶-۰۵-۲۸ | No earlier than 2027-05-28 |
gemini-3.1-flash-image-preview | Preview version of Gemini 3.x Flash Image (aka "Nano Banana 2") | پیشنمایش | ۲۶-۰۲-۲۰۲۶ | As early as ۲۰۲۶-۰۶-۲۵ |
Gemini 3.x Flash‑Lite Image model names (aka "Nano Banana 2 Lite")
Requires the pay-as-you-go Blaze pricing plan regardless of your Gemini API provider.
| نام مدل | توضیحات | Release stage | تاریخ انتشار | تاریخ خاموش شدن |
|---|---|---|---|---|
gemini-3.1-flash-lite-image | Stable version of Gemini 3.x Flash‑Lite Image (aka "Nano Banana 2 Lite") This is a short-term availability model (see Model versioning and naming patterns ). | پایدار | ۲۰۲۶-۰۶-۳۰ | To be determined |
Audio - text-to-speech (TTS) models
Gemini 3.x Flash TTS model names
Does not require the pay-as-you-go Blaze pricing plan if you're using the Gemini Developer API .
| نام مدل | توضیحات | Release stage | تاریخ انتشار | تاریخ خاموش شدن |
|---|---|---|---|---|
gemini-3.1-flash-tts-preview | Preview version for Gemini 3.x Flash TTS | پیشنمایش | 2026-06-17 | To be determined |
Audio - Live API models
Gemini 3.x Flash Live model names
| رابط برنامهنویسی کاربردی (API) توسعهدهندگان جمینی نام مدل | توضیحات | Release stage | تاریخ انتشار | تاریخ خاموش شدن |
|---|---|---|---|---|
gemini-3.1-flash-live-preview 1 | Preview version for the Live API on the Gemini Developer API | پیشنمایش | ۲۰۲۶-۰۳-۲۶ | To be determined |
| Agent Platform Gemini API (formerly Vertex AI) نام مدل | توضیحات | Release stage | تاریخ انتشار | تاریخ خاموش شدن |
|---|---|---|---|---|
| The Agent Platform Gemini API (formerly Vertex AI) does not support any Gemini Live 3.x models. | ||||
1 Only supported by the Gemini Developer API . Also, even though this is a preview model, it's available on the "free tier" of the Gemini Developer API .
Gemini 2.5 Flash Live model names
Does not require the pay-as-you-go Blaze pricing plan if you're using the Gemini Developer API (usually preview models require a paid plan).
Even though the following models have different model names depending on the Gemini API provider, the features of the model are the same.
| رابط برنامهنویسی کاربردی (API) توسعهدهندگان جمینی نام مدل | توضیحات | Release stage | تاریخ انتشار | تاریخ خاموش شدن |
|---|---|---|---|---|
gemini-2.5-flash-native-audio-preview-12-2025 1 | Preview version for the Live API on the Gemini Developer API | پیشنمایش | ۲۰۲۵-۱۲-۱۲ | To be determined |
gemini-2.5-flash-native-audio-preview-09-2025 1 | Initial preview version for the Live API on the Gemini Developer API | پیشنمایش | ۲۰۲۵-۰۹-۱۸ | To be determined |
| Agent Platform Gemini API (formerly Vertex AI) نام مدل | توضیحات | Release stage | تاریخ انتشار | تاریخ خاموش شدن |
|---|---|---|---|---|
gemini-live-2.5-flash-native-audio 2 | Stable version for the Live API on the Agent Platform Gemini API (formerly Vertex AI) | پایدار | ۲۰۲۵-۱۲-۱۲ | No earlier than 2026-12-12 |
gemini-live-2.5-flash-preview-native-audio-09-2025 2 | Preview version for the Live API on the Agent Platform Gemini API (formerly Vertex AI) | پیشنمایش | ۲۰۲۵-۰۹-۱۸ | To be determined |
1 Only supported by the Gemini Developer API . Also, even though these are preview models, they're available on the "free tier" of the Gemini Developer API .
2 Only supported by the Agent Platform Gemini API (formerly Vertex AI) . Also, these models are not available in the global location.
زبانهای پشتیبانیشده
All the Gemini models can understand and respond in the following languages:
Arabic (ar), Bengali (bn), Bulgarian (bg), Chinese simplified and traditional (zh), Croatian (hr), Czech (cs), Danish (da), Dutch (nl), English (en), Estonian (et), Finnish (fi), French (fr), German (de), Greek (el), Hebrew (iw), Hindi (hi), Hungarian (hu), Indonesian (id), Italian (it), Japanese (ja), Korean (ko), Latvian (lv), Lithuanian (lt), Norwegian (no), Polish (pl), Portuguese (pt), Romanian (ro), Russian (ru), Serbian (sr), Slovak (sk), Slovenian (sl), Spanish (es), Swahili (sw), Swedish (sv), Thai (th), Turkish (tr), Ukrainian (uk), Vietnamese (vi)
Gemini 2.0 Flash , Gemini 1.5 Pro and Gemini 1.5 Flash models can understand and respond in the following additional languages:
Afrikaans (af), Amharic (am), Assamese (as), Azerbaijani (az), Belarusian (be), Bosnian (bs), Catalan (ca), Cebuano (ceb), Corsican (co), Welsh (cy), Dhivehi (dv), Esperanto (eo), Basque (eu), Persian (fa), Filipino (Tagalog) (fil), Frisian (fy), Irish (ga), Scots Gaelic (gd), Galician (gl), Gujarati (gu), Hausa (ha), Hawaiian (haw), Hmong (hmn), Haitian Creole (ht), Armenian (hy), Igbo (ig), Icelandic (is), Javanese (jv), Georgian (ka), Kazakh (kk), Khmer (km), Kannada (kn), Krio (kri), Kurdish (ku), Kyrgyz (ky), Latin (la), Luxembourgish (lb), Lao (lo), Malagasy (mg), Maori (mi), Macedonian (mk), Malayalam (ml), Mongolian (mn), Meiteilon (Manipuri) (mni-Mtei), Marathi (mr), Malay (ms), Maltese (mt), Myanmar (Burmese) (my), Nepali (ne), Nyanja (Chichewa) (ny), Odia (Oriya) (or), Punjabi (pa), Pashto (ps), Sindhi (sd), Sinhala (Sinhalese) (si), Samoan (sm), Shona (sn), Somali (so), Albanian (sq), Sesotho (st), Sundanese (su), Tamil (ta), Telugu (te), Tajik (tg), Uyghur (ug), Urdu (ur), Uzbek (uz), Xhosa (xh), Yiddish (yi), Yoruba (yo), Zulu (zu)
Information about previous models
The following are active, but previous generation models. We recommend using one of the latest models instead when possible.
If you can't find the information you're looking for in the following sub-sections, you can find even more information in your chosen API provider documentation: Gemini Developer API orAgent Platform Gemini API (formerly Vertex AI)
Older Gemini models
-
gemini-2.5-pro -
gemini-2.5-flash -
gemini-2.5-flash-lite -
gemini-2.5-flash-image(aka "Nano Banana") -
gemini-2.0-flash-001(and its auto-updated aliasgemini-2.0-flash) -
gemini-2.0-flash-lite-001(and its auto-updated aliasgemini-2.0-flash-lite)
For information about older Gemini Live API models, see the Gemini API provider documentation:
Older Imagen models
-
imagen-4.0-ultra-generate-001 -
imagen-4.0-generate-001 -
imagen-4.0-fast-generate-001 -
imagen-3.0-capability-001 -
imagen-3.0-generate-002 -
imagen-3.0-generate-001 -
imagen-3.0-fast-generate-001
View details about about previous models
These are the input and output types when using each model with Firebase AI Logic :
| Gemini 2.5 Pro, Flash, Flash‑Lite | Gemini 2.5 Flash Image | Gemini 2.0 Flash | Gemini 2.0 Flash‑Lite | ایمیجِن (generate) | ایمیجِن (capability) | |
|---|---|---|---|---|---|---|
| Input types | ||||||
| متن | ||||||
| کد | ||||||
| اسناد (PDFs or plain-text) | ||||||
| تصاویر | ||||||
| ویدئو | ||||||
| صوتی | ||||||
| Audio (streaming) | ||||||
| انواع خروجی | ||||||
| متن | ||||||
| Text (streaming) | ||||||
| کد | ||||||
| خروجی ساختاریافته (like JSON) | ||||||
| تصاویر | ||||||
| ویدئو | ||||||
| صوتی | ||||||
| Audio (streaming) | ||||||
These are the capabilities and features when using each model with Firebase AI Logic :
| Gemini 2.5 Pro, Flash, Flash‑Lite | Gemini 2.5 Flash Image | Gemini 2.0 Flash | Gemini 2.0 Flash‑Lite | ایمیجِن (generate) | ایمیجِن (capability) | |
|---|---|---|---|---|---|---|
| تفکر | ||||||
| Generate text from text-only or multimodal inputs | interleaved or as part of image | |||||
| Generate images ( Gemini or Imagen ) | ||||||
| Edit images ( Gemini or Imagen ) | ||||||
| Generate audio | ||||||
| Generate structured output (like JSON) | ||||||
| Analyze documents (PDFs or plain-text) ( text-output | image-output ) | ||||||
| Analyze images ( text-output | image-output ) | ||||||
| Analyze video ( text-output | image-output ) | ||||||
| Analyze audio | ||||||
| Multi-turn chat | ||||||
| Bidirectional multimodal streaming | ||||||
| ابزارهای پشتیبانی شده | ||||||
| فراخوانی تابع | ||||||
| اجرای کد | ||||||
| زمینه URL | ||||||
| اتصال به زمین با | ||||||
| اتصال به زمین با | ||||||
| دستورالعملهای سیستم | ||||||
| تعداد توکنها | ||||||
These are the specifications and limitations when using each model with Firebase AI Logic :
| ملک | Gemini 2.5 Pro, Flash, Flash‑Lite | Gemini 2.5 Flash Image | Gemini 2.0 Flash | Gemini 2.0 Flash‑Lite | ایمیجِن (generate) | ایمیجِن (capability) |
|---|---|---|---|---|---|---|
| Input token limit * | 1,048,576 tokens | ۳۲۷۶۸ توکن | 1,048,576 tokens | 1,048,576 tokens | 480 tokens | 480 tokens |
| Output token limit * | 65,536 tokens | 8,192 tokens | 8,192 tokens | 8,192 tokens | --- | --- |
| Knowledge cutoff date | ژانویه ۲۰۲۵ | --- | ژوئن ۲۰۲۴ | ژوئن ۲۰۲۴ | --- | --- |
| PDFs (per request) | ||||||
| Max number of input PDF files ** | ۳۰۰۰ فایل | ۳ فایل | ۳۰۰۰ فایل | ۳۰۰۰ فایل | --- | --- |
| Max number of pages per input PDF file ** | 1,000 pages | ۳ صفحه | 1,000 pages | 1,000 pages | --- | --- |
| حداکثر اندازه per input PDF file | ۵۰ مگابایت | ۵۰ مگابایت | ۵۰ مگابایت | ۵۰ مگابایت | --- | --- |
| Images (per request) | ||||||
| Max number of input images | 3,000 images | 3 images | 3,000 images | 3,000 images | --- | 4 images |
| Max number of output images | --- | Up to output token limit | --- | --- | 4 images | 4 images |
| حداکثر اندازه per input base64-encoded image | 7 MB | 7 MB | 7 MB | 7 MB | --- | --- |
| Video (per request) | ||||||
| Max number of input video files | ۱۰ فایل | --- | ۱۰ فایل | ۱۰ فایل | --- | --- |
| Max length of all input video (frames only) | ~60 minutes | --- | ~60 minutes | ~60 minutes | --- | --- |
| Max length of all input video (frames+audio) | ~45 minutes | --- | ~45 minutes | ~45 minutes | --- | --- |
| Audio (per request) | ||||||
| Max number of input audio files | 1 file | --- | 1 file | 1 file | --- | --- |
| Max number of output audio files | --- | --- | --- | --- | --- | --- |
| Max length of all input audio | ~8.4 hours | --- | ~8.4 hours | ~8.4 hours | --- | --- |
| Max length of all output audio | --- | --- | --- | --- | --- | --- |
* For all Gemini models, a token is equivalent to about 4 characters, so 100 tokens are about 60-80 English words. For Gemini models, you can determine the total count of tokens in your requests using countTokens .
** PDFs are treated as images, so a single page of a PDF is treated as one image. The number of pages allowed in a request is limited to the number of images the model can support.
Model names are the explicit values that you include in your code during initialization of the model.
مدلهای جمینی
Gemini 2.5 Pro model names
Does not require the pay-as-you-go Blaze pricing plan if you're using the Gemini Developer API .
| نام مدل | توضیحات | Release stage | تاریخ انتشار | تاریخ خاموش شدن |
|---|---|---|---|---|
gemini-2.5-pro | Stable version of Gemini 2.5 Pro | پایدار | ۲۰۲۵-۰۶-۱۷ | As early as 2026-10-16 |
Gemini 2.5 Flash model names
Does not require the pay-as-you-go Blaze pricing plan if you're using the Gemini Developer API .
| نام مدل | توضیحات | Release stage | تاریخ انتشار | تاریخ خاموش شدن |
|---|---|---|---|---|
gemini-2.5-flash | Stable version of Gemini 2.5 Flash | پایدار | ۲۰۲۵-۰۶-۱۷ | As early as 2026-10-16 |
Gemini 2.5 Flash‑Lite model names
Does not require the pay-as-you-go Blaze pricing plan if you're using the Gemini Developer API .
| نام مدل | توضیحات | Release stage | تاریخ انتشار | تاریخ خاموش شدن |
|---|---|---|---|---|
gemini-2.5-flash-lite | Stable version of Gemini 2.5 Flash‑Lite | پایدار | ۲۰۲۵-۰۷-۲۲ | As early as 2026-10-16 |
Gemini 2.5 Flash Image model names (aka "Nano Banana")
Requires the pay-as-you-go Blaze pricing plan regardless of your Gemini API provider.
| نام مدل | توضیحات | Release stage | تاریخ انتشار | تاریخ خاموش شدن |
|---|---|---|---|---|
gemini-2.5-flash-image | Stable version of Gemini 2.5 Flash Image (aka "Nano Banana") | پایدار | ۲۰۲۵-۱۰-۰۲ | ۲۰۲۶-۱۰-۰۲ |
Gemini 2.0 Flash model names
| نام مدل | توضیحات | Release stage | تاریخ انتشار | تاریخ خاموش شدن |
|---|---|---|---|---|
gemini-2.0-flash-001 | Latest stable version of Gemini 2.0 Flash | پایدار | ۲۰۲۵-۰۲-۰۵ | 2026-06-01 |
gemini-2.0-flash | Auto-updated alias pointing to the latest stable version of Gemini 2.0 Flash (currently gemini-2.0-flash-001 ) | پایدار | 2025-02-10 | 2026-06-01 |
Gemini 2.0 Flash‑Lite model names
| نام مدل | توضیحات | Release stage | تاریخ انتشار | تاریخ خاموش شدن |
|---|---|---|---|---|
gemini-2.0-flash-lite-001 | Latest stable version of Gemini 2.0 Flash‑Lite | پایدار | 2025-02-25 | 2026-06-01 |
gemini-2.0-flash-lite | Auto-updated alias pointing to the latest stable version of Gemini 2.0 Flash‑Lite (currently gemini-2.0-flash-lite-001 ) | پایدار | 2025-02-25 | 2026-06-01 |
مدلهای ایمیجن
Imagen 4 model names
| نام مدل | توضیحات | Release stage | تاریخ انتشار | تاریخ خاموش شدن |
|---|---|---|---|---|
imagen-4.0-generate-001 | Stable version of Imagen 4 | پایدار | ۱۴-۰۸-۲۰۲۵ | ۲۰۲۶-۰۶-۳۰ |
Imagen 4 Fast model names
| نام مدل | توضیحات | Release stage | تاریخ انتشار | تاریخ خاموش شدن |
|---|---|---|---|---|
imagen-4.0-fast-generate-001 | Stable version of Imagen 4 Fast | پایدار | ۱۴-۰۸-۲۰۲۵ | ۲۰۲۶-۰۶-۳۰ |
Imagen 4 Ultra model names
| نام مدل | توضیحات | Release stage | تاریخ انتشار | تاریخ خاموش شدن |
|---|---|---|---|---|
imagen-4.0-ultra-generate-001 | Stable version of Imagen 4 Ultra | پایدار | ۱۴-۰۸-۲۰۲۵ | ۲۰۲۶-۰۶-۳۰ |
Imagen 3 Capability model names
| نام مدل | توضیحات | Release stage | تاریخ انتشار | تاریخ خاموش شدن |
|---|---|---|---|---|
imagen-3.0-capability-001 | Initial stable version of Imagen 3 Capability | پایدار | ۲۰۲۴-۱۲-۱۰ | ۲۰۲۶-۰۶-۳۰ |
Imagen 3 model names
| نام مدل | توضیحات | Release stage | تاریخ انتشار | تاریخ خاموش شدن |
|---|---|---|---|---|
imagen-3.0-generate-002 | Latest stable version of Imagen 3 | پایدار | ۲۰۲۵-۰۱-۲۳ | ۲۰۲۶-۰۶-۳۰ |
imagen-3.0-generate-001 | Initial stable version of Imagen 3 | پایدار | 2024-07-31 | ۲۰۲۶-۰۶-۳۰ |
Imagen 3 Fast model names
| نام مدل | توضیحات | Release stage | تاریخ انتشار | تاریخ خاموش شدن |
|---|---|---|---|---|
imagen-3.0-fast-generate-001 | Initial stable version of Imagen 3 Fast | پایدار | 2024-07-31 | ۲۰۲۶-۰۶-۳۰ |
مراحل بعدی
Try out the capabilities of the Gemini API
- Build multi-turn conversations (chat) .
- Generate text from text-only prompts .
- Generate text by prompting with various file types, like images , PDFs , video , and audio .
- Generate structured output (like JSON) from both text and multimodal prompts.
- Generate and edit images from both text and multimodal prompts.
- Generate speech (both single- and multiple-speakers) using Gemini text-to-speech (TTS) models.
- Stream input and output (including audio) using the Gemini Live API .
- Use tools (like function calling and Grounding with
Google Search orGoogle Maps ) to connect a Gemini model to other parts of your app and external systems and information.
For mobile and web apps, the Firebase AI Logic SDKs let you interact with the supported Gemini models directly from your app.
Gemini models are considered multimodal because they're capable of processing and even generating multiple modalities, including text, code, PDFs, images, video, and audio.
Also, review our FAQ about all the models that Firebase AI Logic supports and does not support.
Featured models
فلش جمینی ۳.۷
gemini-3.7-flash
Frontier-class performance rivaling larger models at a fraction of the cost.
جمینی ۳.۵ فلش-لایت
gemini-3.5-flash-lite
High-volume, cost-sensitive workhorse model with the performance and quality of the Gemini 3 series.
Gemini 3.1 Flash Image ( Nano Banana 2 )
gemini-3.1-flash-image
Powerful, high-efficiency image generation and editing model, optimized for speed and high-volume use cases.
General-use models
Go to tables with model details
جمینی ۳.۱ پرو
gemini-3.1-pro-preview
Advanced intelligence, complex problem-solving skills, and powerful agentic and vibe coding capabilities.
فلش جمینی ۳.۷
gemini-3.7-flash
Frontier-class performance rivaling larger models at a fraction of the cost.
جمینی ۳.۵ فلش-لایت
gemini-3.5-flash-lite
High-volume, cost-sensitive workhorse model with the performance and quality of the Gemini 3 series.
Older stable general-use models
Gemini 3.6 Flash (
gemini-3.6-flash): Previous Gemini 3.x Flash model for frontier-class performance rivaling larger models at a fraction of the cost. (billing not required)Gemini 3.5 Flash (
gemini-3.5-flash): Previous Gemini 3.x Flash model for frontier-class performance rivaling larger models at a fraction of the cost. (billing not required)Gemini 3.1 Flash‑Lite (
gemini-3.1-flash-lite): Previous Gemini 3.x Flash‑Lite model for high-volume, cost-sensitive workhorse tasks. (billing not required)
Image-generating models
Go to tables with model details
Gemini 3 Pro Image ( Nano Banana Pro )
gemini-3-pro-image
State-of-the-art image generation and editing model for highly contextual native image creation.
Gemini 3.1 Flash Image ( Nano Banana 2 )
gemini-3.1-flash-image
Powerful, high-efficiency image generation and editing model, optimized for speed and high-volume use cases.
Gemini 3.1 Flash-Lite Image ( Nano Banana 2 Lite )
gemini-3.1-flash-lite-image
Ultra-low latency and cost-effective image generation and editing model, designed for high-volume interactive use cases.
Audio-generating models
Text-to-speech (TTS) models
You can generate speech from text input with Gemini TTS models.
Go to tables with model details
Gemini 3.x Flash TTS
gemini-3.1-flash-tts-preview
Powerful, low-latency speech generation from text input.
مدلهای Live API
You can generate bidirectional streamed audio with models that support the Gemini Live API .
Go to tables with model details
Gemini 3.x Flash with Gemini Live API native audio
Gemini Developer API:
gemini-3.1-flash-live-preview
Agent Platform Gemini API:
not supported
Enables low-latency, real-time voice and video interactions with a Gemini model that is bidirectional .
Gemini 2.5 Flash with Gemini Live API native audio
Gemini Developer API:
gemini-2.5-flash-native-audio-preview-12-2025
Agent Platform Gemini API:
gemini-live-2.5-flash-native-audio
Enables low-latency, real-time voice and video interactions with a Gemini model that is bidirectional .
The remainder of this page provides detailed information about the models supported by Firebase AI Logic .
- Supported input and output
- High-level comparison of the supported capabilities
- Specifications and limitations, for example max input tokens or max length of input video
Description of how models are versioned , specifically their stable , preview , and experimental versions
Lists of available model names to include in your code during initialization
Lists of supported languages for the models
At the bottom of this page, you can view detailed information about previous generation models .
Compare models
Each model has different capabilities to support various use cases. Note that each of tables in this section describe each model when used with Firebase AI Logic . Each model might have additional capabilities that aren't available when using our SDKs.
If you can't find the information you're looking for in the following sub-sections, you can find even more information in your chosen API provider documentation: Gemini Developer API orAgent Platform Gemini API (formerly Vertex AI) .
Supported input and output
The following table lists the supported input and output types when using each model with Firebase AI Logic .
To learn about supported file types, see Supported input files and requirements .
| Gemini 3.x Pro, Flash, Flash‑Lite | Gemini 3.x Pro Image | Gemini 3.x Flash Image | Gemini 3.x Flash‑Lite Image | Gemini 3.x Flash TTS | Gemini Live | |
|---|---|---|---|---|---|---|
| Input types | ||||||
| متن | ||||||
| کد | ||||||
| اسناد (PDFs or plain-text) | ||||||
| Images | ||||||
| ویدئو | ||||||
| صوتی | ||||||
| انواع خروجی | ||||||
| متن | ||||||
| Text (streaming) | ||||||
| کد | ||||||
| خروجی ساختاریافته (like JSON) | ||||||
| Images | ||||||
| ویدئو | ||||||
| صوتی | ||||||
Supported capabilities and features
The following table lists the supported capabilities and features when using each model with Firebase AI Logic .
| Gemini 3.x Pro, Flash, Flash‑Lite | Gemini 3.x Pro Image | Gemini 3.x Flash Image | Gemini 3.x Flash‑Lite Image | Gemini 3.x Flash TTS | Gemini Live | |
|---|---|---|---|---|---|---|
| تفکر | ||||||
| Generate text from text-only or multimodal inputs | interleaved or as part of image | interleaved or as part of image | interleaved or as part of image | transcription of audio | ||
| Generate images | ||||||
| Edit images | ||||||
| Generate audio | گفتار | streamed audio | ||||
| Generate structured output (like JSON) | ||||||
| Analyze documents (PDFs or plain-text) ( text-output | image-output ) | ||||||
| Analyze images ( text-output | image-output ) | ||||||
| Analyze video ( text-output | image-output ) | streamed video | |||||
| Analyze audio | streamed audio | |||||
| Multi-turn chat | ||||||
| Bidirectional multimodal streaming | ||||||
| ابزارهای پشتیبانی شده | ||||||
| فراخوانی تابع | ||||||
| Code execution | ||||||
| زمینه URL | ||||||
| اتصال به زمین با | ||||||
| اتصال به زمین با | ||||||
Specifications and limitations
The following table lists the specifications and limitations when using each model with Firebase AI Logic .
| ملک | Gemini 3.x Pro, Flash, Flash‑Lite | Gemini 3.x Pro Image | Gemini 3.x Flash Image | Gemini 3.x Flash‑Lite Image | Gemini 3.x Flash TTS | Gemini Live |
|---|---|---|---|---|---|---|
| Input token limit * | 1,048,576 tokens | 65,536 tokens | 131,072 tokens | 65,536 tokens | 8,192 tokens | see docs |
| Output token limit * | 65,536 tokens | ۳۲۷۶۸ توکن | ۳۲۷۶۸ توکن | 4,096 tokens | 16,384 tokens | |
| PDFs (per request) | ||||||
| Max number of input PDF files ** | 900 files | 14 files | 14 files | 14 files | --- | see docs |
| Max number of pages per input PDF file ** | 900 pages | ۱۴ صفحه | ۱۴ صفحه | ۱۴ صفحه | --- | |
| حداکثر اندازه per input PDF file | ۵۰ مگابایت | ۵۰ مگابایت | ۵۰ مگابایت | ۵۰ مگابایت | --- | |
| Images (per request) | ||||||
| Max number of input images | 1,000 images | 14 images | 14 images | 14 images | --- | see docs |
| حداکثر اندازه per input base64-encoded image | 7 MB | 7 MB | 7 MB | 7 MB | --- | |
| Max number of output images | --- | Up to output token limit | Up to output token limit | Up to output token limit | --- | |
| Video (per request) | ||||||
| Max number of input video files | ۱۰ فایل | --- | Up to input token limit | Up to input token limit | --- | see docs |
| Max length of all input video (frames only) | ~60 minutes | --- | ~25 minutes | ~12 minutes | --- | |
| Max length of all input video (frames+audio) | ~45 minutes | --- | --- | --- | --- | |
| Audio (per request) | ||||||
| Max number of input audio files | 1 file | --- | --- | --- | --- | see docs |
| Max length of all input audio | ~8.4 hours | --- | --- | --- | --- | |
* For all Gemini models, a token is equivalent to about 4 characters, so 100 tokens are about 60-80 English words. For Gemini models, you can determine the total count of tokens in your requests using countTokens .
** PDFs are treated as images, so a single page of a PDF is treated as one image. The number of pages allowed in a request is limited to the number of images the model can support.
Find additional detailed information
Quotas and pricing are different for each model. Pricing also depends on input and output.
Learn about supported input file types, how to specify MIME type, and how to make sure that your input files and multimodal requests meet the requirements and follow best practices in Supported input files and requirements .
Model versioning and naming patterns
Models are offered in stable , preview , and experimental versions. For convenience, the-latest aliases are supported, but they're not recommended due to the inconsistency in which model version they point to.
Make sure to review our best practices for using model versions .
To find specific model names to use in your code, see the "available model names" section later on this page.
| Version type / Release stage | توضیحات | Model name pattern | |
|---|---|---|---|
| پایدار | Stable versions are available and supported for production use starting on the release date.
| Model names of stable versions have no suffix مثال: | |
| پیشنمایش | Preview versions have new capabilities and are considered not stable .
| Model names of preview versions are appended with مثال: | |
| تجربی | Experimental versions have new capabilities and are considered not stable .
| Model names of experimental versions are appended with مثال: | |
| Shutdown (retired) | Shutdown (retired) versions are past their shutdown (retirement) date and have been permanently deactivated.
| --- | |
Best practices for using model versions
In your production apps , use the explicit model name for the most recent stable version.
For Agent Platform Gemini API (formerly Vertex AI) , if you choose to use a short-term availability model in your production app, it's even more critical that you use Firebase Remote Config or server prompt templates to control the model name used for your AI feature.
Use preview and experimental versions only during prototyping . We recommend using a stable version when you start developing and testing for a production use case.
We do not recommend using the
-latestaliases (even during development). This alias points to the latest release for a specific model variation (which could be a stable, preview, or experimental version). This alias will get hot-swapped with every new release of a specific model variation, and only for breaking changes will a 2-week-prior notification email be sent. This instability of which model you're actually using can lead to unexpected behavior changes for your AI feature.
Available model names
Model names are the explicit values that you include in your code during initialization of the model.
General-use models (like
gemini-3.7-flash)Image-generating models (like
gemini-3.1-flash-image, aka the "Nano Banana" models)Audio-generating models :
- Text-to-speech (TTS) models (like
gemini-3.1-flash-tts-preview) - Live API models (like
gemini-live-2.5-flash-native-audio)
- Text-to-speech (TTS) models (like
For initialization examples for your platform, see the getting started guide .
For details about the release stages (especially for use cases, billing, and shutdown), see model versioning and naming patterns .
Programmatically list all available models
You can list all available models names using the REST API:
Gemini Developer API : Call the
models.listendpointAgent Platform Gemini API (formerly Vertex AI) : Call the
publishers.models.listendpoint
Note that this returned list will include all models supported by the API providers, but Firebase AI Logic only supports the Gemini models described on this page.
General-use models
Gemini 3.x Pro model names
Requires the pay-as-you-go Blaze pricing plan regardless of your Gemini API provider.
| نام مدل | توضیحات | Release stage | تاریخ انتشار | تاریخ خاموش شدن |
|---|---|---|---|---|
gemini-3.1-pro-preview | Latest preview version of Gemini 3.x Pro | پیشنمایش | ۲۰۲۶-۰۲-۱۹ | To be determined |
Gemini 3.x Flash model names
Does not require the pay-as-you-go Blaze pricing plan if you're using the Gemini Developer API .
| نام مدل | توضیحات | Release stage | تاریخ انتشار | تاریخ خاموش شدن |
|---|---|---|---|---|
gemini-3.7-flash | Latest stable version of Gemini 3.x Flash This is a short-term availability model (see Model versioning and naming patterns ). | پایدار | ۲۰۲۶-۰۸-۱۳ | To be determined |
gemini-3.6-flash | Stable version of Gemini 3.x Flash This is a short-term availability model (see Model versioning and naming patterns ). | پایدار | ۲۰۲۶-۰۷-۲۱ | To be determined |
gemini-3.5-flash | Stable version of Gemini 3.x Flash | پایدار | ۲۰۲۶-۰۵-۱۹ | No earlier than 2027-05-19 |
gemini-3-flash-preview | Preview version of Gemini 3.x Flash | پیشنمایش | ۲۰۲۵-۱۲-۱۷ | To be determined |
Gemini 3.x Flash‑Lite model names
Does not require the pay-as-you-go Blaze pricing plan if you're using the Gemini Developer API .
| نام مدل | توضیحات | Release stage | تاریخ انتشار | تاریخ خاموش شدن |
|---|---|---|---|---|
gemini-3.5-flash-lite | Latest stable version of Gemini 3.x Flash‑Lite | پایدار | ۲۰۲۶-۰۷-۲۱ | No earlier than 2027-07-21 |
gemini-3.1-flash-lite | Stable version of Gemini 3.x Flash‑Lite | پایدار | ۲۰۲۶-۰۵-۰۷ | No earlier than 2027-05-07 |
Image-generating models
Gemini 3.x Pro Image model names (aka "Nano Banana Pro")
Requires the pay-as-you-go Blaze pricing plan regardless of your Gemini API provider.
| نام مدل | توضیحات | Release stage | تاریخ انتشار | تاریخ خاموش شدن |
|---|---|---|---|---|
gemini-3-pro-image | Stable version of Gemini 3.x Pro Image (aka "Nano Banana Pro") | پایدار | ۲۰۲۶-۰۵-۲۸ | No earlier than 2027-05-28 |
gemini-3-pro-image-preview | Preview version of Gemini 3.x Pro Image (aka "Nano Banana Pro") | پیشنمایش | ۲۰۲۵-۱۱-۲۰ | As early as ۲۰۲۶-۰۶-۲۵ |
Gemini 3.x Flash Image model names (aka "Nano Banana 2")
Requires the pay-as-you-go Blaze pricing plan regardless of your Gemini API provider.
| نام مدل | توضیحات | Release stage | تاریخ انتشار | تاریخ خاموش شدن |
|---|---|---|---|---|
gemini-3.1-flash-image | Stable version of Gemini 3.x Flash Image (aka "Nano Banana 2") | پایدار | ۲۰۲۶-۰۵-۲۸ | No earlier than 2027-05-28 |
gemini-3.1-flash-image-preview | Preview version of Gemini 3.x Flash Image (aka "Nano Banana 2") | پیشنمایش | ۲۶-۰۲-۲۰۲۶ | As early as ۲۰۲۶-۰۶-۲۵ |
Gemini 3.x Flash‑Lite Image model names (aka "Nano Banana 2 Lite")
Requires the pay-as-you-go Blaze pricing plan regardless of your Gemini API provider.
| نام مدل | توضیحات | Release stage | تاریخ انتشار | تاریخ خاموش شدن |
|---|---|---|---|---|
gemini-3.1-flash-lite-image | Stable version of Gemini 3.x Flash‑Lite Image (aka "Nano Banana 2 Lite") This is a short-term availability model (see Model versioning and naming patterns ). | پایدار | ۲۰۲۶-۰۶-۳۰ | To be determined |
Audio - text-to-speech (TTS) models
Gemini 3.x Flash TTS model names
Does not require the pay-as-you-go Blaze pricing plan if you're using the Gemini Developer API .
| نام مدل | توضیحات | Release stage | تاریخ انتشار | تاریخ خاموش شدن |
|---|---|---|---|---|
gemini-3.1-flash-tts-preview | Preview version for Gemini 3.x Flash TTS | پیشنمایش | 2026-06-17 | To be determined |
Audio - Live API models
Gemini 3.x Flash Live model names
| رابط برنامهنویسی کاربردی (API) توسعهدهندگان جمینی نام مدل | توضیحات | Release stage | تاریخ انتشار | تاریخ خاموش شدن |
|---|---|---|---|---|
gemini-3.1-flash-live-preview 1 | Preview version for the Live API on the Gemini Developer API | پیشنمایش | ۲۰۲۶-۰۳-۲۶ | To be determined |
| Agent Platform Gemini API (formerly Vertex AI) نام مدل | توضیحات | Release stage | تاریخ انتشار | تاریخ خاموش شدن |
|---|---|---|---|---|
| The Agent Platform Gemini API (formerly Vertex AI) does not support any Gemini Live 3.x models. | ||||
1 Only supported by the Gemini Developer API . Also, even though this is a preview model, it's available on the "free tier" of the Gemini Developer API .
Gemini 2.5 Flash Live model names
Does not require the pay-as-you-go Blaze pricing plan if you're using the Gemini Developer API (usually preview models require a paid plan).
Even though the following models have different model names depending on the Gemini API provider, the features of the model are the same.
| رابط برنامهنویسی کاربردی (API) توسعهدهندگان جمینی نام مدل | توضیحات | Release stage | تاریخ انتشار | تاریخ خاموش شدن |
|---|---|---|---|---|
gemini-2.5-flash-native-audio-preview-12-2025 1 | Preview version for the Live API on the Gemini Developer API | پیشنمایش | ۲۰۲۵-۱۲-۱۲ | To be determined |
gemini-2.5-flash-native-audio-preview-09-2025 1 | Initial preview version for the Live API on the Gemini Developer API | پیشنمایش | ۲۰۲۵-۰۹-۱۸ | To be determined |
| Agent Platform Gemini API (formerly Vertex AI) نام مدل | توضیحات | Release stage | تاریخ انتشار | تاریخ خاموش شدن |
|---|---|---|---|---|
gemini-live-2.5-flash-native-audio 2 | Stable version for the Live API on the Agent Platform Gemini API (formerly Vertex AI) | پایدار | ۲۰۲۵-۱۲-۱۲ | No earlier than 2026-12-12 |
gemini-live-2.5-flash-preview-native-audio-09-2025 2 | Preview version for the Live API on the Agent Platform Gemini API (formerly Vertex AI) | پیشنمایش | ۲۰۲۵-۰۹-۱۸ | To be determined |
1 Only supported by the Gemini Developer API . Also, even though these are preview models, they're available on the "free tier" of the Gemini Developer API .
2 Only supported by the Agent Platform Gemini API (formerly Vertex AI) . Also, these models are not available in the global location.
زبانهای پشتیبانیشده
All the Gemini models can understand and respond in the following languages:
Arabic (ar), Bengali (bn), Bulgarian (bg), Chinese simplified and traditional (zh), Croatian (hr), Czech (cs), Danish (da), Dutch (nl), English (en), Estonian (et), Finnish (fi), French (fr), German (de), Greek (el), Hebrew (iw), Hindi (hi), Hungarian (hu), Indonesian (id), Italian (it), Japanese (ja), Korean (ko), Latvian (lv), Lithuanian (lt), Norwegian (no), Polish (pl), Portuguese (pt), Romanian (ro), Russian (ru), Serbian (sr), Slovak (sk), Slovenian (sl), Spanish (es), Swahili (sw), Swedish (sv), Thai (th), Turkish (tr), Ukrainian (uk), Vietnamese (vi)
Gemini 2.0 Flash , Gemini 1.5 Pro and Gemini 1.5 Flash models can understand and respond in the following additional languages:
Afrikaans (af), Amharic (am), Assamese (as), Azerbaijani (az), Belarusian (be), Bosnian (bs), Catalan (ca), Cebuano (ceb), Corsican (co), Welsh (cy), Dhivehi (dv), Esperanto (eo), Basque (eu), Persian (fa), Filipino (Tagalog) (fil), Frisian (fy), Irish (ga), Scots Gaelic (gd), Galician (gl), Gujarati (gu), Hausa (ha), Hawaiian (haw), Hmong (hmn), Haitian Creole (ht), Armenian (hy), Igbo (ig), Icelandic (is), Javanese (jv), Georgian (ka), Kazakh (kk), Khmer (km), Kannada (kn), Krio (kri), Kurdish (ku), Kyrgyz (ky), Latin (la), Luxembourgish (lb), Lao (lo), Malagasy (mg), Maori (mi), Macedonian (mk), Malayalam (ml), Mongolian (mn), Meiteilon (Manipuri) (mni-Mtei), Marathi (mr), Malay (ms), Maltese (mt), Myanmar (Burmese) (my), Nepali (ne), Nyanja (Chichewa) (ny), Odia (Oriya) (or), Punjabi (pa), Pashto (ps), Sindhi (sd), Sinhala (Sinhalese) (si), Samoan (sm), Shona (sn), Somali (so), Albanian (sq), Sesotho (st), Sundanese (su), Tamil (ta), Telugu (te), Tajik (tg), Uyghur (ug), Urdu (ur), Uzbek (uz), Xhosa (xh), Yiddish (yi), Yoruba (yo), Zulu (zu)
Information about previous models
The following are active, but previous generation models. We recommend using one of the latest models instead when possible.
If you can't find the information you're looking for in the following sub-sections, you can find even more information in your chosen API provider documentation: Gemini Developer API orAgent Platform Gemini API (formerly Vertex AI)
Older Gemini models
-
gemini-2.5-pro -
gemini-2.5-flash -
gemini-2.5-flash-lite -
gemini-2.5-flash-image(aka "Nano Banana") -
gemini-2.0-flash-001(and its auto-updated aliasgemini-2.0-flash) -
gemini-2.0-flash-lite-001(and its auto-updated aliasgemini-2.0-flash-lite)
For information about older Gemini Live API models, see the Gemini API provider documentation:
Older Imagen models
-
imagen-4.0-ultra-generate-001 -
imagen-4.0-generate-001 -
imagen-4.0-fast-generate-001 -
imagen-3.0-capability-001 -
imagen-3.0-generate-002 -
imagen-3.0-generate-001 -
imagen-3.0-fast-generate-001
View details about about previous models
These are the input and output types when using each model with Firebase AI Logic :
| Gemini 2.5 Pro, Flash, Flash‑Lite | Gemini 2.5 Flash Image | Gemini 2.0 Flash | Gemini 2.0 Flash‑Lite | ایمیجِن (generate) | ایمیجِن (capability) | |
|---|---|---|---|---|---|---|
| Input types | ||||||
| متن | ||||||
| کد | ||||||
| اسناد (PDFs or plain-text) | ||||||
| Images | ||||||
| ویدئو | ||||||
| صوتی | ||||||
| Audio (streaming) | ||||||
| انواع خروجی | ||||||
| متن | ||||||
| Text (streaming) | ||||||
| کد | ||||||
| خروجی ساختاریافته (like JSON) | ||||||
| Images | ||||||
| ویدئو | ||||||
| صوتی | ||||||
| Audio (streaming) | ||||||
These are the capabilities and features when using each model with Firebase AI Logic :
| Gemini 2.5 Pro, Flash, Flash‑Lite | Gemini 2.5 Flash Image | Gemini 2.0 Flash | Gemini 2.0 Flash‑Lite | ایمیجِن (generate) | ایمیجِن (capability) | |
|---|---|---|---|---|---|---|
| تفکر | ||||||
| Generate text from text-only or multimodal inputs | interleaved or as part of image | |||||
| Generate images ( Gemini or Imagen ) | ||||||
| Edit images ( Gemini or Imagen ) | ||||||
| Generate audio | ||||||
| Generate structured output (like JSON) | ||||||
| Analyze documents (PDFs or plain-text) ( text-output | image-output ) | ||||||
| Analyze images ( text-output | image-output ) | ||||||
| Analyze video ( text-output | image-output ) | ||||||
| Analyze audio | ||||||
| Multi-turn chat | ||||||
| Bidirectional multimodal streaming | ||||||
| ابزارهای پشتیبانی شده | ||||||
| فراخوانی تابع | ||||||
| Code execution | ||||||
| زمینه URL | ||||||
| اتصال به زمین با | ||||||
| اتصال به زمین با | ||||||
| دستورالعملهای سیستم | ||||||
| تعداد توکنها | ||||||
These are the specifications and limitations when using each model with Firebase AI Logic :
| ملک | Gemini 2.5 Pro, Flash, Flash‑Lite | Gemini 2.5 Flash Image | Gemini 2.0 Flash | Gemini 2.0 Flash‑Lite | ایمیجِن (generate) | ایمیجِن (capability) |
|---|---|---|---|---|---|---|
| Input token limit * | 1,048,576 tokens | ۳۲۷۶۸ توکن | 1,048,576 tokens | 1,048,576 tokens | 480 tokens | 480 tokens |
| Output token limit * | 65,536 tokens | 8,192 tokens | 8,192 tokens | 8,192 tokens | --- | --- |
| Knowledge cutoff date | ژانویه ۲۰۲۵ | --- | ژوئن ۲۰۲۴ | ژوئن ۲۰۲۴ | --- | --- |
| PDFs (per request) | ||||||
| Max number of input PDF files ** | ۳۰۰۰ فایل | ۳ فایل | ۳۰۰۰ فایل | ۳۰۰۰ فایل | --- | --- |
| Max number of pages per input PDF file ** | 1,000 pages | ۳ صفحه | 1,000 pages | 1,000 pages | --- | --- |
| حداکثر اندازه per input PDF file | ۵۰ مگابایت | ۵۰ مگابایت | ۵۰ مگابایت | ۵۰ مگابایت | --- | --- |
| Images (per request) | ||||||
| Max number of input images | 3,000 images | 3 images | 3,000 images | 3,000 images | --- | 4 images |
| Max number of output images | --- | Up to output token limit | --- | --- | 4 images | 4 images |
| حداکثر اندازه per input base64-encoded image | 7 MB | 7 MB | 7 MB | 7 MB | --- | --- |
| Video (per request) | ||||||
| Max number of input video files | ۱۰ فایل | --- | ۱۰ فایل | ۱۰ فایل | --- | --- |
| Max length of all input video (frames only) | ~60 minutes | --- | ~60 minutes | ~60 minutes | --- | --- |
| Max length of all input video (frames+audio) | ~45 minutes | --- | ~45 minutes | ~45 minutes | --- | --- |
| Audio (per request) | ||||||
| Max number of input audio files | 1 file | --- | 1 file | 1 file | --- | --- |
| Max number of output audio files | --- | --- | --- | --- | --- | --- |
| Max length of all input audio | ~8.4 hours | --- | ~8.4 hours | ~8.4 hours | --- | --- |
| Max length of all output audio | --- | --- | --- | --- | --- | --- |
* For all Gemini models, a token is equivalent to about 4 characters, so 100 tokens are about 60-80 English words. For Gemini models, you can determine the total count of tokens in your requests using countTokens .
** PDFs are treated as images, so a single page of a PDF is treated as one image. The number of pages allowed in a request is limited to the number of images the model can support.
Model names are the explicit values that you include in your code during initialization of the model.
مدلهای جمینی
Gemini 2.5 Pro model names
Does not require the pay-as-you-go Blaze pricing plan if you're using the Gemini Developer API .
| نام مدل | توضیحات | Release stage | تاریخ انتشار | تاریخ خاموش شدن |
|---|---|---|---|---|
gemini-2.5-pro | Stable version of Gemini 2.5 Pro | پایدار | ۲۰۲۵-۰۶-۱۷ | As early as 2026-10-16 |
Gemini 2.5 Flash model names
Does not require the pay-as-you-go Blaze pricing plan if you're using the Gemini Developer API .
| نام مدل | توضیحات | Release stage | تاریخ انتشار | تاریخ خاموش شدن |
|---|---|---|---|---|
gemini-2.5-flash | Stable version of Gemini 2.5 Flash | پایدار | ۲۰۲۵-۰۶-۱۷ | As early as 2026-10-16 |
Gemini 2.5 Flash‑Lite model names
Does not require the pay-as-you-go Blaze pricing plan if you're using the Gemini Developer API .
| نام مدل | توضیحات | Release stage | تاریخ انتشار | تاریخ خاموش شدن |
|---|---|---|---|---|
gemini-2.5-flash-lite | Stable version of Gemini 2.5 Flash‑Lite | پایدار | ۲۰۲۵-۰۷-۲۲ | As early as 2026-10-16 |
Gemini 2.5 Flash Image model names (aka "Nano Banana")
Requires the pay-as-you-go Blaze pricing plan regardless of your Gemini API provider.
| نام مدل | توضیحات | Release stage | تاریخ انتشار | تاریخ خاموش شدن |
|---|---|---|---|---|
gemini-2.5-flash-image | Stable version of Gemini 2.5 Flash Image (aka "Nano Banana") | پایدار | ۲۰۲۵-۱۰-۰۲ | ۲۰۲۶-۱۰-۰۲ |
Gemini 2.0 Flash model names
| نام مدل | توضیحات | Release stage | تاریخ انتشار | تاریخ خاموش شدن |
|---|---|---|---|---|
gemini-2.0-flash-001 | Latest stable version of Gemini 2.0 Flash | پایدار | ۲۰۲۵-۰۲-۰۵ | 2026-06-01 |
gemini-2.0-flash | Auto-updated alias pointing to the latest stable version of Gemini 2.0 Flash (currently gemini-2.0-flash-001 ) | پایدار | 2025-02-10 | 2026-06-01 |
Gemini 2.0 Flash‑Lite model names
| نام مدل | توضیحات | Release stage | تاریخ انتشار | تاریخ خاموش شدن |
|---|---|---|---|---|
gemini-2.0-flash-lite-001 | Latest stable version of Gemini 2.0 Flash‑Lite | پایدار | 2025-02-25 | 2026-06-01 |
gemini-2.0-flash-lite | Auto-updated alias pointing to the latest stable version of Gemini 2.0 Flash‑Lite (currently gemini-2.0-flash-lite-001 ) | پایدار | 2025-02-25 | 2026-06-01 |
مدلهای ایمیجن
Imagen 4 model names
| نام مدل | توضیحات | Release stage | تاریخ انتشار | تاریخ خاموش شدن |
|---|---|---|---|---|
imagen-4.0-generate-001 | Stable version of Imagen 4 | پایدار | ۱۴-۰۸-۲۰۲۵ | ۲۰۲۶-۰۶-۳۰ |
Imagen 4 Fast model names
| نام مدل | توضیحات | Release stage | تاریخ انتشار | تاریخ خاموش شدن |
|---|---|---|---|---|
imagen-4.0-fast-generate-001 | Stable version of Imagen 4 Fast | پایدار | ۱۴-۰۸-۲۰۲۵ | ۲۰۲۶-۰۶-۳۰ |
Imagen 4 Ultra model names
| نام مدل | توضیحات | Release stage | تاریخ انتشار | تاریخ خاموش شدن |
|---|---|---|---|---|
imagen-4.0-ultra-generate-001 | Stable version of Imagen 4 Ultra | پایدار | ۱۴-۰۸-۲۰۲۵ | ۲۰۲۶-۰۶-۳۰ |
Imagen 3 Capability model names
| نام مدل | توضیحات | Release stage | تاریخ انتشار | تاریخ خاموش شدن |
|---|---|---|---|---|
imagen-3.0-capability-001 | Initial stable version of Imagen 3 Capability | پایدار | ۲۰۲۴-۱۲-۱۰ | ۲۰۲۶-۰۶-۳۰ |
Imagen 3 model names
| نام مدل | توضیحات | Release stage | تاریخ انتشار | تاریخ خاموش شدن |
|---|---|---|---|---|
imagen-3.0-generate-002 | Latest stable version of Imagen 3 | پایدار | ۲۰۲۵-۰۱-۲۳ | ۲۰۲۶-۰۶-۳۰ |
imagen-3.0-generate-001 | Initial stable version of Imagen 3 | پایدار | 2024-07-31 | ۲۰۲۶-۰۶-۳۰ |
Imagen 3 Fast model names
| نام مدل | توضیحات | Release stage | تاریخ انتشار | تاریخ خاموش شدن |
|---|---|---|---|---|
imagen-3.0-fast-generate-001 | Initial stable version of Imagen 3 Fast | پایدار | 2024-07-31 | ۲۰۲۶-۰۶-۳۰ |
مراحل بعدی
Try out the capabilities of the Gemini API
- Build multi-turn conversations (chat) .
- Generate text from text-only prompts .
- Generate text by prompting with various file types, like images , PDFs , video , and audio .
- Generate structured output (like JSON) from both text and multimodal prompts.
- Generate and edit images from both text and multimodal prompts.
- Generate speech (both single- and multiple-speakers) using Gemini text-to-speech (TTS) models.
- Stream input and output (including audio) using the Gemini Live API .
- Use tools (like function calling and Grounding with
Google Search orGoogle Maps ) to connect a Gemini model to other parts of your app and external systems and information.
For mobile and web apps, the Firebase AI Logic SDKs let you interact with the supported Gemini models directly from your app.
Gemini models are considered multimodal because they're capable of processing and even generating multiple modalities, including text, code, PDFs, images, video, and audio.
Also, review our FAQ about all the models that Firebase AI Logic supports and does not support.
Featured models
فلش جمینی ۳.۷
gemini-3.7-flash
Frontier-class performance rivaling larger models at a fraction of the cost.
جمینی ۳.۵ فلش-لایت
gemini-3.5-flash-lite
High-volume, cost-sensitive workhorse model with the performance and quality of the Gemini 3 series.
Gemini 3.1 Flash Image ( Nano Banana 2 )
gemini-3.1-flash-image
Powerful, high-efficiency image generation and editing model, optimized for speed and high-volume use cases.
General-use models
Go to tables with model details
جمینی ۳.۱ پرو
gemini-3.1-pro-preview
Advanced intelligence, complex problem-solving skills, and powerful agentic and vibe coding capabilities.
فلش جمینی ۳.۷
gemini-3.7-flash
Frontier-class performance rivaling larger models at a fraction of the cost.
جمینی ۳.۵ فلش-لایت
gemini-3.5-flash-lite
High-volume, cost-sensitive workhorse model with the performance and quality of the Gemini 3 series.
Older stable general-use models
Gemini 3.6 Flash (
gemini-3.6-flash): Previous Gemini 3.x Flash model for frontier-class performance rivaling larger models at a fraction of the cost. (billing not required)Gemini 3.5 Flash (
gemini-3.5-flash): Previous Gemini 3.x Flash model for frontier-class performance rivaling larger models at a fraction of the cost. (billing not required)Gemini 3.1 Flash‑Lite (
gemini-3.1-flash-lite): Previous Gemini 3.x Flash‑Lite model for high-volume, cost-sensitive workhorse tasks. (billing not required)
Image-generating models
Go to tables with model details
Gemini 3 Pro Image ( Nano Banana Pro )
gemini-3-pro-image
State-of-the-art image generation and editing model for highly contextual native image creation.
Gemini 3.1 Flash Image ( Nano Banana 2 )
gemini-3.1-flash-image
Powerful, high-efficiency image generation and editing model, optimized for speed and high-volume use cases.
Gemini 3.1 Flash-Lite Image ( Nano Banana 2 Lite )
gemini-3.1-flash-lite-image
Ultra-low latency and cost-effective image generation and editing model, designed for high-volume interactive use cases.
Audio-generating models
Text-to-speech (TTS) models
You can generate speech from text input with Gemini TTS models.
Go to tables with model details
Gemini 3.x Flash TTS
gemini-3.1-flash-tts-preview
Powerful, low-latency speech generation from text input.
مدلهای Live API
You can generate bidirectional streamed audio with models that support the Gemini Live API .
Go to tables with model details
Gemini 3.x Flash with Gemini Live API native audio
Gemini Developer API:
gemini-3.1-flash-live-preview
Agent Platform Gemini API:
not supported
Enables low-latency, real-time voice and video interactions with a Gemini model that is bidirectional .
Gemini 2.5 Flash with Gemini Live API native audio
Gemini Developer API:
gemini-2.5-flash-native-audio-preview-12-2025
Agent Platform Gemini API:
gemini-live-2.5-flash-native-audio
Enables low-latency, real-time voice and video interactions with a Gemini model that is bidirectional .
The remainder of this page provides detailed information about the models supported by Firebase AI Logic .
- Supported input and output
- High-level comparison of the supported capabilities
- Specifications and limitations, for example max input tokens or max length of input video
Description of how models are versioned , specifically their stable , preview , and experimental versions
Lists of available model names to include in your code during initialization
Lists of supported languages for the models
At the bottom of this page, you can view detailed information about previous generation models .
Compare models
Each model has different capabilities to support various use cases. Note that each of tables in this section describe each model when used with Firebase AI Logic . Each model might have additional capabilities that aren't available when using our SDKs.
If you can't find the information you're looking for in the following sub-sections, you can find even more information in your chosen API provider documentation: Gemini Developer API orAgent Platform Gemini API (formerly Vertex AI) .
Supported input and output
The following table lists the supported input and output types when using each model with Firebase AI Logic .
To learn about supported file types, see Supported input files and requirements .
| Gemini 3.x Pro, Flash, Flash‑Lite | Gemini 3.x Pro Image | Gemini 3.x Flash Image | Gemini 3.x Flash‑Lite Image | Gemini 3.x Flash TTS | Gemini Live | |
|---|---|---|---|---|---|---|
| Input types | ||||||
| متن | ||||||
| کد | ||||||
| اسناد (PDFs or plain-text) | ||||||
| Images | ||||||
| ویدئو | ||||||
| صوتی | ||||||
| انواع خروجی | ||||||
| متن | ||||||
| Text (streaming) | ||||||
| کد | ||||||
| خروجی ساختاریافته (like JSON) | ||||||
| تصاویر | ||||||
| ویدئو | ||||||
| صوتی | ||||||
Supported capabilities and features
The following table lists the supported capabilities and features when using each model with Firebase AI Logic .
| Gemini 3.x Pro, Flash, Flash‑Lite | Gemini 3.x Pro Image | Gemini 3.x Flash Image | Gemini 3.x Flash‑Lite Image | Gemini 3.x Flash TTS | Gemini Live | |
|---|---|---|---|---|---|---|
| تفکر | ||||||
| Generate text from text-only or multimodal inputs | interleaved or as part of image | interleaved or as part of image | interleaved or as part of image | transcription of audio | ||
| Generate images | ||||||
| Edit images | ||||||
| Generate audio | گفتار | streamed audio | ||||
| Generate structured output (like JSON) | ||||||
| Analyze documents (PDFs or plain-text) ( text-output | image-output ) | ||||||
| Analyze images ( text-output | image-output ) | ||||||
| Analyze video ( text-output | image-output ) | streamed video | |||||
| Analyze audio | streamed audio | |||||
| Multi-turn chat | ||||||
| Bidirectional multimodal streaming | ||||||
| ابزارهای پشتیبانی شده | ||||||
| فراخوانی تابع | ||||||
| Code execution | ||||||
| زمینه URL | ||||||
| اتصال به زمین با | ||||||
| اتصال به زمین با | ||||||
Specifications and limitations
The following table lists the specifications and limitations when using each model with Firebase AI Logic .
| ملک | Gemini 3.x Pro, Flash, Flash‑Lite | Gemini 3.x Pro Image | Gemini 3.x Flash Image | Gemini 3.x Flash‑Lite Image | Gemini 3.x Flash TTS | Gemini Live |
|---|---|---|---|---|---|---|
| Input token limit * | 1,048,576 tokens | 65,536 tokens | 131,072 tokens | 65,536 tokens | 8,192 tokens | see docs |
| Output token limit * | 65,536 tokens | ۳۲۷۶۸ توکن | ۳۲۷۶۸ توکن | 4,096 tokens | 16,384 tokens | |
| PDFs (per request) | ||||||
| Max number of input PDF files ** | 900 files | 14 files | 14 files | 14 files | --- | see docs |
| Max number of pages per input PDF file ** | 900 pages | ۱۴ صفحه | ۱۴ صفحه | ۱۴ صفحه | --- | |
| حداکثر اندازه per input PDF file | ۵۰ مگابایت | ۵۰ مگابایت | ۵۰ مگابایت | ۵۰ مگابایت | --- | |
| Images (per request) | ||||||
| Max number of input images | 1,000 images | 14 images | 14 images | 14 images | --- | see docs |
| حداکثر اندازه per input base64-encoded image | 7 MB | 7 MB | 7 MB | 7 MB | --- | |
| Max number of output images | --- | Up to output token limit | Up to output token limit | Up to output token limit | --- | |
| Video (per request) | ||||||
| Max number of input video files | ۱۰ فایل | --- | Up to input token limit | Up to input token limit | --- | see docs |
| Max length of all input video (frames only) | ~60 minutes | --- | ~25 minutes | ~12 minutes | --- | |
| Max length of all input video (frames+audio) | ~45 minutes | --- | --- | --- | --- | |
| Audio (per request) | ||||||
| Max number of input audio files | 1 file | --- | --- | --- | --- | see docs |
| Max length of all input audio | ~8.4 hours | --- | --- | --- | --- | |
* For all Gemini models, a token is equivalent to about 4 characters, so 100 tokens are about 60-80 English words. For Gemini models, you can determine the total count of tokens in your requests using countTokens .
** PDFs are treated as images, so a single page of a PDF is treated as one image. The number of pages allowed in a request is limited to the number of images the model can support.
Find additional detailed information
Quotas and pricing are different for each model. Pricing also depends on input and output.
Learn about supported input file types, how to specify MIME type, and how to make sure that your input files and multimodal requests meet the requirements and follow best practices in Supported input files and requirements .
Model versioning and naming patterns
Models are offered in stable , preview , and experimental versions. For convenience, the-latest aliases are supported, but they're not recommended due to the inconsistency in which model version they point to.
Make sure to review our best practices for using model versions .
To find specific model names to use in your code, see the "available model names" section later on this page.
| Version type / Release stage | توضیحات | Model name pattern | |
|---|---|---|---|
| پایدار | Stable versions are available and supported for production use starting on the release date.
| Model names of stable versions have no suffix مثال: | |
| پیشنمایش | Preview versions have new capabilities and are considered not stable .
| Model names of preview versions are appended with مثال: | |
| تجربی | Experimental versions have new capabilities and are considered not stable .
| Model names of experimental versions are appended with مثال: | |
| Shutdown (retired) | Shutdown (retired) versions are past their shutdown (retirement) date and have been permanently deactivated.
| --- | |
Best practices for using model versions
In your production apps , use the explicit model name for the most recent stable version.
For Agent Platform Gemini API (formerly Vertex AI) , if you choose to use a short-term availability model in your production app, it's even more critical that you use Firebase Remote Config or server prompt templates to control the model name used for your AI feature.
Use preview and experimental versions only during prototyping . We recommend using a stable version when you start developing and testing for a production use case.
We do not recommend using the
-latestaliases (even during development). This alias points to the latest release for a specific model variation (which could be a stable, preview, or experimental version). This alias will get hot-swapped with every new release of a specific model variation, and only for breaking changes will a 2-week-prior notification email be sent. This instability of which model you're actually using can lead to unexpected behavior changes for your AI feature.
Available model names
Model names are the explicit values that you include in your code during initialization of the model.
General-use models (like
gemini-3.7-flash)Image-generating models (like
gemini-3.1-flash-image, aka the "Nano Banana" models)Audio-generating models :
- Text-to-speech (TTS) models (like
gemini-3.1-flash-tts-preview) - Live API models (like
gemini-live-2.5-flash-native-audio)
- Text-to-speech (TTS) models (like
For initialization examples for your platform, see the getting started guide .
For details about the release stages (especially for use cases, billing, and shutdown), see model versioning and naming patterns .
Programmatically list all available models
You can list all available models names using the REST API:
Gemini Developer API : Call the
models.listendpointAgent Platform Gemini API (formerly Vertex AI) : Call the
publishers.models.listendpoint
Note that this returned list will include all models supported by the API providers, but Firebase AI Logic only supports the Gemini models described on this page.
General-use models
Gemini 3.x Pro model names
Requires the pay-as-you-go Blaze pricing plan regardless of your Gemini API provider.
| نام مدل | توضیحات | Release stage | تاریخ انتشار | تاریخ خاموش شدن |
|---|---|---|---|---|
gemini-3.1-pro-preview | Latest preview version of Gemini 3.x Pro | پیشنمایش | ۲۰۲۶-۰۲-۱۹ | To be determined |
Gemini 3.x Flash model names
Does not require the pay-as-you-go Blaze pricing plan if you're using the Gemini Developer API .
| نام مدل | توضیحات | Release stage | تاریخ انتشار | تاریخ خاموش شدن |
|---|---|---|---|---|
gemini-3.7-flash | Latest stable version of Gemini 3.x Flash This is a short-term availability model (see Model versioning and naming patterns ). | پایدار | ۲۰۲۶-۰۸-۱۳ | To be determined |
gemini-3.6-flash | Stable version of Gemini 3.x Flash This is a short-term availability model (see Model versioning and naming patterns ). | پایدار | ۲۰۲۶-۰۷-۲۱ | To be determined |
gemini-3.5-flash | Stable version of Gemini 3.x Flash | پایدار | ۲۰۲۶-۰۵-۱۹ | No earlier than 2027-05-19 |
gemini-3-flash-preview | Preview version of Gemini 3.x Flash | پیشنمایش | ۲۰۲۵-۱۲-۱۷ | To be determined |
Gemini 3.x Flash‑Lite model names
Does not require the pay-as-you-go Blaze pricing plan if you're using the Gemini Developer API .
| نام مدل | توضیحات | Release stage | تاریخ انتشار | تاریخ خاموش شدن |
|---|---|---|---|---|
gemini-3.5-flash-lite | Latest stable version of Gemini 3.x Flash‑Lite | پایدار | ۲۰۲۶-۰۷-۲۱ | No earlier than 2027-07-21 |
gemini-3.1-flash-lite | Stable version of Gemini 3.x Flash‑Lite | پایدار | ۲۰۲۶-۰۵-۰۷ | No earlier than 2027-05-07 |
Image-generating models
Gemini 3.x Pro Image model names (aka "Nano Banana Pro")
Requires the pay-as-you-go Blaze pricing plan regardless of your Gemini API provider.
| نام مدل | توضیحات | Release stage | تاریخ انتشار | تاریخ خاموش شدن |
|---|---|---|---|---|
gemini-3-pro-image | Stable version of Gemini 3.x Pro Image (aka "Nano Banana Pro") | پایدار | ۲۰۲۶-۰۵-۲۸ | No earlier than 2027-05-28 |
gemini-3-pro-image-preview | Preview version of Gemini 3.x Pro Image (aka "Nano Banana Pro") | پیشنمایش | ۲۰۲۵-۱۱-۲۰ | As early as ۲۰۲۶-۰۶-۲۵ |
Gemini 3.x Flash Image model names (aka "Nano Banana 2")
Requires the pay-as-you-go Blaze pricing plan regardless of your Gemini API provider.
| نام مدل | توضیحات | Release stage | تاریخ انتشار | تاریخ خاموش شدن |
|---|---|---|---|---|
gemini-3.1-flash-image | Stable version of Gemini 3.x Flash Image (aka "Nano Banana 2") | پایدار | ۲۰۲۶-۰۵-۲۸ | No earlier than 2027-05-28 |
gemini-3.1-flash-image-preview | Preview version of Gemini 3.x Flash Image (aka "Nano Banana 2") | پیشنمایش | ۲۶-۰۲-۲۰۲۶ | As early as ۲۰۲۶-۰۶-۲۵ |
Gemini 3.x Flash‑Lite Image model names (aka "Nano Banana 2 Lite")
Requires the pay-as-you-go Blaze pricing plan regardless of your Gemini API provider.
| نام مدل | توضیحات | Release stage | تاریخ انتشار | تاریخ خاموش شدن |
|---|---|---|---|---|
gemini-3.1-flash-lite-image | Stable version of Gemini 3.x Flash‑Lite Image (aka "Nano Banana 2 Lite") This is a short-term availability model (see Model versioning and naming patterns ). | پایدار | ۲۰۲۶-۰۶-۳۰ | To be determined |
Audio - text-to-speech (TTS) models
Gemini 3.x Flash TTS model names
Does not require the pay-as-you-go Blaze pricing plan if you're using the Gemini Developer API .
| نام مدل | توضیحات | Release stage | تاریخ انتشار | تاریخ خاموش شدن |
|---|---|---|---|---|
gemini-3.1-flash-tts-preview | Preview version for Gemini 3.x Flash TTS | پیشنمایش | 2026-06-17 | To be determined |
Audio - Live API models
Gemini 3.x Flash Live model names
| رابط برنامهنویسی کاربردی (API) توسعهدهندگان جمینی نام مدل | توضیحات | Release stage | تاریخ انتشار | تاریخ خاموش شدن |
|---|---|---|---|---|
gemini-3.1-flash-live-preview 1 | Preview version for the Live API on the Gemini Developer API | پیشنمایش | ۲۰۲۶-۰۳-۲۶ | To be determined |
| Agent Platform Gemini API (formerly Vertex AI) نام مدل | توضیحات | Release stage | تاریخ انتشار | تاریخ خاموش شدن |
|---|---|---|---|---|
| The Agent Platform Gemini API (formerly Vertex AI) does not support any Gemini Live 3.x models. | ||||
1 Only supported by the Gemini Developer API . Also, even though this is a preview model, it's available on the "free tier" of the Gemini Developer API .
Gemini 2.5 Flash Live model names
Does not require the pay-as-you-go Blaze pricing plan if you're using the Gemini Developer API (usually preview models require a paid plan).
Even though the following models have different model names depending on the Gemini API provider, the features of the model are the same.
| رابط برنامهنویسی کاربردی (API) توسعهدهندگان جمینی نام مدل | توضیحات | Release stage | تاریخ انتشار | تاریخ خاموش شدن |
|---|---|---|---|---|
gemini-2.5-flash-native-audio-preview-12-2025 1 | Preview version for the Live API on the Gemini Developer API | پیشنمایش | ۲۰۲۵-۱۲-۱۲ | To be determined |
gemini-2.5-flash-native-audio-preview-09-2025 1 | Initial preview version for the Live API on the Gemini Developer API | پیشنمایش | ۲۰۲۵-۰۹-۱۸ | To be determined |
| Agent Platform Gemini API (formerly Vertex AI) نام مدل | توضیحات | Release stage | تاریخ انتشار | تاریخ خاموش شدن |
|---|---|---|---|---|
gemini-live-2.5-flash-native-audio 2 | Stable version for the Live API on the Agent Platform Gemini API (formerly Vertex AI) | پایدار | ۲۰۲۵-۱۲-۱۲ | No earlier than 2026-12-12 |
gemini-live-2.5-flash-preview-native-audio-09-2025 2 | Preview version for the Live API on the Agent Platform Gemini API (formerly Vertex AI) | پیشنمایش | ۲۰۲۵-۰۹-۱۸ | To be determined |
1 Only supported by the Gemini Developer API . Also, even though these are preview models, they're available on the "free tier" of the Gemini Developer API .
2 Only supported by the Agent Platform Gemini API (formerly Vertex AI) . Also, these models are not available in the global location.
زبانهای پشتیبانیشده
All the Gemini models can understand and respond in the following languages:
Arabic (ar), Bengali (bn), Bulgarian (bg), Chinese simplified and traditional (zh), Croatian (hr), Czech (cs), Danish (da), Dutch (nl), English (en), Estonian (et), Finnish (fi), French (fr), German (de), Greek (el), Hebrew (iw), Hindi (hi), Hungarian (hu), Indonesian (id), Italian (it), Japanese (ja), Korean (ko), Latvian (lv), Lithuanian (lt), Norwegian (no), Polish (pl), Portuguese (pt), Romanian (ro), Russian (ru), Serbian (sr), Slovak (sk), Slovenian (sl), Spanish (es), Swahili (sw), Swedish (sv), Thai (th), Turkish (tr), Ukrainian (uk), Vietnamese (vi)
Gemini 2.0 Flash , Gemini 1.5 Pro and Gemini 1.5 Flash models can understand and respond in the following additional languages:
Afrikaans (af), Amharic (am), Assamese (as), Azerbaijani (az), Belarusian (be), Bosnian (bs), Catalan (ca), Cebuano (ceb), Corsican (co), Welsh (cy), Dhivehi (dv), Esperanto (eo), Basque (eu), Persian (fa), Filipino (Tagalog) (fil), Frisian (fy), Irish (ga), Scots Gaelic (gd), Galician (gl), Gujarati (gu), Hausa (ha), Hawaiian (haw), Hmong (hmn), Haitian Creole (ht), Armenian (hy), Igbo (ig), Icelandic (is), Javanese (jv), Georgian (ka), Kazakh (kk), Khmer (km), Kannada (kn), Krio (kri), Kurdish (ku), Kyrgyz (ky), Latin (la), Luxembourgish (lb), Lao (lo), Malagasy (mg), Maori (mi), Macedonian (mk), Malayalam (ml), Mongolian (mn), Meiteilon (Manipuri) (mni-Mtei), Marathi (mr), Malay (ms), Maltese (mt), Myanmar (Burmese) (my), Nepali (ne), Nyanja (Chichewa) (ny), Odia (Oriya) (or), Punjabi (pa), Pashto (ps), Sindhi (sd), Sinhala (Sinhalese) (si), Samoan (sm), Shona (sn), Somali (so), Albanian (sq), Sesotho (st), Sundanese (su), Tamil (ta), Telugu (te), Tajik (tg), Uyghur (ug), Urdu (ur), Uzbek (uz), Xhosa (xh), Yiddish (yi), Yoruba (yo), Zulu (zu)
Information about previous models
The following are active, but previous generation models. We recommend using one of the latest models instead when possible.
If you can't find the information you're looking for in the following sub-sections, you can find even more information in your chosen API provider documentation: Gemini Developer API orAgent Platform Gemini API (formerly Vertex AI)
Older Gemini models
-
gemini-2.5-pro -
gemini-2.5-flash -
gemini-2.5-flash-lite -
gemini-2.5-flash-image(aka "Nano Banana") -
gemini-2.0-flash-001(and its auto-updated aliasgemini-2.0-flash) -
gemini-2.0-flash-lite-001(and its auto-updated aliasgemini-2.0-flash-lite)
For information about older Gemini Live API models, see the Gemini API provider documentation:
Older Imagen models
-
imagen-4.0-ultra-generate-001 -
imagen-4.0-generate-001 -
imagen-4.0-fast-generate-001 -
imagen-3.0-capability-001 -
imagen-3.0-generate-002 -
imagen-3.0-generate-001 -
imagen-3.0-fast-generate-001
View details about about previous models
These are the input and output types when using each model with Firebase AI Logic :
| Gemini 2.5 Pro, Flash, Flash‑Lite | Gemini 2.5 Flash Image | Gemini 2.0 Flash | Gemini 2.0 Flash‑Lite | ایمیجِن (generate) | ایمیجِن (capability) | |
|---|---|---|---|---|---|---|
| Input types | ||||||
| متن | ||||||
| کد | ||||||
| اسناد (PDFs or plain-text) | ||||||
| Images | ||||||
| ویدئو | ||||||
| صوتی | ||||||
| Audio (streaming) | ||||||
| انواع خروجی | ||||||
| متن | ||||||
| Text (streaming) | ||||||
| کد | ||||||
| خروجی ساختاریافته (like JSON) | ||||||
| Images | ||||||
| ویدئو | ||||||
| صوتی | ||||||
| Audio (streaming) | ||||||
These are the capabilities and features when using each model with Firebase AI Logic :
| Gemini 2.5 Pro, Flash, Flash‑Lite | Gemini 2.5 Flash Image | Gemini 2.0 Flash | Gemini 2.0 Flash‑Lite | ایمیجِن (generate) | ایمیجِن (capability) | |
|---|---|---|---|---|---|---|
| تفکر | ||||||
| Generate text from text-only or multimodal inputs | interleaved or as part of image | |||||
| Generate images ( Gemini or Imagen ) | ||||||
| Edit images ( Gemini or Imagen ) | ||||||
| Generate audio | ||||||
| Generate structured output (like JSON) | ||||||
| Analyze documents (PDFs or plain-text) ( text-output | image-output ) | ||||||
| Analyze images ( text-output | image-output ) | ||||||
| Analyze video ( text-output | image-output ) | ||||||
| Analyze audio | ||||||
| Multi-turn chat | ||||||
| Bidirectional multimodal streaming | ||||||
| ابزارهای پشتیبانی شده | ||||||
| فراخوانی تابع | ||||||
| Code execution | ||||||
| زمینه URL | ||||||
| اتصال به زمین با | ||||||
| اتصال به زمین با | ||||||
| دستورالعملهای سیستم | ||||||
| تعداد توکنها | ||||||
These are the specifications and limitations when using each model with Firebase AI Logic :
| ملک | Gemini 2.5 Pro, Flash, Flash‑Lite | Gemini 2.5 Flash Image | Gemini 2.0 Flash | Gemini 2.0 Flash‑Lite | ایمیجِن (generate) | ایمیجِن (capability) |
|---|---|---|---|---|---|---|
| Input token limit * | 1,048,576 tokens | ۳۲۷۶۸ توکن | 1,048,576 tokens | 1,048,576 tokens | 480 tokens | 480 tokens |
| Output token limit * | 65,536 tokens | 8,192 tokens | 8,192 tokens | 8,192 tokens | --- | --- |
| Knowledge cutoff date | ژانویه ۲۰۲۵ | --- | ژوئن ۲۰۲۴ | ژوئن ۲۰۲۴ | --- | --- |
| PDFs (per request) | ||||||
| Max number of input PDF files ** | ۳۰۰۰ فایل | ۳ فایل | ۳۰۰۰ فایل | ۳۰۰۰ فایل | --- | --- |
| Max number of pages per input PDF file ** | 1,000 pages | ۳ صفحه | 1,000 pages | 1,000 pages | --- | --- |
| حداکثر اندازه per input PDF file | ۵۰ مگابایت | ۵۰ مگابایت | ۵۰ مگابایت | ۵۰ مگابایت | --- | --- |
| Images (per request) | ||||||
| Max number of input images | 3,000 images | 3 images | 3,000 images | 3,000 images | --- | 4 images |
| Max number of output images | --- | Up to output token limit | --- | --- | 4 images | 4 images |
| حداکثر اندازه per input base64-encoded image | 7 MB | 7 MB | 7 MB | 7 MB | --- | --- |
| Video (per request) | ||||||
| Max number of input video files | ۱۰ فایل | --- | ۱۰ فایل | ۱۰ فایل | --- | --- |
| Max length of all input video (frames only) | ~60 minutes | --- | ~60 minutes | ~60 minutes | --- | --- |
| Max length of all input video (frames+audio) | ~45 minutes | --- | ~45 minutes | ~45 minutes | --- | --- |
| Audio (per request) | ||||||
| Max number of input audio files | 1 file | --- | 1 file | 1 file | --- | --- |
| Max number of output audio files | --- | --- | --- | --- | --- | --- |
| Max length of all input audio | ~8.4 hours | --- | ~8.4 hours | ~8.4 hours | --- | --- |
| Max length of all output audio | --- | --- | --- | --- | --- | --- |
* For all Gemini models, a token is equivalent to about 4 characters, so 100 tokens are about 60-80 English words. For Gemini models, you can determine the total count of tokens in your requests using countTokens .
** PDFs are treated as images, so a single page of a PDF is treated as one image. The number of pages allowed in a request is limited to the number of images the model can support.
Model names are the explicit values that you include in your code during initialization of the model.
مدلهای جمینی
Gemini 2.5 Pro model names
Does not require the pay-as-you-go Blaze pricing plan if you're using the Gemini Developer API .
| نام مدل | توضیحات | Release stage | تاریخ انتشار | تاریخ خاموش شدن |
|---|---|---|---|---|
gemini-2.5-pro | Stable version of Gemini 2.5 Pro | پایدار | ۲۰۲۵-۰۶-۱۷ | As early as 2026-10-16 |
Gemini 2.5 Flash model names
Does not require the pay-as-you-go Blaze pricing plan if you're using the Gemini Developer API .
| نام مدل | توضیحات | Release stage | تاریخ انتشار | تاریخ خاموش شدن |
|---|---|---|---|---|
gemini-2.5-flash | Stable version of Gemini 2.5 Flash | پایدار | ۲۰۲۵-۰۶-۱۷ | As early as 2026-10-16 |
Gemini 2.5 Flash‑Lite model names
Does not require the pay-as-you-go Blaze pricing plan if you're using the Gemini Developer API .
| نام مدل | توضیحات | Release stage | تاریخ انتشار | تاریخ خاموش شدن |
|---|---|---|---|---|
gemini-2.5-flash-lite | Stable version of Gemini 2.5 Flash‑Lite | پایدار | ۲۰۲۵-۰۷-۲۲ | As early as 2026-10-16 |
Gemini 2.5 Flash Image model names (aka "Nano Banana")
Requires the pay-as-you-go Blaze pricing plan regardless of your Gemini API provider.
| نام مدل | توضیحات | Release stage | تاریخ انتشار | تاریخ خاموش شدن |
|---|---|---|---|---|
gemini-2.5-flash-image | Stable version of Gemini 2.5 Flash Image (aka "Nano Banana") | پایدار | ۲۰۲۵-۱۰-۰۲ | ۲۰۲۶-۱۰-۰۲ |
Gemini 2.0 Flash model names
| نام مدل | توضیحات | Release stage | تاریخ انتشار | تاریخ خاموش شدن |
|---|---|---|---|---|
gemini-2.0-flash-001 | Latest stable version of Gemini 2.0 Flash | پایدار | ۲۰۲۵-۰۲-۰۵ | 2026-06-01 |
gemini-2.0-flash | Auto-updated alias pointing to the latest stable version of Gemini 2.0 Flash (currently gemini-2.0-flash-001 ) | پایدار | 2025-02-10 | 2026-06-01 |
Gemini 2.0 Flash‑Lite model names
| نام مدل | توضیحات | Release stage | تاریخ انتشار | تاریخ خاموش شدن |
|---|---|---|---|---|
gemini-2.0-flash-lite-001 | Latest stable version of Gemini 2.0 Flash‑Lite | پایدار | 2025-02-25 | 2026-06-01 |
gemini-2.0-flash-lite | Auto-updated alias pointing to the latest stable version of Gemini 2.0 Flash‑Lite (currently gemini-2.0-flash-lite-001 ) | پایدار | 2025-02-25 | 2026-06-01 |
مدلهای ایمیجن
Imagen 4 model names
| نام مدل | توضیحات | Release stage | تاریخ انتشار | تاریخ خاموش شدن |
|---|---|---|---|---|
imagen-4.0-generate-001 | Stable version of Imagen 4 | پایدار | ۱۴-۰۸-۲۰۲۵ | ۲۰۲۶-۰۶-۳۰ |
Imagen 4 Fast model names
| نام مدل | توضیحات | Release stage | تاریخ انتشار | تاریخ خاموش شدن |
|---|---|---|---|---|
imagen-4.0-fast-generate-001 | Stable version of Imagen 4 Fast | پایدار | ۱۴-۰۸-۲۰۲۵ | ۲۰۲۶-۰۶-۳۰ |
Imagen 4 Ultra model names
| نام مدل | توضیحات | Release stage | تاریخ انتشار | تاریخ خاموش شدن |
|---|---|---|---|---|
imagen-4.0-ultra-generate-001 | Stable version of Imagen 4 Ultra | پایدار | ۱۴-۰۸-۲۰۲۵ | ۲۰۲۶-۰۶-۳۰ |
Imagen 3 Capability model names
| نام مدل | توضیحات | Release stage | تاریخ انتشار | تاریخ خاموش شدن |
|---|---|---|---|---|
imagen-3.0-capability-001 | Initial stable version of Imagen 3 Capability | پایدار | ۲۰۲۴-۱۲-۱۰ | ۲۰۲۶-۰۶-۳۰ |
Imagen 3 model names
| نام مدل | توضیحات | Release stage | تاریخ انتشار | تاریخ خاموش شدن |
|---|---|---|---|---|
imagen-3.0-generate-002 | Latest stable version of Imagen 3 | پایدار | ۲۰۲۵-۰۱-۲۳ | ۲۰۲۶-۰۶-۳۰ |
imagen-3.0-generate-001 | Initial stable version of Imagen 3 | پایدار | 2024-07-31 | ۲۰۲۶-۰۶-۳۰ |
Imagen 3 Fast model names
| نام مدل | توضیحات | Release stage | تاریخ انتشار | تاریخ خاموش شدن |
|---|---|---|---|---|
imagen-3.0-fast-generate-001 | Initial stable version of Imagen 3 Fast | پایدار | 2024-07-31 | ۲۰۲۶-۰۶-۳۰ |
مراحل بعدی
Try out the capabilities of the Gemini API
- Build multi-turn conversations (chat) .
- Generate text from text-only prompts .
- Generate text by prompting with various file types, like images , PDFs , video , and audio .
- Generate structured output (like JSON) from both text and multimodal prompts.
- Generate and edit images from both text and multimodal prompts.
- Generate speech (both single- and multiple-speakers) using Gemini text-to-speech (TTS) models.
- Stream input and output (including audio) using the Gemini Live API .
- Use tools (like function calling and Grounding with
Google Search orGoogle Maps ) to connect a Gemini model to other parts of your app and external systems and information.