Skip to content
Docs menu

LLM models

Gemini 3.6 Flash

On this page

Google's newest speed-optimized multimodal model — fast text, vision, and reasoning.

modelIdgemini-3-6-flash
Modalitytext
PricingSee this model on the Pricing page for the current per-call price (with your markup).

Operations

OperationmodelIdEndpointRequired input
Defaultgemini-3-6-flashPOST /api/v1/generateprompt
Poll taskGET /api/v1/task/{id}?model=gemini-3-6-flash

Input parameters

FieldTypeRequiredValues / example
promptstringYesText value (default: If gravity took the day off, describe in one playful paragraph what the morning commute would look like.)
memorybooleanNoAutomatically append previous messages to maintain multi-turn context. May increase token usage. (true/false) (default: false)
function_callingbooleanNoLets you define functions that Gemini can call (true/false) (default: false)
google_searchbooleanNoUse Google Search (true/false) (default: true)
thinkingbooleanNoInclude the model's thinking process in the response (true/false) (default: true)
systemstringNoSystem instruction for the model. Defaults to "You are Gemini 3.6 Flash, developed by Google." when omitted.

Example request

bash
curl -X POST https://you.bot/api/v1/generate \
  -H "Authorization: Bearer $YOUBOT_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"modelId":"gemini-3-6-flash","input":{"prompt":"Explain how diffusion models generate images, in two short sentences.","memory":false,"function_calling":false,"google_search":true,"thinking":true}}'