Google Gemini 3.5 Flash Lite is the lightweight tier of the Gemini 3.5 family, tuned for low cost and fast responses - roughly one fifth the cost of Gemini 3.5 Flash. Great for high-volume calls, latency-sensitive apps, and cost-sensitive workloads such as everyday chat, translation, short content generation, and classification/labeling tasks. Supports multimodal input (text, images, PDF), optional web search for up-to-date answers, and a thinking (reasoning) mode. Switch to Gemini 3.5 Flash or Gemini 3.1 Pro when you need deeper reasoning or more polished output. If you need to search the web for information, please turn on the "Web Search" switch. Maximum context window (input and output combined): about 1,048,576 tokens. Maximum output per reply: about 65,536 tokens (model reasoning, if any, counts toward this limit). Maximum total upload size per request: 256MB (all files and message content combined). Accepted document formats: pdf, docx, xlsx, pptx, txt, csv (file content is provided to the model as extracted text). Accepted image formats: heic, heif, jpeg, jpg, png, webp. Maximum number of images: 3600. Maximum size per image: 14MB. Maximum total image size per message: 14MB (text and history share the same request size limit; uploads close to the limit may still be rejected). Image dimensions: at least 200 pixels per side. Maximum image resolution: 36,000,000 pixels in total. Output content formats: Text This module supports continuous conversations. AI can make mistakes. Please verify important information.