How do I choose between the two Gemini 3.5 Flash endpoints?
Choose /gemini/chat/completions when you need to manage messages yourself, handle tool calls, and parse choices. Choose /aichat2/conversations when you want to save conversation history, continue asking follow-up questions by id, or use file content blocks; its standard JSON response includes answer and id.
How can I make Gemini 3.5 Flash analyze images?
In Chat Completions, write message content as an array of content blocks, combining text and image_url, and put an accessible image address or Base64 data URI in image_url.url. The text should specify the object of interest and output requirements, such as extracting fields, comparing differences, or explaining charts.
Can I have it read PDFs?
Gemini 3.5 Flash natively supports PDF understanding. When using this platform, you can provide a PDF link through file_url in an AI Chat v2 message and combine it with file reading for analysis; do not submit a PDF link as a regular image block, and do not equate native capabilities with the upload methods of all endpoints.
Why is there token usage but no final text?
Gemini 3.5 Flash uses reasoning tokens, and an output budget that is too small may be exhausted before a final answer is produced. The usage guide recommends setting max_tokens to 512 or higher; complex tasks should allow more room, and you should use finish_reason, usage, and the actual text to determine whether the budget needs to be increased.
Is it the same invocation name as gemini-3-flash-preview?
No. Google lists gemini-3.5-flash as Stable and gemini-3-flash-preview as Preview. To invoke this model, use the full ID gemini-3.5-flash; when replacing models, especially retest structured results, tool loops, and long-task performance to avoid deploying it directly after only changing the name.