Skip to main content

Minimal request

cURL example

Python example

Node.js example

Best practices

  • Start with a Flash model for lower-latency chat
  • Keep the parts model intact so media can be added later without redesign
  • If you do not need extra reasoning cost, set thinkingBudget to 0 on Gemini 2.5 Flash