Gemini 3.5 Flash API
Vollständige API-Referenz für Gemini 3.5 Flash-große Sprachmodelle-Modelle von Google.
Modellvarianten
Diese API unterstützt drei Gemini 3.5 Flash-Modellvarianten:
Schnellvergleich
Gemini 3.5 Flash Text to Text
Endpunkt
POST /api/v1/chat/completions
Anfrageparameter
Erforderlich
Nachrichtenunterstützung
Optional
Beispielanfrage
const response = await fetch('https://api.flaq.ai/api/v1/chat/completions', {
method: 'POST',
headers: {
'Authorization': 'Bearer YOUR_API_KEY',
'Accept': 'text/event-stream',
'Content-Type': 'application/json'
},
body: JSON.stringify({
model: 'gemini-3.5-flash-text-to-text',
messages: [
{
role: 'user',
content: 'Write a concise summary of how transformer models work.'
}
],
stream: true,
max_tokens: 2048
})
});
Gemini 3.5 Flash Image to Text
Endpunkt
POST /api/v1/chat/completions
Anfrageparameter
Erforderlich
Nachrichtenunterstützung
Optional
Beispielanfrage
const response = await fetch('https://api.flaq.ai/api/v1/chat/completions', {
method: 'POST',
headers: {
'Authorization': 'Bearer YOUR_API_KEY',
'Accept': 'text/event-stream',
'Content-Type': 'application/json'
},
body: JSON.stringify({
model: 'gemini-3.5-flash-image-to-text',
messages: [
{
role: 'user',
content: [
{ type: 'text', text: 'Describe the image and extract any visible text.' },
{
type: 'image_url',
image_url: {
url: 'https://example.com/sample-image.jpg'
}
}
]
}
],
stream: true,
max_tokens: 2048
})
});
Gemini 3.5 Flash File Analysis
Endpunkt
POST /api/v1/chat/completions
Anfrageparameter
Erforderlich
Nachrichtenunterstützung
Optional
Beispielanfrage
const response = await fetch('https://api.flaq.ai/api/v1/chat/completions', {
method: 'POST',
headers: {
'Authorization': 'Bearer YOUR_API_KEY',
'Accept': 'text/event-stream',
'Content-Type': 'application/json'
},
body: JSON.stringify({
model: 'gemini-3.5-flash-file-analysis',
messages: [
{
role: 'user',
content: [
{ type: 'text', text: 'Summarize this file and extract the key risks.' },
{
type: 'file',
file: {
filename: 'demo.pdf',
file_data: 'https://example.com/demo.pdf'
}
}
]
}
],
stream: true,
max_tokens: 4096
})
});
Gemini 3.5 Flash LLM-Modelle geben Google-kompatible Completion-Antworten zurück. Mit stream: true ist die Antwort Server-Sent Events; mit stream: false ist die Antwort ein einzelnes JSON-Objekt.
Erfolgreiche Antwort
data: {"id":"chatcmpl-...","object":"chat.completion.chunk","created":1710000000,"model":"gemini-3.5-flash-text-to-text","choices":[{"index":0,"delta":{"role":"assistant"},"finish_reason":null}]}
data: {"id":"chatcmpl-...","object":"chat.completion.chunk","created":1710000000,"model":"gemini-3.5-flash-text-to-text","choices":[{"index":0,"delta":{"content":"Here is a concise summary"},"finish_reason":null}]}
data: {"id":"chatcmpl-...","object":"chat.completion.chunk","created":1710000000,"model":"gemini-3.5-flash-text-to-text","choices":[{"index":0,"delta":{},"finish_reason":"stop"}]}
data: {"id":"chatcmpl-...","object":"chat.completion.chunk","created":1710000000,"model":"gemini-3.5-flash-text-to-text","choices":[],"usage":{"prompt_tokens":12,"completion_tokens":8,"total_tokens":20}}
data: [DONE]
Fehlerantwort
event: error
data: {"error":{"message":"API requests too frequent, exceeding rate limit","type":"rate_limit_error","code":"1302","param":null}}
Bewährte Praktiken
- Strukturierte Nachrichten verwenden: Sende den Gesprächsverlauf über
messages[], anstatt den Kontext in einem einzigen Prompt zusammenzufassen.
- Die richtige Modellvariante wählen: Verwende Text to Text für reine Textaufgaben, Image to Text für Bildverständnis und File Analysis, wenn Dokumente oder gemischte Inhalte benötigt werden.
- SSE-Ereignisse verarbeiten: Hänge
choices[0].delta.content für die Streaming-Anzeige an und behandle data: [DONE] als erfolgreichen Abschluss.
- Nur unterstützte Eingaben senden: Textmodelle ignorieren im aktuellen Client nicht unterstützte Datei- und Bildteile.