API Gemini 3.7 Flash
Tài liệu tham khảo API đầy đủ cho các mô hình LLM Gemini 3.7 Flash do Google cung cấp.
Các biến thể mô hình
API này hỗ trợ ba biến thể mô hình Gemini 3.7 Flash:
So sánh nhanh
Gemini 3.7 Flash Text to Text
Điểm cuối
POST /api/v1/chat/completions
Tham số yêu cầu
Bắt buộc
Hỗ trợ tin nhắn
Không bắt buộc
Yêu cầu mẫu
const response = await fetch('https://api.flaq.ai/api/v1/chat/completions', {
method: 'POST',
headers: {
'Authorization': 'Bearer YOUR_API_KEY',
'Accept': 'text/event-stream',
'Content-Type': 'application/json'
},
body: JSON.stringify({
model: 'gemini-3.7-flash-text-to-text',
messages: [
{
role: 'user',
content: 'Write a concise summary of how transformer models work.'
}
],
stream: true,
max_tokens: 2048
})
});
Gemini 3.7 Flash Image to Text
Điểm cuối
POST /api/v1/chat/completions
Tham số yêu cầu
Bắt buộc
Hỗ trợ tin nhắn
Không bắt buộc
Yêu cầu mẫu
const response = await fetch('https://api.flaq.ai/api/v1/chat/completions', {
method: 'POST',
headers: {
'Authorization': 'Bearer YOUR_API_KEY',
'Accept': 'text/event-stream',
'Content-Type': 'application/json'
},
body: JSON.stringify({
model: 'gemini-3.7-flash-image-to-text',
messages: [
{
role: 'user',
content: [
{ type: 'text', text: 'Describe the image and extract any visible text.' },
{
type: 'image_url',
image_url: {
url: 'https://example.com/sample-image.jpg'
}
}
]
}
],
stream: true,
max_tokens: 2048
})
});
Gemini 3.7 Flash File Analysis
Điểm cuối
POST /api/v1/chat/completions
Tham số yêu cầu
Bắt buộc
Hỗ trợ tin nhắn
Không bắt buộc
Yêu cầu mẫu
const response = await fetch('https://api.flaq.ai/api/v1/chat/completions', {
method: 'POST',
headers: {
'Authorization': 'Bearer YOUR_API_KEY',
'Accept': 'text/event-stream',
'Content-Type': 'application/json'
},
body: JSON.stringify({
model: 'gemini-3.7-flash-file-analysis',
messages: [
{
role: 'user',
content: [
{ type: 'text', text: 'Summarize this file and extract the key risks.' },
{
type: 'file',
file: {
filename: 'demo.pdf',
file_data: 'https://example.com/demo.pdf'
}
}
]
}
],
stream: true,
max_tokens: 4096
})
});
Định dạng phản hồi
Các mô hình LLM Gemini 3.7 Flash trả về phản hồi completion tương thích với OpenAI. Với stream: true, phản hồi là Server-Sent Events; với stream: false, phản hồi là một đối tượng JSON duy nhất.
Phản hồi thành công
data: {"id":"chatcmpl-...","object":"chat.completion.chunk","created":1710000000,"model":"gemini-3.7-flash-text-to-text","choices":[{"index":0,"delta":{"role":"assistant"},"finish_reason":null}]}
data: {"id":"chatcmpl-...","object":"chat.completion.chunk","created":1710000000,"model":"gemini-3.7-flash-text-to-text","choices":[{"index":0,"delta":{"content":"Here is a concise summary"},"finish_reason":null}]}
data: {"id":"chatcmpl-...","object":"chat.completion.chunk","created":1710000000,"model":"gemini-3.7-flash-text-to-text","choices":[{"index":0,"delta":{},"finish_reason":"stop"}]}
data: {"id":"chatcmpl-...","object":"chat.completion.chunk","created":1710000000,"model":"gemini-3.7-flash-text-to-text","choices":[],"usage":{"prompt_tokens":12,"completion_tokens":8,"total_tokens":20}}
data: [DONE]
Phản hồi lỗi
event: error
data: {"error":{"message":"API requests too frequent, exceeding rate limit","type":"rate_limit_error","code":"1302","param":null}}
Các phương pháp hay nhất
- Sử dụng tin nhắn có cấu trúc: Gửi lịch sử hội thoại qua
messages[] thay vì gộp ngữ cảnh vào một prompt.
- Chọn đúng biến thể mô hình: Sử dụng Text to Text cho các tác vụ chỉ có văn bản, Image to Text để hiểu hình ảnh và File Analysis khi cần xử lý tài liệu hoặc nội dung hỗn hợp.
- Tiếp nhận sự kiện SSE: Nối thêm
choices[0].delta.content để hiển thị truyền phát và coi data: [DONE] là hoàn tất thành công.
- Chỉ gửi đầu vào được hỗ trợ: Các mô hình văn bản bỏ qua phần tệp và hình ảnh không được hỗ trợ trong ứng dụng khách hiện tại.