Claude Code Guide
Set up Flaq AI Claude models and explore Claude Code skills
Free to try GPT 5.6 Luna API for fast chat, writing, coding, summaries, and affordable high-throughput OpenAI production workflows. Lower cost for coding than Claude Opus 4.8.
const response = await fetch('https://api.flaq.ai/api/v1/chat/completions', {
method: 'POST',
headers: {
Authorization: 'Bearer YOUR_API_KEY',
Accept: 'text/event-stream',
'Content-Type': 'application/json'
},
body: JSON.stringify({
model: 'gpt-5.6-luna-text-to-text',
messages: [
{
role: 'user',
content: 'Write a concise product update for a developer audience.'
}
],
stream: true,
max_tokens: 2048
})
});
const reader = response.body.getReader();
const decoder = new TextDecoder();
let buffer = '';
let assistantText = '';
while (true) {
const { done, value } = await reader.read();
if (done) break;
buffer += decoder.decode(value, { stream: true });
const frames = buffer.split('\n\n');
buffer = frames.pop() || '';
for (const frame of frames) {
const lines = frame.split('\n').filter(Boolean);
let eventName = 'message';
const dataLines = [];
for (const line of lines) {
if (line.startsWith('event:')) {
eventName = line.slice(6).trim();
} else if (line.startsWith('data:')) {
dataLines.push(line.replace(/^data:\s*/, ''));
}
}
const raw = dataLines.join('\n').trim();
if (raw === '[DONE]') {
console.log('\nFinal text:', assistantText);
continue;
}
let payload;
try {
payload = JSON.parse(raw);
} catch {
continue;
}
if (eventName === 'error' || payload.error) {
const msg = payload.error?.message ?? payload.message ?? 'Chat request failed';
throw new Error(msg);
}
const delta = payload.choices?.[0]?.delta;
if (delta?.content) {
assistantText += delta.content;
console.log(assistantText);
}
}
}
import json
import requests
response = requests.post(
'https://api.flaq.ai/api/v1/chat/completions',
headers={
'Authorization': 'Bearer YOUR_API_KEY',
'Accept': 'text/event-stream',
'Content-Type': 'application/json',
},
json={
'model': 'gpt-5.6-luna-text-to-text',
'messages': [
{
'role': 'user',
'content': 'Write a concise product update for a developer audience.',
}
],
'stream': True,
'max_tokens': 2048,
},
stream=True,
)
response.raise_for_status()
event_name = 'message'
assistant_text = ''
for raw_line in response.iter_lines(decode_unicode=True):
if not raw_line:
event_name = 'message'
continue
if raw_line.startswith('event:'):
event_name = raw_line.replace('event:', '', 1).strip()
continue
if raw_line.startswith('data:'):
raw_data = raw_line.replace('data:', '', 1).strip()
if raw_data == '[DONE]':
print('\nFinal text:', assistant_text)
continue
payload = json.loads(raw_data)
if event_name == 'error' or payload.get('error'):
error = payload.get('error') or payload
raise RuntimeError(error.get('message', 'Chat request failed'))
choices = payload.get('choices') or []
if choices:
delta = choices[0].get('delta') or {}
content = delta.get('content')
if content:
assistant_text += content
print(content, end='', flush=True)
curl -N -X POST "https://api.flaq.ai/api/v1/chat/completions" \
-H "Authorization: Bearer YOUR_API_KEY" \
-H "Accept: text/event-stream" \
-H "Content-Type: application/json" \
-d '{
"model": "gpt-5.6-luna-text-to-text",
"messages": [
{
"role": "user",
"content": "Write a concise product update for a developer audience."
}
],
"stream": true,
"max_tokens": 2048
}'
| Parameters | Price | Original Price | Discount |
|---|
GPT 5.6 Luna Text-to-Text API on Flaq AI delivers the fastest and most cost-efficient experience in OpenAI's GPT-5.6 family. Designed for responsive and high-throughput production workflows, Luna helps developers build chat, writing, summarization, coding, and knowledge features that need reliable output at scale. With multi-turn conversation context, streaming responses, and flexible output controls, teams can integrate GPT capabilities without managing model infrastructure or provider-specific plumbing.
Note Please ensure your prompts and application workflows comply with OpenAI's safety and usage guidelines. If an error occurs, review the input for restricted content, simplify the request, and try again.
GPT 5.6 Luna vs. GPT 5.6 Sol
GPT 5.6 Sol is the flagship choice for the most complex reasoning workloads. GPT 5.6 Luna prioritizes faster, more cost-efficient text processing for scalable applications.
GPT 5.6 Luna vs. GPT 5.6 Terra
GPT 5.6 Terra balances strong professional capability with lower cost. GPT 5.6 Luna goes further toward speed and operating efficiency for high-throughput use.
GPT 5.6 Luna vs. GPT 5.5
GPT 5.5 supports complex professional work across a broad range of tasks. GPT 5.6 Luna provides a faster, more cost-efficient path for recurring production text workloads.
GPT 5.6 Luna vs. Claude Sonnet 5
Claude Sonnet 5 focuses on strong agentic reasoning and coding within the Anthropic ecosystem. GPT 5.6 Luna offers a fast OpenAI-native option for scalable text applications.
GPT 5.6 Luna vs. Open Models
Open models provide deployment control but require serving and optimization work. GPT 5.6 Luna offers managed, cost-efficient text generation through Flaq AI's API workflow.
Set up Flaq AI Claude models and explore Claude Code skills

Set up Flaq AI GPT models and explore Codex skills
Use Flaq AI LLM models in Hermes Agent
Use GLM 5.2, Kimi K3, and DeepSeek v4 in ZCode
Run DeepSeek Harness with Flaq AI DeepSeek models
Connect AI agents to Flaq image and video generation tools