Integriere Vortixy Chat mit einem Bearer-Schlüssel in deine Projekte — Chat-, Bild- und Nutzungs-Endpunkte mit denselben Quotas wie dein Plan. Doku und Beispiele inklusive. Schlüssel erstellen in Dashboard → Deine API
Sende eine Nachricht mit optionalem Gesprächsverlauf, aktiviere die Websuche oder passe die Denktiefe an. Wird auf das tägliche Nachrichten- und Token-Limit deines Plans angerechnet.
Chatbot-bereit: füge system für deine Bot-Persona hinzu (8k Zeichen), stream:true für NDJSON-Deltas, webSearch für Live-Quellen. Jeder Aufruf zieht Tokens, Requests und web-Suchen ab — geteilt mit Web (Reset 00:00 UTC).
Alle JSON-Felder für POST /api/v1/chat. Nur message ist erforderlich — alles andere optional.
| Feld | Typ | Erforderlich | Beschreibung |
|---|---|---|---|
| message | string | yes | Nutzer-Nachricht (1–32.000 Zeichen). |
| history | array | no | Vorherige Turns: [{ role: 'user' | 'assistant', content: string }] (max. 20). |
| system | string | no | Bot-Persona / Anweisungen (max. 8k). Akzeptiert auch Alias instructions. |
| provider | string | no | gemini (verwaltet, Standard). Für BYOK speichere deinen Schlüssel unter Dashboard → Einstellungen → KI-Schlüssel und nutze dann provider openai / anthropic / deepseek / gemini mit einer BYOK-Model-ID. |
| model | string | no | Model-ID — siehe Tabelle unten. Muss für Provider + Plan erlaubt sein. |
| chatMode | enum | no | Chat-Modus: fast | smart | auto (auto wählt pro Anfrage). Bevorzugt gegenüber model — model weglassen. Akzeptiert auch den Alias engineMode. |
| thinkingLevel | enum | no | none | low | medium | high | xhigh | max (xhigh/max nur für GPT-5.6). |
| webSearch | boolean | no | true für Live-Websuche. Zählt zum täglichen Such-Kontingent. |
| stream | boolean | no | true für NDJSON-Streaming (delta → done). Standard false (JSON). |
Tipp: provider/model weglassen für bestes Modell deines Plans. Setze stream:true für Chatbots.
Das Reasoning ist jetzt pro Modus automatisch — Fast nutzt low, Smart nutzt high, Auto wählt intelligent passend zu deiner Frage. Du musst keine Stufe mehr wählen. Die API akzeptiert thinkingLevel weiterhin, aber es ist optional und der Modus überschreibt es.
| Stufe | Beschreibung |
|---|---|
| none | Kein Nachdenken — am schnellsten, geringster Verbrauch. |
| low | Leichtes Nachdenken — schnelle Antworten mit etwas Überlegung. |
| medium | Ausgewogen — gut für die meisten Aufgaben. |
| high | Tiefe Überlegung — für komplexe Analyse und Planung. |
| xhigh | Sehr tief — nur Frontier-Modelle, für schwere Probleme. |
| max | Maximum — nutzt volles Thinking-Budget, am langsamsten und teuersten. |
Fast → low/medium, Smart → medium/high, Auto → wählt low→high pro Anfrage. Kein UI-Schalter nötig — das System ist in jedem Moment intelligent.
Wähle einen Modus, kein Modell. Vortixy routet intelligent, damit du deine Quote optimal nutzt — ganz ohne Modellwahl. Alle Modi teilen die Quoten deines Tarifs (00:00 UTC).
| Modus | Wann nutzen | Thinking | Tempo |
|---|---|---|---|
| Fast | Schnelle Antworten für einfache Aufgaben — am schnellsten und sparsamsten. | low / medium | Am schnellsten |
| Smart | Tiefes Reasoning für komplexe Aufgaben — gründlicher, Schritt für Schritt. | medium / high | Gründlich |
| Auto | Vortixy wählt automatisch den besten Modus pro Frage (Heuristik, ohne Mehrkosten). | auto (low → high) | Auto-Wahl |
Smart zeigt nach Auswahl ein animiertes Glühen (Shimmer-Rahmen). Auto ist Standard — für die meisten Aufgaben empfohlen.
Speichere deinen Key einmal in Dashboard → Settings → AI keys (AES-256-GCM, nie exponiert) — dann mit provider + model aufrufen. BYOK-Turns ohne Modell-Token; Rest wie managed (00:00 UTC).
| Modell | Anbieter | Zugriff | Thinking | Kontext |
|---|---|---|---|---|
| gemini-3.8-flash:byok | gemini | Free — Schlüssel nötig | none, low, medium, high | 1049k |
| gemini-3.7-flash:byok | gemini | Free — Schlüssel nötig | none, low, medium, high | 1049k |
| gemini-3.6-flash:byok | gemini | Free — Schlüssel nötig | none, low, medium, high | 1049k |
| gemini-3.5-flash-lite:byok | gemini | Free — Schlüssel nötig | none, low, medium, high | 1049k |
| gemini-3.1-flash-lite:byok | gemini | Free — Schlüssel nötig | none, low, medium, high | 1049k |
| gpt-6-astra | openai | Free — Schlüssel nötig | low, medium, high, xhigh, max | 1050k |
| gpt-6-sol | openai | Free — Schlüssel nötig | none, low, medium, high, xhigh, max | 1049k |
| gpt-6-luna | openai | Free — Schlüssel nötig | none, low, medium, high, xhigh, max | 1049k |
| gpt-5.6-sol | openai | Free — Schlüssel nötig | none, low, medium, high, xhigh, max | 1049k |
| gpt-5.6-terra | openai | Free — Schlüssel nötig | none, low, medium, high, xhigh, max | 1049k |
| gpt-5.6-luna | openai | Free — Schlüssel nötig | none, low, medium, high, xhigh, max | 1049k |
| claude-fable-5-1 | anthropic | Free — Schlüssel nötig | none, high | 1049k |
| claude-opus-5-5 | anthropic | Free — Schlüssel nötig | none, high | 1049k |
| claude-sonnet-5-5 | anthropic | Free — Schlüssel nötig | none, high | 1049k |
| claude-fable-5 | anthropic | Free — Schlüssel nötig | none, high | 1049k |
| claude-opus-5 | anthropic | Free — Schlüssel nötig | none, high | 1049k |
| claude-sonnet-5 | anthropic | Free — Schlüssel nötig | none, high | 1049k |
| claude-haiku-4-5 | anthropic | Free — Schlüssel nötig | none, high | 200k |
| claude-sonnet-4-6 | anthropic | Free — Schlüssel nötig | none, high | 200k |
| deepseek-flash | deepseek | Free — Schlüssel nötig | none, high | 1049k |
| deepseek-v4-pro | deepseek | Free — Schlüssel nötig | none, high | 1049k |
| deepseek-v4-flash | deepseek | Free — Schlüssel nötig | none, high | 1049k |
BYOK-Katalog: gpt-6-astra/gpt-6-sol/gpt-6-luna/gpt-5.6-sol/terra/luna (openai), claude-fable-5-1/opus-5-5/sonnet-5-5/fable-5/opus-5/sonnet-5/haiku-4-5 (anthropic), deepseek-flash/v4-pro/v4-flash (deepseek), gemini-3.8/3.7/3.6-flash:byok + flash-lites:byok (gemini). Alle in Free verfügbar, sobald ein Schlüssel hinterlegt ist — kein Upgrade nötig.
Modell-Kontextfenster variieren und reichen bis zu 1M (3.8/3.7) / 256k (Rest) Tokens. Jede Anfrage ist außerdem durch den Tarif begrenzt: Free 30k, Go 30k, Pro 35k, Max 50k oder Max+ 100k Tokens. Oberhalb von 85% des geltenden Budgets werden ältere Turns im Chat automatisch kompaktiert. Die öffentliche API ist zustandslos; sende die History bei jedem Aufruf erneut.
Der nutzbare Kontext ist der kleinere Wert aus Tariflimit (30k/30k/35k/50k/100k Tokens) und Modellfenster (bis zu 1M (3.8/3.7) / 256k (Rest)). Sende bei Integrationen nur passende History und fasse ältere Turns selbst zusammen; zu große Anfragen können abgelehnt werden.
// Basic — only required field
{ "message": "Explain quantum computing simply" }
// With system persona (chatbot) — 8k chars
{
"message": "What do you sell?",
"system": "You are a friendly shop assistant for Acme. Answer briefly in Spanish."
}
// With history — multi-turn conversation
{
"message": "And what about entanglement?",
"history": [
{ "role": "user", "content": "Hi" },
{ "role": "assistant", "content": "Hello! How can I help?" },
{ "role": "user", "content": "Explain quantum computing" }
]
}
// Full request — all variants together
{
"message": "Latest AI breakthroughs with sources",
"history": [{ "role": "user", "content": "Hi" }],
"system": "You are Acme support. Be concise.",
"provider": "gemini", // managed (default) — BYOK dashboard-only for other providers
"model": "gemini-3.5-flash", // any available managed model from the table above
"thinkingLevel": "high", // none | low | medium | high | xhigh | max
"webSearch": true, // true = live web grounding + sources[]
"stream": false // false = JSON, true = NDJSON deltas
}// 200 OK — non-streaming response
{
"answer": "Quantum computing uses qubits...",
"model": "gemini-3.5-flash",
"usage": { "tokensEstimated": 42, "tokensReserved": 42, "webSearch": true },
"sources": [{ "title": "Quantum computing — Wikipedia", "url": "https://en.wikipedia.org/wiki/Quantum_computing" }],
"plan": "pro",
"resetsInSeconds": 43200,
"resetsAt": "2026-09-25T00:00:00.000Z"
} // quotas reset daily at 00:00 UTC — streaming uses RateLimit-* headers// Streaming (NDJSON) — ideal for chatbots
const res = await fetch("https://www.vortixy.net/api/v1/chat", {
method: "POST",
headers: { Authorization: "Bearer " + process.env.VORTIXY_API_KEY, "Content-Type": "application/json" },
body: JSON.stringify({ message: "Hello!", system: "Be concise.", stream: true, thinkingLevel: "medium" })
});
for await (const line of res.body.pipeThrough(new TextDecoderStream()).pipeThrough(splitNDJSON())) {
const evt = JSON.parse(line);
if (evt.type === "delta") process.stdout.write(evt.delta);
if (evt.type === "sources") console.log("sources", evt.sources);
if (evt.type === "done") console.log("\nusage", evt.usage);
}
// curl variants — pick the one you need
// 1) Basic
curl -X POST https://www.vortixy.net/api/v1/chat -H "Authorization: Bearer $VORTIXY_API_KEY" -H "Content-Type: application/json" -d '{"message":"Hello"}'
// 2) With web search + high thinking
curl -X POST https://www.vortixy.net/api/v1/chat -H "Authorization: Bearer $VORTIXY_API_KEY" -H "Content-Type: application/json" -d '{"message":"Latest AI news","webSearch":true,"thinkingLevel":"high"}'
// 3) Chatbot persona + streaming (NDJSON)
curl -N -X POST https://www.vortixy.net/api/v1/chat -H "Authorization: Bearer $VORTIXY_API_KEY" -H "Content-Type: application/json" -d '{"message":"Hi","system":"You are Acme bot.","stream":true}'Streaming sendet delta → done (NDJSON, Content-Type: application/x-ndjson). Mit webSearch:true erhältst du zusätzlich geerdete Quellen via Websuche (zählt zum täglichen Kontingent).
Monatlich oder jährlich?
Sie sparen 17% mit jährlicher Abrechnung.