Integra Vortixy Chat nei tuoi progetti con una chiave Bearer. Stesse quote del tuo piano. Crea chiave in Dashboard → La tua API
Invia un messaggio con cronologia opzionale, abilita la ricerca web o regola la profondità di ragionamento. Conta verso i limiti giornalieri di messaggi e token del tuo piano.
Pronto per chatbot: aggiungi system per la persona del tuo bot (8k chars), stream:true per delta NDJSON, webSearch per fonti live. Ogni chiamata scala token, richieste e ricerche web — condivisi con il web (reset 00:00 UTC).
Tutti i campi JSON per POST /api/v1/chat. Solo message è obbligatorio — il resto è opzionale.
| Campo | Tipo | Obbligatorio | Descrizione |
|---|---|---|---|
| message | string | yes | Messaggio utente (1–32.000 caratteri). |
| history | array | no | Turni precedenti: [{ role: 'user' | 'assistant', content: string }] (max 20). |
| system | string | no | Persona del bot / istruzioni (max 8k). Accetta anche alias instructions. |
| provider | string | no | gemini (gestito, predefinito). Per BYOK, salva la chiave in Dashboard → Impostazioni → Chiavi IA e poi usa provider openai / anthropic / deepseek / gemini con un model id BYOK. |
| model | string | no | ID modello — vedi tabella sotto. Deve essere consentito per provider + piano. |
| chatMode | enum | no | Modalità chat: fast | smart | auto (auto sceglie per richiesta). Preferita a model — ometti model se impostata. Accetta anche l'alias engineMode. |
| thinkingLevel | enum | no | none | low | medium | high | xhigh | max (xhigh/max solo per GPT-5.6). |
| webSearch | boolean | no | true per attivare la ricerca web live. Conta nella quota giornaliera. |
| stream | boolean | no | true per streaming NDJSON (delta → done). Default false (JSON). |
Suggerimento: ometti provider/model per usare il modello migliore del tuo piano. Imposta stream:true per i chatbot.
Il ragionamento ora è automatico per modalità — Fast usa low, Smart usa high, Auto sceglie con intelligenza in base alla tua domanda. Non devi più scegliere un livello. L'API accetta ancora thinkingLevel, ma è opzionale e la modalità lo sovrascrive.
| Livello | Descrizione |
|---|---|
| none | Nessun ragionamento — più veloce, consumo minimo. |
| low | Ragionamento leggero — risposte rapide con un po' di riflessione. |
| medium | Ragionamento bilanciato — buono per la maggior parte dei task. |
| high | Ragionamento profondo — per analisi e pianificazione complesse. |
| xhigh | Molto profondo — solo modelli frontier, per problemi difficili. |
| max | Massimo — usa tutto il budget di ragionamento, più lento e costoso. |
Fast → low/medium, Smart → medium/high, Auto → sceglie low→high per richiesta. Nessun selettore per l'utente — il sistema è intelligente in ogni momento.
Scegli una modalità, non un modello. Vortixy instrada con intelligenza per sfruttare al meglio la tua quota — senza scegliere un modello specifico. Tutte le modalità condividono le quote del tuo piano (00:00 UTC).
| Modalità | Quando usarla | Ragionamento | Velocità |
|---|---|---|---|
| Fast | Risposte rapide per attività semplici — il più veloce ed efficiente. | low / medium | Velocissima |
| Smart | Ragionamento approfondito per attività complesse — più profondo, passo passo. | medium / high | Approfondita |
| Auto | Vortixy sceglie da solo la modalità migliore per ogni domanda (euristica, senza costi extra). | auto (low → high) | Auto-sceglie |
Smart mostra un bagliore animato quando selezionato (bordo shimmer). Auto è il default — consigliato per la maggior parte delle attività.
Salva la chiave una volta in Dashboard → Settings → AI keys (AES-256-GCM, mai esposta) — poi chiama con provider + model. Turni BYOK senza token modello; resto come managed (00:00 UTC).
| Modello | Provider | Accesso | Ragionamento | Contesto |
|---|---|---|---|---|
| gemini-3.8-flash:byok | gemini | Free — serve la chiave | none, low, medium, high | 1049k |
| gemini-3.7-flash:byok | gemini | Free — serve la chiave | none, low, medium, high | 1049k |
| gemini-3.6-flash:byok | gemini | Free — serve la chiave | none, low, medium, high | 1049k |
| gemini-3.5-flash-lite:byok | gemini | Free — serve la chiave | none, low, medium, high | 1049k |
| gemini-3.1-flash-lite:byok | gemini | Free — serve la chiave | none, low, medium, high | 1049k |
| gpt-6-astra | openai | Free — serve la chiave | low, medium, high, xhigh, max | 1050k |
| gpt-6-sol | openai | Free — serve la chiave | none, low, medium, high, xhigh, max | 1049k |
| gpt-6-luna | openai | Free — serve la chiave | none, low, medium, high, xhigh, max | 1049k |
| gpt-5.6-sol | openai | Free — serve la chiave | none, low, medium, high, xhigh, max | 1049k |
| gpt-5.6-terra | openai | Free — serve la chiave | none, low, medium, high, xhigh, max | 1049k |
| gpt-5.6-luna | openai | Free — serve la chiave | none, low, medium, high, xhigh, max | 1049k |
| claude-fable-5-1 | anthropic | Free — serve la chiave | none, high | 1049k |
| claude-opus-5-5 | anthropic | Free — serve la chiave | none, high | 1049k |
| claude-sonnet-5-5 | anthropic | Free — serve la chiave | none, high | 1049k |
| claude-fable-5 | anthropic | Free — serve la chiave | none, high | 1049k |
| claude-opus-5 | anthropic | Free — serve la chiave | none, high | 1049k |
| claude-sonnet-5 | anthropic | Free — serve la chiave | none, high | 1049k |
| claude-haiku-4-5 | anthropic | Free — serve la chiave | none, high | 200k |
| claude-sonnet-4-6 | anthropic | Free — serve la chiave | none, high | 200k |
| deepseek-flash | deepseek | Free — serve la chiave | none, high | 1049k |
| deepseek-v4-pro | deepseek | Free — serve la chiave | none, high | 1049k |
| deepseek-v4-flash | deepseek | Free — serve la chiave | none, high | 1049k |
Catalogo BYOK: gpt-6-astra/gpt-6-sol/gpt-6-luna/gpt-5.6-sol/terra/luna (openai), claude-fable-5-1/opus-5-5/sonnet-5-5/fable-5/opus-5/sonnet-5/haiku-4-5 (anthropic), deepseek-flash/v4-pro/v4-flash (deepseek), gemini-3.8/3.7/3.6-flash:byok + flash-lites:byok (gemini). Tutti disponibili in Free con chiave presente — nessun upgrade.
Le finestre di contesto variano in base al modello e arrivano a 1M (3.8/3.7) / 256k token. Ogni richiesta è inoltre limitata dal piano: Free 30k, Go 30k, Pro 35k, Max 50k o Max+ 100k token. Oltre l’85% del budget applicabile, i turni meno recenti vengono compattati automaticamente. L’API pubblica è stateless: reinvia la cronologia in `history` a ogni chiamata.
Il contesto utilizzabile è il valore minore tra il limite del piano (30k/30k/35k/50k/100k token) e la finestra del modello (fino a 1M (3.8/3.7) / 256k (altri)). Nelle integrazioni, invia solo la cronologia che rientra nel budget e riassumi i turni più vecchi; richieste troppo grandi possono essere rifiutate.
// Basic — only required field
{ "message": "Explain quantum computing simply" }
// With system persona (chatbot) — 8k chars
{
"message": "What do you sell?",
"system": "You are a friendly shop assistant for Acme. Answer briefly in Spanish."
}
// With history — multi-turn conversation
{
"message": "And what about entanglement?",
"history": [
{ "role": "user", "content": "Hi" },
{ "role": "assistant", "content": "Hello! How can I help?" },
{ "role": "user", "content": "Explain quantum computing" }
]
}
// Full request — all variants together
{
"message": "Latest AI breakthroughs with sources",
"history": [{ "role": "user", "content": "Hi" }],
"system": "You are Acme support. Be concise.",
"provider": "gemini", // managed (default) — BYOK dashboard-only for other providers
"model": "gemini-3.5-flash", // any available managed model from the table above
"thinkingLevel": "high", // none | low | medium | high | xhigh | max
"webSearch": true, // true = live web grounding + sources[]
"stream": false // false = JSON, true = NDJSON deltas
}// 200 OK — non-streaming response
{
"answer": "Quantum computing uses qubits...",
"model": "gemini-3.5-flash",
"usage": { "tokensEstimated": 42, "tokensReserved": 42, "webSearch": true },
"sources": [{ "title": "Quantum computing — Wikipedia", "url": "https://en.wikipedia.org/wiki/Quantum_computing" }],
"plan": "pro",
"resetsInSeconds": 43200,
"resetsAt": "2026-09-25T00:00:00.000Z"
} // quotas reset daily at 00:00 UTC — streaming uses RateLimit-* headers// Streaming (NDJSON) — ideal for chatbots
const res = await fetch("https://www.vortixy.net/api/v1/chat", {
method: "POST",
headers: { Authorization: "Bearer " + process.env.VORTIXY_API_KEY, "Content-Type": "application/json" },
body: JSON.stringify({ message: "Hello!", system: "Be concise.", stream: true, thinkingLevel: "medium" })
});
for await (const line of res.body.pipeThrough(new TextDecoderStream()).pipeThrough(splitNDJSON())) {
const evt = JSON.parse(line);
if (evt.type === "delta") process.stdout.write(evt.delta);
if (evt.type === "sources") console.log("sources", evt.sources);
if (evt.type === "done") console.log("\nusage", evt.usage);
}
// curl variants — pick the one you need
// 1) Basic
curl -X POST https://www.vortixy.net/api/v1/chat -H "Authorization: Bearer $VORTIXY_API_KEY" -H "Content-Type: application/json" -d '{"message":"Hello"}'
// 2) With web search + high thinking
curl -X POST https://www.vortixy.net/api/v1/chat -H "Authorization: Bearer $VORTIXY_API_KEY" -H "Content-Type: application/json" -d '{"message":"Latest AI news","webSearch":true,"thinkingLevel":"high"}'
// 3) Chatbot persona + streaming (NDJSON)
curl -N -X POST https://www.vortixy.net/api/v1/chat -H "Authorization: Bearer $VORTIXY_API_KEY" -H "Content-Type: application/json" -d '{"message":"Hi","system":"You are Acme bot.","stream":true}'Lo streaming emette delta → done (NDJSON, Content-Type: application/x-ndjson). Con webSearch:true ricevi anche fonti ancorate via ricerca web (conta nella quota giornaliera).
Mensile o annuale?
Risparmi il 17% con la fatturazione annuale.