Integra Vortixy Chat en tus propios proyectos con una clave Bearer. Mismas cuotas que tu plan. Crear clave en Dashboard → Tu API
Envía un mensaje con historial de conversación, activa la búsqueda web o ajusta la profundidad de análisis. Se descuenta del límite diario de mensajes y tokens de tu plan.
Listo para chatbots: añade system para la persona de tu bot (8k caracteres), stream:true para deltas NDJSON, webSearch para fuentes en vivo. Cada llamada descuenta tokens, solicitudes y búsquedas web — compartidas con la web (reinicio 00:00 UTC).
Todos los campos JSON para POST /api/v1/chat. Solo message es obligatorio — el resto es opcional.
| Campo | Tipo | Obligatorio | Descripción |
|---|---|---|---|
| message | string | yes | Mensaje del usuario (1–32.000 caracteres). |
| history | array | no | Turnos previos: [{ role: 'user' | 'assistant', content: string }] (máx. 20). |
| system | string | no | Persona del bot / instrucciones (máx. 8k). También acepta alias instructions. |
| provider | string | no | gemini (gestionado, por defecto). Para BYOK, guarda tu clave en Dashboard → Ajustes → Claves de IA y luego usa provider openai / anthropic / deepseek / gemini con un modelo BYOK. |
| model | string | no | ID del modelo — ver tabla abajo. Debe estar permitido para tu proveedor y plan. |
| chatMode | enum | no | Modo de chat: fast | smart | auto (auto elige por consulta). Preferido sobre model — omite model al usarlo. También acepta el alias engineMode. |
| thinkingLevel | enum | no | none | low | medium | high | xhigh | max (xhigh/max solo para GPT-5.6). |
| webSearch | boolean | no | true para activar la búsqueda web en vivo. Cuenta en tu cuota diaria. |
| stream | boolean | no | true para streaming NDJSON (delta → done). Por defecto false (JSON). |
Consejo: omite provider/model para usar el mejor modelo de tu plan. Usa stream:true para chatbots.
El razonamiento ya es automático por modo — Fast usa low, Smart usa high, Auto elige con inteligencia según tu pregunta. Ya no necesitas elegir nivel. La API aún acepta thinkingLevel, pero es opcional y el modo lo sobrescribe.
| Nivel | Descripción |
|---|---|
| none | Sin razonamiento — más rápido, menor consumo de tokens. |
| low | Razonamiento ligero — respuestas rápidas con un poco de reflexión. |
| medium | Razonamiento equilibrado — bueno para la mayoría de tareas. |
| high | Razonamiento profundo — para análisis y planificación complejos. |
| xhigh | Muy profundo — solo modelos frontera, para problemas difíciles. |
| max | Máximo — usa todo el presupuesto de razonamiento, más lento y costoso. |
Fast → low/medium, Smart → medium/high, Auto → elige low→high por consulta. Sin selector para el usuario — el sistema es inteligente en cada momento.
Elige un modo, no un modelo. Vortixy enruta con inteligencia para aprovechar tu cuota — sin elegir un modelo concreto. Todos los modos comparten las cuotas de tu plan (00:00 UTC).
| Modo | Cuándo usarlo | Pensamiento | Velocidad |
|---|---|---|---|
| Fast | Respuestas rápidas para tareas simples — lo más veloz y eficiente. | low / medium | El más rápido |
| Smart | Razonamiento a fondo para tareas complejas — más profundo, paso a paso. | medium / high | Exhaustivo |
| Auto | Vortixy elige el mejor modo solo para cada pregunta (heurística, sin costo extra). | auto (low → high) | Auto-elige |
Smart muestra un brillo animado al seleccionarlo (borde shimmer). Auto viene por defecto — recomendado para la mayoría de tareas.
Guarda tu clave de proveedor una vez en Dashboard → Ajustes → Claves de IA (cifrado AES-256-GCM, nunca expuesta) — luego llama con provider + model. Los turnos BYOK no consumen tokens del modelo; las demás cuotas igual que los gestionados (00:00 UTC).
| Modelo | Proveedor | Acceso | Pensamiento | Contexto |
|---|---|---|---|---|
| gemini-3.8-flash:byok | gemini | Free — necesita clave | none, low, medium, high | 1049k |
| gemini-3.7-flash:byok | gemini | Free — necesita clave | none, low, medium, high | 1049k |
| gemini-3.6-flash:byok | gemini | Free — necesita clave | none, low, medium, high | 1049k |
| gemini-3.5-flash-lite:byok | gemini | Free — necesita clave | none, low, medium, high | 1049k |
| gemini-3.1-flash-lite:byok | gemini | Free — necesita clave | none, low, medium, high | 1049k |
| gpt-6-astra | openai | Free — necesita clave | low, medium, high, xhigh, max | 1050k |
| gpt-6-sol | openai | Free — necesita clave | none, low, medium, high, xhigh, max | 1049k |
| gpt-6-luna | openai | Free — necesita clave | none, low, medium, high, xhigh, max | 1049k |
| gpt-5.6-sol | openai | Free — necesita clave | none, low, medium, high, xhigh, max | 1049k |
| gpt-5.6-terra | openai | Free — necesita clave | none, low, medium, high, xhigh, max | 1049k |
| gpt-5.6-luna | openai | Free — necesita clave | none, low, medium, high, xhigh, max | 1049k |
| claude-fable-5-1 | anthropic | Free — necesita clave | none, high | 1049k |
| claude-opus-5-5 | anthropic | Free — necesita clave | none, high | 1049k |
| claude-sonnet-5-5 | anthropic | Free — necesita clave | none, high | 1049k |
| claude-fable-5 | anthropic | Free — necesita clave | none, high | 1049k |
| claude-opus-5 | anthropic | Free — necesita clave | none, high | 1049k |
| claude-sonnet-5 | anthropic | Free — necesita clave | none, high | 1049k |
| claude-haiku-4-5 | anthropic | Free — necesita clave | none, high | 200k |
| claude-sonnet-4-6 | anthropic | Free — necesita clave | none, high | 200k |
| deepseek-flash | deepseek | Free — necesita clave | none, high | 1049k |
| deepseek-v4-pro | deepseek | Free — necesita clave | none, high | 1049k |
| deepseek-v4-flash | deepseek | Free — necesita clave | none, high | 1049k |
Catálogo BYOK: gpt-6-astra/gpt-6-sol/gpt-6-luna/gpt-5.6-sol/terra/luna (openai), claude-fable-5-1/opus-5-5/sonnet-5-5/fable-5/opus-5/sonnet-5/haiku-4-5 (anthropic), deepseek-flash/v4-pro/v4-flash (deepseek), gemini-3.8/3.7/3.6-flash:byok + flash-lites:byok (gemini). Todos disponibles en Free con clave presente — sin subir de plan.
Las ventanas de contexto varían según el modelo y llegan hasta 1M (3.8/3.7) / 256k (resto) tokens. Cada petición también tiene el límite de tu plan: Free 30k, Go 30k, Pro 35k, Max 50k o Max+ 100k tokens. En Vortixy Chat (/chat), los turnos antiguos se compactan automáticamente al superar el 85% del presupuesto aplicable. La API pública (/api/v1/chat) no guarda estado: conserva el historial en tu app y reenvíalo en `history` en cada llamada.
El contexto útil por petición es el menor entre el límite de tu plan (30k/30k/35k/50k/100k tokens) y la ventana del modelo elegido (hasta 1M (3.8/3.7) / 256k (resto)). En integraciones, envía solo el historial que quepa en ese presupuesto y resume tú los turnos antiguos; las peticiones demasiado grandes pueden rechazarse en vez de aceptarse silenciosamente.
// Basic — only required field
{ "message": "Explain quantum computing simply" }
// With system persona (chatbot) — 8k chars
{
"message": "What do you sell?",
"system": "You are a friendly shop assistant for Acme. Answer briefly in Spanish."
}
// With history — multi-turn conversation
{
"message": "And what about entanglement?",
"history": [
{ "role": "user", "content": "Hi" },
{ "role": "assistant", "content": "Hello! How can I help?" },
{ "role": "user", "content": "Explain quantum computing" }
]
}
// Full request — all variants together
{
"message": "Latest AI breakthroughs with sources",
"history": [{ "role": "user", "content": "Hi" }],
"system": "You are Acme support. Be concise.",
"provider": "gemini", // managed (default) — BYOK dashboard-only for other providers
"model": "gemini-3.5-flash", // any available managed model from the table above
"thinkingLevel": "high", // none | low | medium | high | xhigh | max
"webSearch": true, // true = live web grounding + sources[]
"stream": false // false = JSON, true = NDJSON deltas
}// 200 OK — non-streaming response
{
"answer": "Quantum computing uses qubits...",
"model": "gemini-3.5-flash",
"usage": { "tokensEstimated": 42, "tokensReserved": 42, "webSearch": true },
"sources": [{ "title": "Quantum computing — Wikipedia", "url": "https://en.wikipedia.org/wiki/Quantum_computing" }],
"plan": "pro",
"resetsInSeconds": 43200,
"resetsAt": "2026-09-25T00:00:00.000Z"
} // quotas reset daily at 00:00 UTC — streaming uses RateLimit-* headers// Streaming (NDJSON) — ideal for chatbots
const res = await fetch("https://www.vortixy.net/api/v1/chat", {
method: "POST",
headers: { Authorization: "Bearer " + process.env.VORTIXY_API_KEY, "Content-Type": "application/json" },
body: JSON.stringify({ message: "Hello!", system: "Be concise.", stream: true, thinkingLevel: "medium" })
});
for await (const line of res.body.pipeThrough(new TextDecoderStream()).pipeThrough(splitNDJSON())) {
const evt = JSON.parse(line);
if (evt.type === "delta") process.stdout.write(evt.delta);
if (evt.type === "sources") console.log("sources", evt.sources);
if (evt.type === "done") console.log("\nusage", evt.usage);
}
// curl variants — pick the one you need
// 1) Basic
curl -X POST https://www.vortixy.net/api/v1/chat -H "Authorization: Bearer $VORTIXY_API_KEY" -H "Content-Type: application/json" -d '{"message":"Hello"}'
// 2) With web search + high thinking
curl -X POST https://www.vortixy.net/api/v1/chat -H "Authorization: Bearer $VORTIXY_API_KEY" -H "Content-Type: application/json" -d '{"message":"Latest AI news","webSearch":true,"thinkingLevel":"high"}'
// 3) Chatbot persona + streaming (NDJSON)
curl -N -X POST https://www.vortixy.net/api/v1/chat -H "Authorization: Bearer $VORTIXY_API_KEY" -H "Content-Type: application/json" -d '{"message":"Hi","system":"You are Acme bot.","stream":true}'El streaming emite delta → done (NDJSON, Content-Type: application/x-ndjson). Con webSearch:true también recibes fuentes verificadas vía búsqueda web (cuenta en tu cuota diaria).
¿Mensual o anual?
Ahorras 17% con facturación anual.