NoticiasNews

Claude Haiku 5.5: modelo rápido y económico con 1 millón de tokens para clasificación y enrutamientoClaude Haiku 5.5: Fast, Low-Cost Model With 1M Tokens for Classification and Routing

2026-10-09

Anthropic lanzó el 7 de octubre de 2026 Claude Haiku 5.5, su modelo más rápido, orientado a tareas de alto volumen y baja latencia como clasificación, enrutamiento, extracción de datos y subagentes. Según la documentación oficial, ofrece una ventana de contexto de 1 millón de tokens y hasta 128K tokens de salida.

Datos clave

  • Precio: US$0.10 por millón de tokens de entrada y US$0.50 de salida en prompts de hasta 100K tokens; la API por lotes da 50% de descuento.
  • Entrada de texto e imágenes, salida de texto; corte de conocimiento en junio de 2026.
  • Razonamiento adaptativo activado por defecto, con parámetro effort para ajustar profundidad.
  • Disponible en la API de Claude, Google Cloud, Microsoft Foundry, Amazon Bedrock y Claude Platform on AWS.

En la familia 5.5 es la opción más barata, por debajo de Sonnet 5.5 (US$2/US$10), Opus 5.5 (US$4/US$20) y Fable 5.1 (US$10/US$50). La documentación advierte que el nuevo tokenizador usa alrededor de 30% más tokens que Haiku 4.5 para el mismo texto, y que no se deben enviar temperature, top_p ni top_k.

Por qué interesa a las empresas

Modelos pequeños y baratos hacen viable clasificar cada ticket, correo o conversación y enrutarlo a la cola correcta en tiempo real, algo valioso en mesas de ayuda y centros de contacto. Fuente: documentación de Claude Haiku 5.5.

En TEKFENIX aplicamos esta idea en Servigo365, con clasificación y atención multicanal asistida por IA, y en Nexturno, para orientar turnos y colas en sucursales según la demanda.

On October 7, 2026, Anthropic released Claude Haiku 5.5, its fastest model, aimed at high-volume, low-latency work such as classification, routing, data extraction and subagents. According to the official documentation, it offers a 1-million-token context window and up to 128K output tokens.

Key facts

  • Pricing: $0.10 per million input tokens and $0.50 output for prompts up to 100K tokens; the batch API gives a 50% discount.
  • Text and image input, text output; knowledge cutoff June 2026.
  • Adaptive thinking on by default, with an effort parameter to tune depth.
  • Available on the Claude API, Google Cloud, Microsoft Foundry, Amazon Bedrock and Claude Platform on AWS.

It is the cheapest option in the 5.5 family, below Sonnet 5.5 ($2/$10), Opus 5.5 ($4/$20) and Fable 5.1 ($10/$50). The documentation warns the new tokenizer uses roughly 30% more tokens than Haiku 4.5 for the same text, and that temperature, top_p and top_k must not be sent.

Why it matters to businesses

Small, cheap models make it feasible to classify every ticket, email or conversation and route it to the right queue in real time, which is valuable for help desks and contact centers. Source: Claude Haiku 5.5 documentation.

At TEKFENIX we apply this idea in Servigo365, with AI-assisted classification and multichannel support, and in Nexturno, to guide queues and appointments in branches according to demand.

← Volver al blog← Back to blog