For the complete documentation index, see llms.txt. This page is also available as Markdown.

LLM Updates

This section provides the latest updates to the LLMs available in Blockbrain. Users can learn about newly added models, model improvements, provider availability, retirements, and replacements.

Newly Added Models

1. GPT 5.6 Sol (Fast)

OpenAI (EU)

GPT 5.6 Sol (Fast) is now available with EU hosting through OpenAI. It delivers the flagship intelligence of GPT 5.6 Sol at up to 2.5× faster speeds, making it a strong choice when both maximum capability and low latency are essential.

With a 1M token context window, It is especially well suited for complex coding, advanced reasoning, long-running agents, research, and other demanding workflows where teams need top-tier results without waiting for standard flagship response times. Fast Mode retains the same intelligence as GPT 5.6 Sol at approximately twice the price.

Model type: LLM Context window: 1M tokens Recommended upgrade from: GPT 5.6 Sol when faster responses are required

2. GPT 5.6 Terra (Fast)

OpenAI (EU)

GPT 5.6 Terra (Fast) is now available with EU hosting through OpenAI. It provides the same balanced intelligence as GPT 5.6 Terra with accelerated response times, offering near-flagship quality for everyday workflows where latency matters.

With a 1M token context window, it is well suited for time-sensitive coding, document analysis, tool use, knowledge work, and multi-step automation. It is a practical middle ground for teams that need stronger capability than lightweight models but do not require the full depth or cost of the Sol tier. Fast Mode is priced at approximately twice the standard version.

Model type: LLM Context window: 1M tokens Recommended upgrade from: GPT 5.6 Terra when faster responses are required

3. GPT 5.6 Luna (Fast)

OpenAI (EU)

GPT 5.6 Luna (Fast) is now available with EU hosting through OpenAI. Built for maximum throughput, it delivers the same intelligence as GPT 5.6 Luna with accelerated response speeds, making it ideal for workloads where fast and consistent output is the priority.

With a 1M token context window, it is especially well suited for high-volume automations, data extraction, classification, customer-facing assistants, document processing, and other time-critical production tasks. It retains Luna’s strong cost efficiency while providing near-instant responses, with Fast Mode priced at approximately twice the standard version.

Model type: LLM Context window: 1M tokens Recommended upgrade from: GPT 5.6 Luna when faster responses are required

4. Gemini 3.5 Flash (Lite)

OpenAI (EU)

Gemini 3.5 Flash-Lite is now available with EU hosting through Google AI. It is the fastest and most cost-effective model in the Gemini 3.5 family, designed for speed, scale, and high-volume production workloads.

With a 1M token context window and output speeds of up to 350 tokens per second, it is a major upgrade over Gemini 3.1 Flash-Lite. It is particularly well suited for agentic search, document processing, data extraction, classification, and other high-throughput workflows where minimizing latency and cost is more important than using a larger general-purpose model.

Model type: LLM Context window: 1M tokens Recommended upgrade from: Gemini 3.1 Flash (Lite)

Newly Added Models

1. Qwen 3VL 235B

StackIT (EU)

Qwen 3VL 235B is now available with EU hosting through StackIt. It is a powerful open-weight vision-language model designed for advanced multimodal work, including long-document processing, extended video understanding, complex visual analysis, and agentic workflows.

With a 218K token context window, it is especially well suited for teams that need high-end multimodal capabilities and have the infrastructure to support a heavyweight model.

Model type: LLM Context window: 218K tokens

2. Qwen3.6 27B

StackIT (EU)

Qwen3.6 27B is now available with EU hosting through StackIt. It is an open-weight multimodal model that delivers particularly strong agentic coding capabilities for its size. The model is well suited for software development, frontend workflows, tool use, visual analysis, and other tasks that combine coding with multimodal understanding.

With a 262K token context window, it offers a strong balance of capability, deployment flexibility, and long-context support.

Model type: LLM Context window: 262K tokens

3. Llama 3.3 70B

StackIT (EU)

Llama 3.3 70B is now available with EU hosting through StackIt. It is a balanced, high-performance open-weight model that provides robust reasoning and consistently reliable output quality across a broad range of tasks.

With a 128K token context window, it is especially well suited for complex analysis, long-form content creation, contextual decision support, and general-purpose enterprise workflows that require dependable performance.

Model type: LLM Context window: 128K tokens

4. GPT OSS 120B

StackIT (EU)

GPT OSS 120B is now available with EU hosting through StackIt. It is OpenAI’s largest open-weight reasoning model, built for advanced coding, reasoning, tool calling, and agentic workflows.

With a 131K token context window, it is especially well suited for enterprise agents, complex development tasks, and specialized deployments where organizations require greater control over their data and model infrastructure.

Model type: LLM Context window: 131K tokens

5. GPT OSS 20B

StackIT (EU)

GPT OSS 20B is now available with EU hosting through StackIt. It is a compact open-weight reasoning model designed for low-latency, private, and specialized deployments. The model supports coding, mathematics, tool use, and agentic workflows while requiring as little as 16 GB of memory.

With a 131K token context window, it is especially well suited for private agents, local inference, and cost-conscious deployments that need capable reasoning without heavyweight infrastructure.

Model type: LLM Context window: 131K tokens

6. Gemma 3 27B

StackIT (EU)

Gemma 3 27B is now available with EU hosting through StackIt. It is a capable open-weight multimodal model supporting text and image inputs, more than 140 languages, function calling, and structured outputs.

With a 131K token context window, it is especially well suited for document analysis, multilingual assistants, visual tasks, and flexible agentic workflows that require reliable multimodal understanding in a comparatively compact model.

Model type: LLM Context window: 131K tokens

Newly Added Models

1. Claude Opus 5

AWS Bedrock (EU)

Claude Opus 5 is now available with EU hosting through AWS Bedrock. It is a major step up from Opus 4.8 at the same price, delivering near–Fable 5 intelligence at approximately half the cost. The model is especially strong at complex professional work and long-running agent workflows where it must review, verify, and improve its own output.

With a 1M token context window, Opus 5 is the recommended upgrade for AWS Bedrock users who need greater intelligence and reliability without an increase in model pricing.

Model type: LLM Context window: 1 million tokens Recommended upgrade from: Claude Opus 4.8

2. Claude Opus 5

Google AI (EU)

Claude Opus 5 is now available with EU hosting through Google AI. It is a major step up from Opus 4.8 at the same price, delivering near–Fable 5 intelligence at approximately half the cost. The model is especially strong at complex professional work and long-running agent workflows where it must review, verify, and improve its own output.

With a 1M token context window, Opus 5 is the recommended upgrade for AWS Bedrock users who need greater intelligence and reliability without an increase in model pricing.

Model type: LLM Context window: 1 million tokens Recommended upgrade from: Claude Opus 4.8

Newly Added Models

1. Grok 4.5

xAI (US)

Grok 4.5 is now available with US hosting through xAI. xAI’s latest flagship model combines strong reasoning with fast response speeds and competitive pricing, with particular strengths in real-world software engineering, coding agents, and agentic task execution.

With a 500k token context window, it is especially well suited for technical workflows, application development, business productivity, and complex knowledge work that requires high capability without the latency and cost of slower flagship models.

Model type: LLM Context window: 500,000 tokens Recommended upgrade from: Grok 4

Newly Added Models

1. Claude Sonnet 5

Google AI (EU)

Claude Sonnet 5 is now available with EU hosting through Vertex AI. It is a major upgrade over Sonnet 4.6, delivering near-flagship performance across reasoning, coding, tool use, and knowledge work at a lower cost than top-tier models.

With a 1M token context window, Sonnet 5 is especially well suited for complex, multi-step tasks, large document and codebase analysis, and automation workflows that require high capability without premium pricing.

Model type: LLM Context window: 1 million tokens Recommended upgrade from: Claude Sonnet 4.6

Newly Added Models

1. Gemini 3.5 Flash

Google AI (EU)

Gemini 3.5 Flash is now available with EU hosting through Google AI. It is a major step up from Gemini 3.1 Pro, combining strong agentic, coding, financial, and multimodal capabilities with the speed expected from a Flash model.

Its 1M token context window makes it particularly effective for fast multi-step task handling, coding workflows, document analysis, and automation.

The model offers Minimal, Low, Medium, and High thinking modes, allowing users to balance response speed with deeper reasoning depending on the task.

Model type: LLM Context window: 1 million tokens Recommended upgrade from: Gemini 2.5 Flash

2. Mistral Medium 3.5

Mistral (EU)

Mistral Medium 3.5 is now available with EU hosting through Mistral. This open-weight model combines Mistral’s previous reasoning and coding lines into a unified model, providing strong performance for agentic coding, tool use, and multi-step technical work.

With a 256k token context window, it is particularly well suited for coding agents, software development workflows, technical automation, and applications that need a capable general-purpose model with open-weight flexibility.

Model type: LLM Context window: 256,000 tokens Recommended upgrade from: Mistral Medium

3. Claude Opus 4.8

AWS Bedrock (EU)

Claude Opus 4.8 is now available with EU hosting through AWS Bedrock. It is a significant upgrade over Opus 4.7, with stronger capabilities across agentic coding, computer use, and complex long-running tasks. Improved honesty and reliability also make it a more dependable partner for workflows that require sustained autonomy and accurate reporting.

With a 1M token context window, it is especially well suited for difficult engineering work, computer-use agents, codebase management, and high-stakes knowledge tasks.

Model type: LLM Context window: 1 million tokens Recommended upgrade from: Claude Opus 4.7reyouki

4. Claude Opus 4.7

AWS Bedrock (EU)

Claude Opus 4.7 is now available with EU hosting through AWS Bedrock. It is a major upgrade over Opus 4.6, designed for difficult hands-on coding and engineering work that can be delegated with greater confidence. Beyond coding, it provides stronger performance across reasoning, tool use, and complex knowledge tasks.

With a 1M token context window, it is particularly well suited for advanced software development, codebase analysis, long-running agents, and demanding multi-step workflows. For AWS Bedrock users of Opus 4.6, Opus 4.7 is the recommended upgrade.

Model type: LLM Context window: 1 million tokens Recommended upgrade from: Claude Opus 4.6

5. Claude Opus 4.6

AWS Bedrock (EU)

Claude Opus 4.6 is now available with EU hosting through AWS Bedrock. It improves on Opus 4.5 across coding, agentic task execution, codebase management, structured reviews, and debugging. It also provides more consistent and efficient performance during demanding analysis, document review, and multitasking workflows.

With a 1M token context window, it is well suited for complex engineering, long-running agents, large-scale document work, and tasks requiring reliable reasoning across multiple steps.

The model offers Low, Medium, High, and Max Reasoning modes, ranging from faster and more efficient responses to maximum reasoning depth for the most complex tasks.

Model type: LLM Context window: 1 million tokens Thinking modes: Low Reasoning, Medium Reasoning, High Reasoning, Max Reasoning Recommended upgrade from: Claude Opus 4.5

6. Claude Sonnet 4.6 (Fast)

AWS Bedrock (EU)

Claude Sonnet 4.6 (Fast) is now available with EU hosting through AWS Bedrock. This configuration operates strictly in Non-Reasoning mode, bypassing extended thinking steps to deliver lower-latency responses while retaining the underlying quality of Sonnet 4.6.

With a 1M context window, it is best suited for responsive chat experiences, high-volume production workloads, straightforward coding assistance, document processing, and agent workflows where speed is more important than deeper reasoning.

Model type: LLM Context window: 1 million tokens

7. Claude Haiku 4.5

AWS Bedrock (EU)

Claude Haiku 4.5 is now available with EU hosting through AWS Bedrock. It is a fast and cost-efficient model that delivers strong coding, reasoning, tool use, and interface interaction capabilities at production scale. Its speed and ability to execute tasks in parallel make it particularly effective as a worker model within multi-agent systems.

With a 200k token context window and an available Fast configuration, it is well suited for backend automation, chat workloads, coding assistance, and high-volume agent systems that prioritize responsiveness and lower costs.

Model type: LLM Context window: 200,000 tokens Recommended upgrade from: Claude Haiku 3.5

Newly Added Models

1. Claude Sonnet 5

AWS Bedrock (EU)

Claude Sonnet 5 is now available with EU hosting through AWS Bedrock. It is a major upgrade over Sonnet 4.6, delivering near-flagship performance across reasoning, coding, tool use, and knowledge work at a lower cost than top-tier models.

With a 1M token context window, Sonnet 5 is especially well suited for complex, multi-step tasks, large document and codebase analysis, and automation workflows that require high capability without premium pricing.

Model type: LLM Context window: 1 million tokens Recommended upgrade from: Claude Sonnet 4.6

2. Gemini 3.1 Flash-Lite Image

Google AI (US)

Gemini 3.1 Flash-Lite Image is now available with US hosting through Google AI. This lightweight image-generation model provides a fast and cost-efficient option for creating visuals directly within the platform. It is particularly well suited for high-volume image generation, rapid creative experimentation, simple marketing assets, and workflows where turnaround time and affordability are more important than the advanced controls of larger image models.

Model type: Image model Provider: Google AI Hosting: US

3. GPT 5.6 Sol

Azure AI (EU)

GPT-5.6 Sol is now available with EU hosting through Azure AI. It is the most capable model in the GPT-5.6 family, designed to deliver the highest-quality results for complex and demanding work.

With a 1M token context window, it is especially well suited for difficult reasoning, long-running tasks, advanced research, complex coding, document-intensive analysis, and workflows where accuracy and depth are more important than speed or cost.

Model type: LLM Context window: 1 million tokens Recommended upgrade from: GPT 5.5

4. GPT 5.6 Terra

Azure AI (EU)

GPT-5.6 Sol is now available with EU hosting through Azure AI. It is the most capable model in the GPT-5.6 family, designed to deliver the highest-quality results for complex and demanding work.

With a 1M token context window, it is especially well suited for difficult reasoning, long-running tasks, advanced research, complex coding, document-intensive analysis, and workflows where accuracy and depth are more important than speed or cost.

Model type: LLM Context window: 1 million tokens Recommended upgrade from: GPT 5.4

5. GPT 5.6 Luna

Azure AI (EU)

GPT-5.6 Sol is now available with EU hosting through Azure AI. It is the most capable model in the GPT-5.6 family, designed to deliver the highest-quality results for complex and demanding work.

With a 1M token context window, it is especially well suited for difficult reasoning, long-running tasks, advanced research, complex coding, document-intensive analysis, and workflows where accuracy and depth are more important than speed or cost.

Model type: LLM Context window: 1 million tokens Recommended upgrade from: GPT 5.4, GPT 5.4 Nano, GPT 5.4 Mini

Newly Added Models

1. GPT 5.6 Sol

OpenAI (EU)

GPT-5.6 Sol is now available with EU hosting through OpenAI. It is the most capable model in the GPT-5.6 family, designed to deliver the highest-quality results for complex and demanding work.

With a 1M token context window, it is especially well suited for difficult reasoning, long-running tasks, advanced research, complex coding, document-intensive analysis, and workflows where accuracy and depth are more important than speed or cost.

Model type: LLM Context window: 1 million tokens Recommended upgrade from: GPT 5.5

2. GPT 5.6 Terra

OpenAI (EU)

GPT-5.6 Sol is now available with EU hosting through OpenAI. It is the most capable model in the GPT-5.6 family, designed to deliver the highest-quality results for complex and demanding work.

With a 1M token context window, it is especially well suited for difficult reasoning, long-running tasks, advanced research, complex coding, document-intensive analysis, and workflows where accuracy and depth are more important than speed or cost.

Model type: LLM Context window: 1 million tokens Recommended upgrade from: GPT 5.4

3. GPT 5.6 Luna

OpenAI (EU)

GPT-5.6 Sol is now available with EU hosting through OpenAI. It is the most capable model in the GPT-5.6 family, designed to deliver the highest-quality results for complex and demanding work.

With a 1M token context window, it is especially well suited for difficult reasoning, long-running tasks, advanced research, complex coding, document-intensive analysis, and workflows where accuracy and depth are more important than speed or cost.

Model type: LLM Context window: 1 million tokens Recommended upgrade from: GPT 5.4, GPT 5.4 Nano, GPT 5.4 Mini

Last updated