Information for administrators
Good to know

Useful information about AI models

Provider and model diversity in nele.aiVariety of Providers and Models in nele.ai

nele.ai is a platform that provides a wide range of generative language models from various providers. Current partners include renowned companies such as OpenAI, Microsoft Azure, Google, Anthropic, as well as AWS Bedrock and Mistral. In addition, new models are continuously integrated to ensure that the latest developments in artificial intelligence are always available.nele.ai is a platform that provides a wide range of generative language models from various providers. Currently available partners include renowned companies such as OpenAI, Microsoft Azure, Google, Anthropic, AWS Bedrock, and Mistral. Furthermore, new models are continuously integrated to always offer the latest developments in artificial intelligence.

‍‍

The currently available models include, among others:The currently available models include the following:

‍‍

  • Advanced models of the GPT-5 family such as GPT-5.5, GPT-5.4, and their variants
  • The proven GPT-4.1 generation with GPT-4.1, GPT-4.1 mini, and GPT-4.1 nano
  • Google Gemini 3.5 Flash as well as Gemini 2.5 Pro and Flash
  • Claude 4.8 Opus, Claude 4.6 Sonnet, and Claude 4.5 Haiku from Anthropic
  • Mistral Large 3, Mistral Medium 3.5, and Mistral Small 4 for specialized applications
    ‍

Through these partnerships, nele.ai always provides the latest AI models from these providers while ensuring compliance with high security and data protection standards.Through these partnerships, nele.ai consistently provides the latest AI models from these providers, ensuring compliance with high security and data protection standards.

European server locations as an alternative to the USAEuropean Server Locations: An Alternative to the US

It is particularly important that high-performance AI models are also offered on European servers. This allows organizations that must keep their data within Europe to use these models without restriction. For this reason, nele.ai offers almost all AI models on European servers as well.It is particularly important that powerful AI models are also available on European servers. This enables organizations that must keep their data within Europe to utilize these models without restriction. For this reason, nele.ai provides nearly all AI models on European servers.

‍‍

Available models on European servers include:Available models on European servers include:

‍‍

  • GPT-5.5, GPT-5.4, GPT-5.1, and the entire GPT-4.1 series via Azure infrastructure
  • Google Gemini 3.5 Flash as well as Gemini 2.5 Pro and Flash
  • Claude 4.6 Sonnet and Claude 4.5 Haiku via Google and AWS Bedrock
  • Mistral Large 3, Mistral Medium 3.5, and Mistral Small 4

Advanced AI model functionalitiesEnhanced AI Model Capabilities

The available AI models offer various specialized functions:The available AI models offer various specialized functionalities:

‍‍

Reasoning capabilitiesReasoning Capabilities

Modern models such as GPT-5.5, GPT-5.4, GPT-5.1, Gemini 3.5 Flash, Claude 4.8 Opus, Claude 4.6 Sonnet, and their variants possess advanced reasoning capabilities that enable complex, logical deductions and multi-step problem solving.Modern models such as GPT-5.5, GPT-5.4, GPT-5.1, Gemini 3.5 Flash, Claude 4.8 Opus, Claude 4.6 Sonnet and their variants possess advanced reasoning capabilities, which enable complex, logical conclusions and multi-step problem-solving.

‍‍

In nele.ai, you can configure the so-called thinking modeThinking Mode for supported models. This determines how intensively the model processes a request before responding. The following levels are available:In nele.ai, for supported models, the so-called can be configured. This determines how thoroughly the model processes a request before responding. The following levels are available:

‍‍

  • Standard configuration – configuration predefined by the provider
  • Brief thinking – a few simple steps for a quick response
  • Moderate thinking – several steps, better reasoned, somewhat more detailed
  • Thorough thinking – many steps, checks assumptions, very detailed
  • Very thorough thinking – many steps, checks assumptions, very detailed

Please note that not all models with reasoning capabilities support the full range of available thinking mode levels.It should be noted that not all models with reasoning functionality support the full spectrum of available thinking mode levels.

‍‍

Vision functionalityVision Functionality

Most current models support vision capabilities for analyzing and processing image content, including all models in the GPT-4.1 and GPT-5 families, Gemini 2.5 and Gemini 3.5 Flash, as well as all Claude models.Most current models support vision capabilities for analyzing and processing image content, including all models from the GPT-4.1 and GPT-5 families, Gemini 2.5 and Gemini 3.5 Flash, as well as all Claude models.

‍‍

Multimodal applicationsMultimodal Applications

In addition to text models, specialized modelsspecialized models are available:In addition to text models, are available:

‍‍

  • Image generation: GPT Image 2 (OpenAI) as well as Gemini 2.5 Flash Image (Google)
  • Audio processing and transcription: Azure Speech - Fast transcription for speech recognition and transcription
  • Audio processing and transcription: Azure Whisper and Azure Speech - Fast Transcription for speech recognition and transcriptionAudio processing: Whisper for voice recognition and transcription

AI models and token sizes in chat contextAI Models and Token Sizes in Chat Context

The AI models offered vary in their token capacities within their chat context (see also our blog post on the difference between knowledge base (RAG) vs. chat contextblog post on the difference between Knowledge Base (RAG) vs. Chat Context), with extended context sizes ranging from 128k to over 1 million tokens. It is important to understand what a token is: the smallest units that make up AI models. These units can be letters, syllables, abbreviations, or entire words. Tokens are comparable to puzzle pieces that are assembled to form answers.The AI models offered vary in their token sizes within their chat context (see also our ), with extended context sizes ranging from 128k to over 1 million tokens. It's important to understand what a token is: the smallest units that make up AI models. These units can be letters, syllables, abbreviations, or entire words. Tokens are comparable to puzzle pieces that are assembled to form responses.

‍‍

Context sizes vary depending on the model:Context sizes vary by model:

‍‍

  • Medium contexts: 128k–256k tokens (Mistral Medium 3.5 as well as Mistral Large 3 and Small 4)
  • Extended contexts: 200k–256k tokens (Claude 4.6 Sonnet, Claude 4.5 Haiku)
  • Large contexts: 400k tokens (GPT-5, GPT-5-mini, GPT-5-nano, and GPT-5.1)
  • Maximum contexts: over 1 million tokens (GPT-5.5, GPT-5.4, the GPT-4.1 series, as well as Gemini 2.5 and Gemini 3.5 Flash)

‍‍

The chat context limits the number of tokens that can be processed in a chat. As a rule of thumb, 750 English words correspond to approximately 1,000 tokens, while in German, 1,000 tokens correspond to about 350 words.The chat context limits the number of tokens that can be processed in a chat. As a rule of thumb, 750 English words correspond to approximately 1,000 tokens, while in German, 1,000 tokens correspond to about 350 words.

Billing via our flexible and transparent pricing modelBilling through our flexible and transparent pricing model

The costs for AI models at nele.ai vary depending on the model used and the number of tokens or words. For language models, billing is based on tokens, while for image models, the price depends on factors such as the desired image resolution. nele.ai has introduced a flexible and transparent pricing model based on credits.The costs for AI models at nele.ai vary depending on the model used and the number of tokens or words. For language models, billing is per token, while for image models, the price depends, for example, on the desired image resolution. nele.ai has introduced a flexible and transparent credit-based pricing model.

‍‍

The key advantage of nele.ai is that it offers usage-based billing instead of fixed monthly fees per member. This allows every employee in an organization to have access without incurring flat-rate costs. Knowing that the need for generative AI varies among employees, this model ensures fair and reasonable costs.A particular advantage of nele.ai is its usage-based billing instead of fixed monthly fees per member. This allows every employee in an organization to gain access without incurring flat-rate costs. Recognizing that the demand for generative AI varies among employees, this model ensures fair and appropriate costs.

‍‍

An important factor in this model is the AI volume consumption factor, which describes the ratio of costs to credit consumption. These factors vary depending on the model and its performance, with newer and more powerful models typically having higher factors.A key factor in this model is the AI volume consumption factor, which describes the ratio of costs to credit consumption. These factors vary depending on the model and its performance, with newer and more powerful models typically having higher factors.

‍‍

This structure enables effective cost optimization and better resource utilization.This structure enables effective cost optimization and improved resource utilization.

‍‍

In addition, the nele.ai administration interface (manage.nele.aimanage.nele.ai) simplifies the management of available AI models and their associated costs. Administrators can define which AI models are available to their team and restrict individual members regarding their AI usage volume.The administration interface () of nele.ai also simplifies the management of available AI models and their associated costs. Administrators can define which AI models are accessible to their team and restrict individual members' AI usage volume.