NewThe new NeoFlux certified email is online: simple, secure and fully legally valid.
Cloud & Dedicated Servers
Call us

The best on-demand open-source AI in a sovereign Cloud

Discover the best open-source alternatives to ChatGPT, Gemini, Midjourney or Claude to process sensitive data in full compliance with European law.

Large language models (LLM)

The best open-source alternatives to ChatGPT, Gemini and Microsoft Copilot to interact with, analyse and generate content with AI.

Qwen/Qwen3.5-122B-A10B-FP8

The most powerful

Beta

  • Designed for complex tasks that require a large context window and greater precision in logical reasoning.

  • Architecture optimised for faster inference and reduced energy consumption, freeing up significant computing resources.

  • Trained on millions of agents and increasingly complex tasks to ensure robust real-world adaptability.

Modality

Image-Text to Text

Max input tokens

200 000

Languages

More than 100 languages

Function calling

Yes

Model category

chat_large

  • Designed for complex tasks that require a large context window and greater precision in logical reasoning.

  • Architecture optimised for faster inference and reduced energy consumption, freeing up significant computing resources.

  • Trained on millions of agents and increasingly complex tasks to ensure robust real-world adaptability.

Modality

Image-Text to Text

Max input tokens

200 000

Languages

More than 100 languages

Function calling

Yes

Model category

chat_large

Apertus-70B-Instruct-2509

The most ethical

Beta

  • Ideal for multilingual services, public administrations and R&D teams looking for a reliable and adaptable model

  • Documented data and methods for unprecedented transparency

  • Compliant with the AI Act and respectful of privacy and intellectual property

  • A 70B version with performance comparable to today's market leaders

Modality

Text to Text

Max input tokens

65 536

Languages

More than 100 languages

Function calling

No

Model category

chat_medium

  • Ideal for multilingual services, public administrations and R&D teams looking for a reliable and adaptable model

  • Documented data and methods for unprecedented transparency

  • Compliant with the AI Act and respectful of privacy and intellectual property

  • A 70B version with performance comparable to today's market leaders

Modality

Text to Text

Max input tokens

65 536

Languages

More than 100 languages

Function calling

No

Model category

chat_medium

google/gemma-4-31B-it

The perfect balance

Beta

  • The ideal trade-off between responsiveness and power, designed to excel at logical reasoning, in-depth document analysis and reliable code generation.

  • Leverages a state-of-the-art architecture to deliver a deep understanding of advanced contexts and complex instructions.

  • Ideal for advanced conversational agents and business workflows that require great versatility without sacrificing execution speed.

Modality

Text-to-Text (optimised for instruction)

Max input tokens

100 000

Languages

More than 140 languages

Function calling

Yes (native and optimised)

Model category

chat_medium

  • The ideal trade-off between responsiveness and power, designed to excel at logical reasoning, in-depth document analysis and reliable code generation.

  • Leverages a state-of-the-art architecture to deliver a deep understanding of advanced contexts and complex instructions.

  • Ideal for advanced conversational agents and business workflows that require great versatility without sacrificing execution speed.

Modality

Text-to-Text (optimised for instruction)

Max input tokens

100 000

Languages

More than 140 languages

Function calling

Yes (native and optimised)

Model category

chat_medium

moonshotai/Kimi-K2.6

The most powerful for vibe coding

Beta

  • Natively multimodal: turns text, images or mockups into fully functional code.

  • Built for large-scale development: features an extended context window of up to 256k tokens to handle complex projects

  • Optimised for vibe coding: a fast, smooth and creative experience designed for developers and product designers

  • Compatible with agentic workflows: automates analysis, code generation and its end-to-end execution

Modality

Image-Text to Text

Max input tokens

256 000

Languages

Multilingual

Function calling

Yes

Model category

code

  • Natively multimodal: turns text, images or mockups into fully functional code.

  • Built for large-scale development: features an extended context window of up to 256k tokens to handle complex projects

  • Optimised for vibe coding: a fast, smooth and creative experience designed for developers and product designers

  • Compatible with agentic workflows: automates analysis, code generation and its end-to-end execution

Modality

Image-Text to Text

Max input tokens

256 000

Languages

Multilingual

Function calling

Yes

Model category

code

mistralai/Ministral-3-14B-Instruct-2512

The most versatile

Beta

  • Optimised for fast and cost-effective deployment, ideal for conversational agents, document analysis and specialised tasks.

  • Delivers performance comparable to Mistral Small 3.2 24B with minimal resources.

  • Able to analyse images and provide information based on visual content as well as text.

Modality

Image-Text to Text

Max input tokens

100 000

Languages

EN, ES, FR, DE, IT...

Function calling

Yes

Model category

chat_small

  • Optimised for fast and cost-effective deployment, ideal for conversational agents, document analysis and specialised tasks.

  • Delivers performance comparable to Mistral Small 3.2 24B with minimal resources.

  • Able to analyse images and provide information based on visual content as well as text.

Modality

Image-Text to Text

Max input tokens

100 000

Languages

EN, ES, FR, DE, IT...

Function calling

Yes

Model category

chat_small

Reranking models

The best compatible open-source alternatives to optimise the relevance of your search results. Optimise your document ranking, improve the accuracy of your RAG systems and ensure smarter, more contextual information retrieval.

BAAI/bge-reranker-v2-m3

The most versatile

  • Advanced multilingual model able to process short queries, paragraphs and long documents of up to 8192 tokens simultaneously

  • Combines lexical analysis (keywords) and semantic analysis (meaning) for unrivalled ranking accuracy on complex corpora

  • The ideal solution for enterprise search engines and RAG applications that require a deep understanding of context

Modality

Text to Text

Max input tokens

8192

Languages

Over 100 languages

Function calling

No

Type

Position

  • Advanced multilingual model able to process short queries, paragraphs and long documents of up to 8192 tokens simultaneously

  • Combines lexical analysis (keywords) and semantic analysis (meaning) for unrivalled ranking accuracy on complex corpora

  • The ideal solution for enterprise search engines and RAG applications that require a deep understanding of context

Modality

Text to Text

Max input tokens

8192

Languages

Over 100 languages

Function calling

No

Type

Position

Qwen/Qwen3-Reranker-0.6B

The most efficient

  • Ultra-lightweight architecture (0.6 billion parameters) designed for very low-latency inference and minimal energy consumption

  • Maintains high relevance accuracy even with a context window extended up to 32768 tokens

  • Ideal for real-time data streams, autonomous agents and large-scale deployments

Modality

Text to Text

Max input tokens

32768

Languages

Over 100 languages

Function calling

No

Type

Position

  • Ultra-lightweight architecture (0.6 billion parameters) designed for very low-latency inference and minimal energy consumption

  • Maintains high relevance accuracy even with a context window extended up to 32768 tokens

  • Ideal for real-time data streams, autonomous agents and large-scale deployments

Modality

Text to Text

Max input tokens

32768

Languages

Over 100 languages

Function calling

No

Type

Position

Embedding models

The best open-source embedding models to turn your data into intelligent vectors. Improve the accuracy of your searches, personalise your recommendations, simplify data analysis, explore semantic connections and classify text easily.

Bge Multilingual Gemma2

The highest quality

  • The most powerful open-source embedding model on the market

  • The benchmark for semantic search and retrieval-augmented generation (RAG) tasks

  • Ideal for advanced use of embedding vectors across various use cases

  • Exceptional performance, regardless of the text language (100+ languages)

Max input tokens

8192

Parameters

9.2 B

Dimensions

3584

Languages

EN, ES, FR, DE, IT...

Type

Text

  • The most powerful open-source embedding model on the market

  • The benchmark for semantic search and retrieval-augmented generation (RAG) tasks

  • Ideal for advanced use of embedding vectors across various use cases

  • Exceptional performance, regardless of the text language (100+ languages)

Max input tokens

8192

Parameters

9.2 B

Dimensions

3584

Languages

EN, ES, FR, DE, IT...

Type

Text

All MiniLM L12 v2

The best value for money

  • This model is the result of joint work based on a model published by Microsoft

  • Excellent value for money, ideal for prototyping and simple tasks with limited resources

  • Attractive performance for relatively simple tasks, regardless of the text language

  • Extreme speed for indexing huge databases or for real-time processing

  • High energy efficiency to reduce environmental impact

Max input tokens

512

Parameters

33 M

Dimensions

384

Languages

EN, ES, FR, DE, IT...

Type

Text

  • This model is the result of joint work based on a model published by Microsoft

  • Excellent value for money, ideal for prototyping and simple tasks with limited resources

  • Attractive performance for relatively simple tasks, regardless of the text language

  • Extreme speed for indexing huge databases or for real-time processing

  • High energy efficiency to reduce environmental impact

Max input tokens

512

Parameters

33 M

Dimensions

384

Languages

EN, ES, FR, DE, IT...

Type

Text

Speech recognition

The best open-source AI to transcribe audio files into text or create realistic human voices.

Whisper V3

For complex transcriptions

  • Model trained on over 1 million hours of data

  • Reduction of transcription errors by up to 20% compared to Whisper V2

  • Better handling of accents, background noise and complex speech (for example, calls or video conferences)

  • Improved multilingual support and translation of transcriptions into languages other than English

Maximum file size

25 MB

Supported formats

mp3, mp4, aac, wav, flac, ogg, opus, wma, m4a

  • Model trained on over 1 million hours of data

  • Reduction of transcription errors by up to 20% compared to Whisper V2

  • Better handling of accents, background noise and complex speech (for example, calls or video conferences)

  • Improved multilingual support and translation of transcriptions into languages other than English

Maximum file size

25 MB

Supported formats

mp3, mp4, aac, wav, flac, ogg, opus, wma, m4a

Image creation and processing

The best open-source alternatives to Midjourney, Microsoft Copilot Designer or Gemini to create, merge or interpret images.

Photomaker V2

Ideal for creating images

  • The best combination of quality and speed in image creation using generative AI

  • Fast creation of photorealistic images in 1, 2, 4 or 8 steps from a prompt

  • Works by distillation, which increases energy efficiency while ensuring excellent quality

  • Optimised for English, with limited knowledge of other languages (FR, DE, ES, IT…)

Max input tokens

77

Max output images

5

Languages

EN

Maximum resolution

1024x1024, 1792x1024, 1024x1792

  • The best combination of quality and speed in image creation using generative AI

  • Fast creation of photorealistic images in 1, 2, 4 or 8 steps from a prompt

  • Works by distillation, which increases energy efficiency while ensuring excellent quality

  • Optimised for English, with limited knowledge of other languages (FR, DE, ES, IT…)

Max input tokens

77

Max output images

5

Languages

EN

Maximum resolution

1024x1024, 1792x1024, 1024x1792

Flux schnell

Ideal for editing and merging portraits of people

  • Creation of photos in multiple styles from one or more profile photos

  • Powerful and flexible: recontextualisation, colourisation, age and gender change, identity mixing...

Max input tokens

77

Max input images

6

Max output images

5

Languages

EN

Maximum resolution

1024x1024, 1792x1024, 1024x1792

  • Creation of photos in multiple styles from one or more profile photos

  • Powerful and flexible: recontextualisation, colourisation, age and gender change, identity mixing...

Max input tokens

77

Max input images

6

Max output images

5

Languages

EN

Maximum resolution

1024x1024, 1792x1024, 1024x1792