The best on-demand open-source AI in a sovereign Cloud
Discover the best open-source alternatives to ChatGPT, Gemini, Midjourney or Claude to process sensitive data in full compliance with European law.
Large language models (LLM)
The best open-source alternatives to ChatGPT, Gemini and Microsoft Copilot to interact with, analyse and generate content with AI.
Qwen/Qwen3.5-122B-A10B-FP8
The most powerful
Beta
- ●
Designed for complex tasks that require a large context window and greater precision in logical reasoning.
- ●
Architecture optimised for faster inference and reduced energy consumption, freeing up significant computing resources.
- ●
Trained on millions of agents and increasingly complex tasks to ensure robust real-world adaptability.
Modality
Image-Text to Text
Max input tokens
200 000
Languages
More than 100 languages
Function calling
Yes
Model category
chat_large
- ●
Designed for complex tasks that require a large context window and greater precision in logical reasoning.
- ●
Architecture optimised for faster inference and reduced energy consumption, freeing up significant computing resources.
- ●
Trained on millions of agents and increasingly complex tasks to ensure robust real-world adaptability.
Modality
Image-Text to Text
Max input tokens
200 000
Languages
More than 100 languages
Function calling
Yes
Model category
chat_large
Apertus-70B-Instruct-2509
The most ethical
Beta
- ●
Ideal for multilingual services, public administrations and R&D teams looking for a reliable and adaptable model
- ●
Documented data and methods for unprecedented transparency
- ●
Compliant with the AI Act and respectful of privacy and intellectual property
- ●
A 70B version with performance comparable to today's market leaders
Modality
Text to Text
Max input tokens
65 536
Languages
More than 100 languages
Function calling
No
Model category
chat_medium
- ●
Ideal for multilingual services, public administrations and R&D teams looking for a reliable and adaptable model
- ●
Documented data and methods for unprecedented transparency
- ●
Compliant with the AI Act and respectful of privacy and intellectual property
- ●
A 70B version with performance comparable to today's market leaders
Modality
Text to Text
Max input tokens
65 536
Languages
More than 100 languages
Function calling
No
Model category
chat_medium
google/gemma-4-31B-it
The perfect balance
Beta
- ●
The ideal trade-off between responsiveness and power, designed to excel at logical reasoning, in-depth document analysis and reliable code generation.
- ●
Leverages a state-of-the-art architecture to deliver a deep understanding of advanced contexts and complex instructions.
- ●
Ideal for advanced conversational agents and business workflows that require great versatility without sacrificing execution speed.
Modality
Text-to-Text (optimised for instruction)
Max input tokens
100 000
Languages
More than 140 languages
Function calling
Yes (native and optimised)
Model category
chat_medium
- ●
The ideal trade-off between responsiveness and power, designed to excel at logical reasoning, in-depth document analysis and reliable code generation.
- ●
Leverages a state-of-the-art architecture to deliver a deep understanding of advanced contexts and complex instructions.
- ●
Ideal for advanced conversational agents and business workflows that require great versatility without sacrificing execution speed.
Modality
Text-to-Text (optimised for instruction)
Max input tokens
100 000
Languages
More than 140 languages
Function calling
Yes (native and optimised)
Model category
chat_medium
moonshotai/Kimi-K2.6
The most powerful for vibe coding
Beta
- ●
Natively multimodal: turns text, images or mockups into fully functional code.
- ●
Built for large-scale development: features an extended context window of up to 256k tokens to handle complex projects
- ●
Optimised for vibe coding: a fast, smooth and creative experience designed for developers and product designers
- ●
Compatible with agentic workflows: automates analysis, code generation and its end-to-end execution
Modality
Image-Text to Text
Max input tokens
256 000
Languages
Multilingual
Function calling
Yes
Model category
code
- ●
Natively multimodal: turns text, images or mockups into fully functional code.
- ●
Built for large-scale development: features an extended context window of up to 256k tokens to handle complex projects
- ●
Optimised for vibe coding: a fast, smooth and creative experience designed for developers and product designers
- ●
Compatible with agentic workflows: automates analysis, code generation and its end-to-end execution
Modality
Image-Text to Text
Max input tokens
256 000
Languages
Multilingual
Function calling
Yes
Model category
code
mistralai/Ministral-3-14B-Instruct-2512
The most versatile
Beta
- ●
Optimised for fast and cost-effective deployment, ideal for conversational agents, document analysis and specialised tasks.
- ●
Delivers performance comparable to Mistral Small 3.2 24B with minimal resources.
- ●
Able to analyse images and provide information based on visual content as well as text.
Modality
Image-Text to Text
Max input tokens
100 000
Languages
EN, ES, FR, DE, IT...
Function calling
Yes
Model category
chat_small
- ●
Optimised for fast and cost-effective deployment, ideal for conversational agents, document analysis and specialised tasks.
- ●
Delivers performance comparable to Mistral Small 3.2 24B with minimal resources.
- ●
Able to analyse images and provide information based on visual content as well as text.
Modality
Image-Text to Text
Max input tokens
100 000
Languages
EN, ES, FR, DE, IT...
Function calling
Yes
Model category
chat_small
Reranking models
The best compatible open-source alternatives to optimise the relevance of your search results. Optimise your document ranking, improve the accuracy of your RAG systems and ensure smarter, more contextual information retrieval.
BAAI/bge-reranker-v2-m3
The most versatile
- ●
Advanced multilingual model able to process short queries, paragraphs and long documents of up to 8192 tokens simultaneously
- ●
Combines lexical analysis (keywords) and semantic analysis (meaning) for unrivalled ranking accuracy on complex corpora
- ●
The ideal solution for enterprise search engines and RAG applications that require a deep understanding of context
Modality
Text to Text
Max input tokens
8192
Languages
Over 100 languages
Function calling
No
Type
Position
- ●
Advanced multilingual model able to process short queries, paragraphs and long documents of up to 8192 tokens simultaneously
- ●
Combines lexical analysis (keywords) and semantic analysis (meaning) for unrivalled ranking accuracy on complex corpora
- ●
The ideal solution for enterprise search engines and RAG applications that require a deep understanding of context
Modality
Text to Text
Max input tokens
8192
Languages
Over 100 languages
Function calling
No
Type
Position
Qwen/Qwen3-Reranker-0.6B
The most efficient
- ●
Ultra-lightweight architecture (0.6 billion parameters) designed for very low-latency inference and minimal energy consumption
- ●
Maintains high relevance accuracy even with a context window extended up to 32768 tokens
- ●
Ideal for real-time data streams, autonomous agents and large-scale deployments
Modality
Text to Text
Max input tokens
32768
Languages
Over 100 languages
Function calling
No
Type
Position
- ●
Ultra-lightweight architecture (0.6 billion parameters) designed for very low-latency inference and minimal energy consumption
- ●
Maintains high relevance accuracy even with a context window extended up to 32768 tokens
- ●
Ideal for real-time data streams, autonomous agents and large-scale deployments
Modality
Text to Text
Max input tokens
32768
Languages
Over 100 languages
Function calling
No
Type
Position
Embedding models
The best open-source embedding models to turn your data into intelligent vectors. Improve the accuracy of your searches, personalise your recommendations, simplify data analysis, explore semantic connections and classify text easily.
Bge Multilingual Gemma2
The highest quality
- ●
The most powerful open-source embedding model on the market
- ●
The benchmark for semantic search and retrieval-augmented generation (RAG) tasks
- ●
Ideal for advanced use of embedding vectors across various use cases
- ●
Exceptional performance, regardless of the text language (100+ languages)
Max input tokens
8192
Parameters
9.2 B
Dimensions
3584
Languages
EN, ES, FR, DE, IT...
Type
Text
- ●
The most powerful open-source embedding model on the market
- ●
The benchmark for semantic search and retrieval-augmented generation (RAG) tasks
- ●
Ideal for advanced use of embedding vectors across various use cases
- ●
Exceptional performance, regardless of the text language (100+ languages)
Max input tokens
8192
Parameters
9.2 B
Dimensions
3584
Languages
EN, ES, FR, DE, IT...
Type
Text
All MiniLM L12 v2
The best value for money
- ●
This model is the result of joint work based on a model published by Microsoft
- ●
Excellent value for money, ideal for prototyping and simple tasks with limited resources
- ●
Attractive performance for relatively simple tasks, regardless of the text language
- ●
Extreme speed for indexing huge databases or for real-time processing
- ●
High energy efficiency to reduce environmental impact
Max input tokens
512
Parameters
33 M
Dimensions
384
Languages
EN, ES, FR, DE, IT...
Type
Text
- ●
This model is the result of joint work based on a model published by Microsoft
- ●
Excellent value for money, ideal for prototyping and simple tasks with limited resources
- ●
Attractive performance for relatively simple tasks, regardless of the text language
- ●
Extreme speed for indexing huge databases or for real-time processing
- ●
High energy efficiency to reduce environmental impact
Max input tokens
512
Parameters
33 M
Dimensions
384
Languages
EN, ES, FR, DE, IT...
Type
Text
Speech recognition
The best open-source AI to transcribe audio files into text or create realistic human voices.
Whisper V3
For complex transcriptions
- ●
Model trained on over 1 million hours of data
- ●
Reduction of transcription errors by up to 20% compared to Whisper V2
- ●
Better handling of accents, background noise and complex speech (for example, calls or video conferences)
- ●
Improved multilingual support and translation of transcriptions into languages other than English
Maximum file size
25 MB
Supported formats
mp3, mp4, aac, wav, flac, ogg, opus, wma, m4a
- ●
Model trained on over 1 million hours of data
- ●
Reduction of transcription errors by up to 20% compared to Whisper V2
- ●
Better handling of accents, background noise and complex speech (for example, calls or video conferences)
- ●
Improved multilingual support and translation of transcriptions into languages other than English
Maximum file size
25 MB
Supported formats
mp3, mp4, aac, wav, flac, ogg, opus, wma, m4a
Image creation and processing
The best open-source alternatives to Midjourney, Microsoft Copilot Designer or Gemini to create, merge or interpret images.
Photomaker V2
Ideal for creating images
- ●
The best combination of quality and speed in image creation using generative AI
- ●
Fast creation of photorealistic images in 1, 2, 4 or 8 steps from a prompt
- ●
Works by distillation, which increases energy efficiency while ensuring excellent quality
- ●
Optimised for English, with limited knowledge of other languages (FR, DE, ES, IT…)
Max input tokens
77
Max output images
5
Languages
EN
Maximum resolution
1024x1024, 1792x1024, 1024x1792
- ●
The best combination of quality and speed in image creation using generative AI
- ●
Fast creation of photorealistic images in 1, 2, 4 or 8 steps from a prompt
- ●
Works by distillation, which increases energy efficiency while ensuring excellent quality
- ●
Optimised for English, with limited knowledge of other languages (FR, DE, ES, IT…)
Max input tokens
77
Max output images
5
Languages
EN
Maximum resolution
1024x1024, 1792x1024, 1024x1792
Flux schnell
Ideal for editing and merging portraits of people
- ●
Creation of photos in multiple styles from one or more profile photos
- ●
Powerful and flexible: recontextualisation, colourisation, age and gender change, identity mixing...
Max input tokens
77
Max input images
6
Max output images
5
Languages
EN
Maximum resolution
1024x1024, 1792x1024, 1024x1792
- ●
Creation of photos in multiple styles from one or more profile photos
- ●
Powerful and flexible: recontextualisation, colourisation, age and gender change, identity mixing...
Max input tokens
77
Max input images
6
Max output images
5
Languages
EN
Maximum resolution
1024x1024, 1792x1024, 1024x1792
