State-of-the-art Danish Models
These models constitute state-of-the-art models for Danish within their respective domain (highlighted below the model).
24B • Updated • 541k • 1.38kNote Among the best performing open-weight ~10-100b generative models which has been instruction-tuned. Determined by EuroEval Danish NLG (2025/11/04).
google/gemma-3-27b-it
Image-Text-to-Text • 27B • Updated • 418k • • 2.04kNote Among the best performing open-weight ~10-100b generative models which has been instruction-tuned. Determined by EuroEval Danish NLG (2025/11/04).
google/gemma-3n-E4B-it
Image-Text-to-Text • 8B • Updated • 13.4k • • 931Note Among the best performing open-weight ~7-9b generative models which has been instruction-tuned. Determined by EuroEval Danish NLG (2025/11/04).
google/gemma-2-9b-it
Text Generation • 9B • Updated • 1.09M • • 1.09kNote Among the best performing open-weight ~7-9b generative models which has been instruction-tuned. Determined by EuroEval Danish NLG (2025/11/04).
google/gemma-2-9b
Text Generation • 9B • Updated • 29.4k • • 739Note Among the best performing open-weight ~7-9b generative models which hasn't been instruction-tuned. Determined by EuroEval Danish NLG (2025/11/04).
KennethEnevoldsen/dfm-sentence-encoder-large
Feature Extraction • 0.4B • Updated • 1.89k • 3Note Among the best large-sized encoder for Danish determined by EuroEval Danish NLU (2025/11/04)
AI-Sweden-Models/roberta-large-1160k
Fill-Mask • 0.4B • Updated • 1.04k • 11Note Among the best large-sized encoder for Danish determined by EuroEval Danish NLU (2025/11/04)
KennethEnevoldsen/dfm-sentence-encoder-medium
Sentence Similarity • Updated • 125Note Among the best medium-sized encoder for Danish determined by EuroEval Danish NLU (2025/11/04)
ltg/norbert3-small
Fill-Mask • Updated • 513 • 2Note Among the best small sized encoder for Danish as determined by EuroEval Danish NLU (2025/11/04)
syvai/hviske-v3-conversation
Automatic Speech Recognition • 2B • Updated • 2.04k • 12Note Automatic speech recognition based on Whisper 3 and fine-tuned on CoRal Obtains the lowest word error rate on CoRal conversations (2025/11/04), might be slightly overfit
openai/whisper-large-v3
Automatic Speech Recognition • 2B • Updated • 4.21M • • 6.54kNote Automatic speech recognition (ASR) Best multilingual ASR model for Danish (2025/11/04)
CoRal-project/roest-v2-wav2vec2-315m
Automatic Speech Recognition • 0.3B • Updated • 815 • 7Note Speech Encoder (Wav2Vec2.0) The encoder which obtains the lowest word error rate on CoRal (2025/11/04). Also exist in a 1B version.
jinaai/jina-embeddings-v3
Feature Extraction • 0.6B • Updated • 2.09M • 1.16kNote Among the best large-sized embedding model with flexible embedding sizes and long-document understanding. Determined by The Scandinavian Embedding Benchmark (SEB) (2025/11/04)
intfloat/multilingual-e5-large-instruct
Feature Extraction • 0.6B • Updated • 1.36M • • 643Note Among the best large-sized embedding model with Instructions. Determined by The Scandinavian Embedding Benchmark (SEB) (2025/11/04)
intfloat/multilingual-e5-large
Feature Extraction • 0.6B • Updated • 7.38M • • 1.26kNote Among the best large-sized embedding model which does not require instructions. Determined by The Scandinavian Embedding Benchmark (SEB) (2025/11/04)
intfloat/multilingual-e5-base
Sentence Similarity • 0.3B • Updated • 7.21M • • 391Note Among the best medium-sized embedding model which does not require instructions. Determined by The Scandinavian Embedding Benchmark (SEB) (2025/11/04)
intfloat/multilingual-e5-small
Sentence Similarity • 0.1B • Updated • 11.7M • • 435Note Among the best small-sized embedding model which does not require instructions. Determined by The Scandinavian Embedding Benchmark (SEB) (2025/11/04)
facebook/seamless-m4t-v2-large
Automatic Speech Recognition • 2B • Updated • 290k • 1.02kNote Machine translation (and other tasks)