Legal
MTEB: Legal est un benchmark public créé par MTEB / MMTEB (embeddings-benchmark) pour évaluer la recherche de documents juridiques dans plusieurs langues. Il couvre notamment la jurisprudence, les statuts, les questions-réponses juridiques et le résumé juridique.
MTEB: Legal est un benchmark public créé par MTEB / MMTEB (embeddings-benchmark) pour évaluer la recherche de documents juridiques dans plusieurs langues. Il couvre notamment la jurisprudence, les statuts, les questions-réponses juridiques et le résumé juridique.
Publié en 2025, cet agrégat de tâches mesure la capacité des modèles à retrouver des contenus pertinents dans différents contextes juridiques. Il fournit ainsi un cadre commun pour comparer des modèles d’embeddings sur des usages spécialisés et multilingues, à l’aide de nDCG@10 et de métriques adaptées à chaque tâche.
Carte d'identité
| Caractéristique | Valeur |
|---|---|
| Éditeur du benchmark | MTEB / MMTEB (embeddings-benchmark) |
| Capacités mesurées | Recherche de documents juridiques (jurisprudence, statuts, Q&R juridique, resume juridique) dans plusieurs langues. |
| Modalité | Texte |
| Type de questions | Recherche de documents juridiques, Q&R et resume |
| Métrique d'évaluation | nDCG@10 et autres metriques selon la tache |
| Accès | Public |
| Langues | Plusieurs langues |
| Taille du jeu | Agregat de taches juridiques (multilingue) |
| Année de publication | 2025 |
| Ressources | Site / dépôt officiel · Article scientifique |
Classement des modèles (156)
| # | Modèle | Éditeur | Licence | Score | Sortie | Fiabilité |
|---|---|---|---|---|---|---|
| 1 | Mira190/Euler-Legal-Embedding-V1 | Mira190 | 🟢 Ouvert | 70,4 % | 6 novembre 2025 | ✅ Mesuré |
| 2 | Hanno-Labs/dinghy-law-0.6b-v1 | Hanno-Labs | 🟢 Ouvert | 65,8 % | 13 juillet 2026 | ✅ Mesuré |
| 3 | voyageai/voyage-law-2 | Voyage AI | 🟢 Ouvert | 65,4 % | 15 avril 2024 | ✅ Mesuré |
| 4 | codefuse-ai/F2LLM-v2-14B | codefuse-ai | 🟢 Ouvert | 64,7 % | 9 mars 2026 | ✅ Mesuré |
| 5 | voyageai/voyage-3 | Voyage AI | 🟢 Ouvert | 64,1 % | 18 septembre 2024 | ✅ Mesuré |
| 6 | infly/inf-retriever-v1 | infly | 🟢 Ouvert | 63,7 % | 24 décembre 2024 | ✅ Mesuré |
| 7 | codefuse-ai/F2LLM-v2-8B | codefuse-ai | 🟢 Ouvert | 63,5 % | 9 mars 2026 | ✅ Mesuré |
| 8 | ICT-TIME-and-Querit/ICT-TIME-and-Querit-embedding-v1 | ICT-TIME-and-Querit | 🟢 Ouvert | 62,2 % | 2 mai 2026 | ✅ Mesuré |
| 9 | infly/inf-retriever-v1-1.5b | infly | 🟢 Ouvert | 62,0 % | 8 février 2025 | ✅ Mesuré |
| 10 | codefuse-ai/F2LLM-v2-4B | codefuse-ai | 🟢 Ouvert | 61,5 % | 9 mars 2026 | ✅ Mesuré |
| 11 | voyageai/voyage-3-lite | Voyage AI | 🟢 Ouvert | 60,8 % | 18 septembre 2024 | ✅ Mesuré |
| 12 | ICT-TIME-and-Querit/BOOM_4B_v1 | ICT-TIME-and-Querit | 🟢 Ouvert | 60,4 % | 31 janvier 2026 | ✅ Mesuré |
| 13 | codefuse-ai/F2LLM-v2-1.7B | codefuse-ai | 🟢 Ouvert | 60,3 % | 9 mars 2026 | ✅ Mesuré |
| 14 | openai/text-embedding-3-large | OpenAI | 🟢 Ouvert | 59,4 % | 25 janvier 2024 | ✅ Mesuré |
| 15 | Snowflake/snowflake-arctic-embed-l-v2.0 | Snowflake | 🟢 Ouvert | 59,0 % | 4 décembre 2024 | ✅ Mesuré |
| 16 | NovaSearch/jasper_en_vision_language_v1 | NovaSearch | 🟢 Ouvert | 58,8 % | 11 décembre 2024 | ✅ Mesuré |
| 17 | codefuse-ai/F2LLM-v2-0.6B | codefuse-ai | 🟢 Ouvert | 57,9 % | 9 mars 2026 | ✅ Mesuré |
| 18 | jinaai/jina-embeddings-v3 | Jina AI | 🟢 Ouvert | 56,9 % | 18 septembre 2024 | ✅ Mesuré |
| 19 | GritLM/GritLM-7B | GritLM | 🟢 Ouvert | 56,7 % | 15 février 2024 | ✅ Mesuré |
| 20 | codefuse-ai/F2LLM-v2-330M | codefuse-ai | 🟢 Ouvert | 56,7 % | 9 mars 2026 | ✅ Mesuré |
| 21 | Snowflake/snowflake-arctic-embed-m-v2.0 | Snowflake | 🟢 Ouvert | 56,6 % | 4 décembre 2024 | ✅ Mesuré |
| 22 | openai/text-embedding-3-small | OpenAI | 🟢 Ouvert | 56,3 % | 25 janvier 2024 | ✅ Mesuré |
| 23 | Salesforce/SFR-Embedding-2_R | Salesforce | 🟢 Ouvert | 55,8 % | 14 juin 2024 | ✅ Mesuré |
| 24 | Alibaba-NLP/gte-Qwen2-1.5B-instruct | Alibaba | 🟢 Ouvert | 55,6 % | 29 juillet 2024 | ✅ Mesuré |
| 25 | Linq-AI-Research/Linq-Embed-Mistral | Linq-AI-Research | 🟢 Ouvert | 55,4 % | 29 mai 2024 | ✅ Mesuré |
| 26 | infgrad/Jasper-Token-Compression-600M | infgrad | 🟢 Ouvert | 55,4 % | 14 novembre 2025 | ✅ Mesuré |
| 27 | Salesforce/SFR-Embedding-Mistral | Salesforce | 🟢 Ouvert | 55,3 % | 24 janvier 2024 | ✅ Mesuré |
| 28 | manu/bge-m3-custom-fr | manu | 🟢 Ouvert | 55,1 % | 11 avril 2024 | ✅ Mesuré |
| 29 | intfloat/multilingual-e5-large-instruct | Intfloat | 🟢 Ouvert | 54,9 % | 8 février 2024 | ✅ Mesuré |
| 30 | HIT-TMG/KaLM-embedding-multilingual-mini-instruct-v1 | HIT-TMG | 🟢 Ouvert | 54,5 % | 23 octobre 2024 | ✅ Mesuré |
| 31 | Alibaba-NLP/gte-multilingual-base | Alibaba | 🟢 Ouvert | 54,5 % | 20 juillet 2024 | ✅ Mesuré |
| 32 | intfloat/e5-mistral-7b-instruct | Intfloat | 🟢 Ouvert | 54,3 % | 8 février 2024 | ✅ Mesuré |
| 33 | Alibaba-NLP/gte-Qwen2-7B-instruct | Alibaba | 🟢 Ouvert | 54,2 % | 15 juin 2024 | ✅ Mesuré |
| 34 | GritLM/GritLM-8x7B | GritLM | 🟢 Ouvert | 54,1 % | 15 février 2024 | ✅ Mesuré |
| 35 | HIT-TMG/KaLM-embedding-multilingual-mini-v1 | HIT-TMG | 🟢 Ouvert | 53,9 % | 27 août 2024 | ✅ Mesuré |
| 36 | OrdalieTech/Solon-embeddings-large-0.1 | OrdalieTech | 🟢 Ouvert | 53,7 % | 9 décembre 2023 | ✅ Mesuré |
| 37 | Lajavaness/bilingual-embedding-large | Lajavaness | 🟢 Ouvert | 52,3 % | 24 juin 2024 | ✅ Mesuré |
| 38 | Cohere/Cohere-embed-multilingual-v3.0 | Cohere | 🟢 Ouvert | 52,1 % | 2 novembre 2023 | ✅ Mesuré |
| 39 | Alibaba-NLP/gte-Qwen1.5-7B-instruct | Alibaba | 🟢 Ouvert | 51,6 % | 20 avril 2024 | ✅ Mesuré |
| 40 | NovaSearch/stella_en_1.5B_v5 | NovaSearch | 🟢 Ouvert | 50,3 % | 12 juillet 2024 | ✅ Mesuré |
| 41 | Cohere/Cohere-embed-multilingual-light-v3.0 | Cohere | 🟢 Ouvert | 50,2 % | 2 novembre 2023 | ✅ Mesuré |
| 42 | codefuse-ai/F2LLM-v2-160M | codefuse-ai | 🟢 Ouvert | 49,9 % | 9 mars 2026 | ✅ Mesuré |
| 43 | Lajavaness/bilingual-embedding-small | Lajavaness | 🟢 Ouvert | 49,4 % | 17 juillet 2024 | ✅ Mesuré |
| 44 | Lajavaness/bilingual-embedding-base | Lajavaness | 🟢 Ouvert | 49,0 % | 26 juin 2024 | ✅ Mesuré |
| 45 | intfloat/multilingual-e5-base | Intfloat | 🟢 Ouvert | 48,5 % | 8 février 2024 | ✅ Mesuré |
| 46 | Intfloat: Multilingual-E5-Large | Intfloat | 🟢 Ouvert | 48,4 % | 18 novembre 2025 | ✅ Mesuré |
| 47 | deepvk/USER-bge-m3 | deepvk | 🟢 Ouvert | 47,4 % | 5 juillet 2024 | ✅ Mesuré |
| 48 | intfloat/multilingual-e5-small | Intfloat | 🟢 Ouvert | 47,4 % | 8 février 2024 | ✅ Mesuré |
| 49 | nvidia/NV-Embed-v1 | NVIDIA | 🟢 Ouvert | 46,5 % | 13 septembre 2024 | ✅ Mesuré |
| 50 | ibm-granite/granite-embedding-278m-multilingual | IBM | 🟢 Ouvert | 45,8 % | 18 décembre 2024 | ✅ Mesuré |
| 51 | nomic-ai/nomic-embed-text-v1-unsupervised | Nomic AI | 🟢 Ouvert | 45,5 % | 15 janvier 2024 | ✅ Mesuré |
| 52 | NovaSearch/stella_en_400M_v5 | NovaSearch | 🟢 Ouvert | 44,9 % | 12 juillet 2024 | ✅ Mesuré |
| 53 | sentence-transformers/static-similarity-mrl-multilingual-v1 | Sentence Transformers | 🟢 Ouvert | 44,8 % | 15 janvier 2025 | ✅ Mesuré |
| 54 | nvidia/NV-Embed-v2 | NVIDIA | 🟢 Ouvert | 44,3 % | 9 septembre 2024 | ✅ Mesuré |
| 55 | manu/sentence_croissant_alpha_v0.4 | manu | 🟢 Ouvert | 43,9 % | 27 avril 2024 | ✅ Mesuré |
| 56 | manu/sentence_croissant_alpha_v0.3 | manu | 🟢 Ouvert | 43,5 % | 26 avril 2024 | ✅ Mesuré |
| 57 | Cohere/Cohere-embed-english-v3.0 | Cohere | 🟢 Ouvert | 43,1 % | 2 novembre 2023 | ✅ Mesuré |
| 58 | codefuse-ai/F2LLM-v2-80M | codefuse-ai | 🟢 Ouvert | 43,0 % | 9 mars 2026 | ✅ Mesuré |
| 59 | Haon-Chen/speed-embedding-7b-instruct | Haon-Chen | 🟢 Ouvert | 43,0 % | 31 octobre 2024 | ✅ Mesuré |
| 60 | manu/sentence_croissant_alpha_v0.2 | manu | 🟢 Ouvert | 42,3 % | 15 mars 2024 | ✅ Mesuré |
| 61 | dwzhu/e5-base-4k | dwzhu | 🟢 Ouvert | 41,9 % | 28 mars 2024 | ✅ Mesuré |
| 62 | ibm-granite/granite-embedding-107m-multilingual | IBM | 🟢 Ouvert | 41,8 % | 18 décembre 2024 | ✅ Mesuré |
| 63 | w601sxs/b1ade-embed | w601sxs | 🟢 Ouvert | 41,8 % | 10 mars 2025 | ✅ Mesuré |
| 64 | nomic-ai/nomic-embed-text-v1 | Nomic AI | 🟢 Ouvert | 41,8 % | 31 janvier 2024 | ✅ Mesuré |
| 65 | Intfloat: E5-Large-v2 | Intfloat | 🟢 Ouvert | 41,7 % | 18 novembre 2025 | ✅ Mesuré |
| 66 | deepfile/embedder-100p | deepfile | 🟢 Ouvert | 41,6 % | 24 juillet 2023 | ✅ Mesuré |
| 67 | Snowflake/snowflake-arctic-embed-l | Snowflake | 🟢 Ouvert | 41,5 % | 12 avril 2024 | ✅ Mesuré |
| 68 | avsolatorio/GIST-large-Embedding-v0 | avsolatorio | 🟢 Ouvert | 41,5 % | 14 février 2024 | ✅ Mesuré |
| 69 | intfloat/e5-large | Intfloat | 🟢 Ouvert | 41,2 % | 26 décembre 2022 | ✅ Mesuré |
| 70 | Thenlper: GTE-Large | Alibaba | 🟢 Ouvert | 41,2 % | 18 novembre 2025 | ✅ Mesuré |
| 71 | jinaai/jina-embedding-b-en-v1 | Jina AI | 🟢 Ouvert | 40,8 % | 7 juillet 2023 | ✅ Mesuré |
| 72 | avsolatorio/GIST-Embedding-v0 | avsolatorio | 🟢 Ouvert | 40,7 % | 31 janvier 2024 | ✅ Mesuré |
| 73 | Cohere/Cohere-embed-english-light-v3.0 | Cohere | 🟢 Ouvert | 40,4 % | 2 novembre 2023 | ✅ Mesuré |
| 74 | ibm-granite/granite-embedding-125m-english | IBM | 🟢 Ouvert | 40,3 % | 18 décembre 2024 | ✅ Mesuré |
| 75 | Mihaiii/Ivysaur | Mihaiii | 🟢 Ouvert | 40,2 % | 27 avril 2024 | ✅ Mesuré |
| 76 | BAAI: bge-base-en-v1.5 | BAAI | 🟢 Ouvert | 40,2 % | 18 novembre 2025 | ✅ Mesuré |
| 77 | Thenlper: GTE-Base | Alibaba | 🟢 Ouvert | 40,0 % | 18 novembre 2025 | ✅ Mesuré |
| 78 | thenlper/gte-small | Alibaba | 🟢 Ouvert | 39,9 % | 27 juillet 2023 | ✅ Mesuré |
| 79 | mixedbread-ai/mxbai-embed-large-v1 | Mixedbread AI | 🟢 Ouvert | 39,9 % | 7 mars 2024 | ✅ Mesuré |
| 80 | Snowflake/snowflake-arctic-embed-m-v1.5 | Snowflake | 🟢 Ouvert | 39,8 % | 8 juillet 2024 | ✅ Mesuré |
| 81 | Snowflake/snowflake-arctic-embed-m | Snowflake | 🟢 Ouvert | 39,8 % | 12 avril 2024 | ✅ Mesuré |
| 82 | BAAI: bge-large-en-v1.5 | BAAI | 🟢 Ouvert | 39,6 % | 18 novembre 2025 | ✅ Mesuré |
| 83 | Intfloat: E5-Base-v2 | Intfloat | 🟢 Ouvert | 39,5 % | 18 novembre 2025 | ✅ Mesuré |
| 84 | avsolatorio/GIST-small-Embedding-v0 | avsolatorio | 🟢 Ouvert | 39,4 % | 3 février 2024 | ✅ Mesuré |
| 85 | sdadas/mmlw-e5-large | sdadas | 🟢 Ouvert | 39,3 % | 17 novembre 2023 | ✅ Mesuré |
| 86 | WhereIsAI/UAE-Large-V1 | WhereIsAI | 🟢 Ouvert | 39,3 % | 4 décembre 2023 | ✅ Mesuré |
| 87 | intfloat/e5-small-v2 | Intfloat | 🟢 Ouvert | 39,3 % | 8 février 2024 | ✅ Mesuré |
| 88 | avsolatorio/GIST-all-MiniLM-L6-v2 | avsolatorio | 🟢 Ouvert | 39,2 % | 3 février 2024 | ✅ Mesuré |
| 89 | omarelshehy/arabic-english-sts-matryoshka | omarelshehy | 🟢 Ouvert | 39,0 % | 13 octobre 2024 | ✅ Mesuré |
| 90 | avsolatorio/NoInstruct-small-Embedding-v0 | avsolatorio | 🟢 Ouvert | 38,9 % | 1 mai 2024 | ✅ Mesuré |
| 91 | infgrad/stella-base-en-v2 | infgrad | 🟢 Ouvert | 38,8 % | 19 octobre 2023 | ✅ Mesuré |
| 92 | BAAI/bge-small-en-v1.5 | BAAI | 🟢 Ouvert | 38,2 % | 12 septembre 2023 | ✅ Mesuré |
| 93 | Snowflake/snowflake-arctic-embed-s | Snowflake | 🟢 Ouvert | 38,0 % | 12 avril 2024 | ✅ Mesuré |
| 94 | jinaai/jina-embedding-s-en-v1 | Jina AI | 🟢 Ouvert | 37,9 % | 7 juillet 2023 | ✅ Mesuré |
| 95 | minishlab/potion-multilingual-128M | minishlab | 🟢 Ouvert | 37,9 % | 23 mai 2025 | ✅ Mesuré |
| 96 | intfloat/e5-small | Intfloat | 🟢 Ouvert | 37,8 % | 8 février 2024 | ✅ Mesuré |
| 97 | sdadas/mmlw-e5-base | sdadas | 🟢 Ouvert | 37,7 % | 17 novembre 2023 | ✅ Mesuré |
| 98 | Snowflake/snowflake-arctic-embed-m-long | Snowflake | 🟢 Ouvert | 37,5 % | 12 avril 2024 | ✅ Mesuré |
| 99 | abhinand/MedEmbed-small-v0.1 | abhinand | 🟢 Ouvert | 37,5 % | 20 octobre 2024 | ✅ Mesuré |
| 100 | Sentence Transformers: all-mpnet-base-v2 | Sentence Transformers | 🟢 Ouvert | 37,4 % | 17 novembre 2025 | ✅ Mesuré |
| 101 | sdadas/mmlw-roberta-large | sdadas | 🟢 Ouvert | 37,3 % | 17 novembre 2023 | ✅ Mesuré |
| 102 | intfloat/e5-base | Intfloat | 🟢 Ouvert | 37,0 % | 26 décembre 2022 | ✅ Mesuré |
| 103 | McGill-NLP/LLM2Vec-Mistral-7B-Instruct-v2-mntp-supervised | McGill-NLP | 🟢 Ouvert | 36,5 % | 9 avril 2024 | ✅ Mesuré |
| 104 | sentence-transformers/paraphrase-multilingual-mpnet-base-v2 | Sentence Transformers | 🟢 Ouvert | 36,4 % | 1 novembre 2019 | ✅ Mesuré |
| 105 | Mihaiii/gte-micro-v4 | Mihaiii | 🟢 Ouvert | 36,3 % | 22 avril 2024 | ✅ Mesuré |
| 106 | Omartificial-Intelligence-Space/Arabic-labse-Matryoshka | Omartificial-Intelligence-Space | 🟢 Ouvert | 35,9 % | 16 juin 2024 | ✅ Mesuré |
| 107 | nomic-ai/nomic-embed-text-v1.5 | Nomic AI | 🟢 Ouvert | 35,8 % | 10 février 2024 | ✅ Mesuré |
| 108 | Mihaiii/Bulbasaur | Mihaiii | 🟢 Ouvert | 35,8 % | 27 avril 2024 | ✅ Mesuré |
| 109 | ibm-granite/granite-embedding-30m-english | IBM | 🟢 Ouvert | 35,7 % | 18 décembre 2024 | ✅ Mesuré |
| 110 | Snowflake/snowflake-arctic-embed-xs | Snowflake | 🟢 Ouvert | 35,7 % | 8 juillet 2024 | ✅ Mesuré |
| 111 | sdadas/mmlw-e5-small | sdadas | 🟢 Ouvert | 35,7 % | 17 novembre 2023 | ✅ Mesuré |
| 112 | Sentence Transformers: all-MiniLM-L6-v2 | Sentence Transformers | 🟢 Ouvert | 35,4 % | 17 novembre 2025 | ✅ Mesuré |
| 113 | brahmairesearch/slx-v0.1 | brahmairesearch | 🟢 Ouvert | 35,4 % | 13 août 2024 | ✅ Mesuré |
| 114 | sdadas/mmlw-roberta-base | sdadas | 🟢 Ouvert | 35,3 % | 17 novembre 2023 | ✅ Mesuré |
| 115 | bigscience/sgpt-bloom-7b1-msmarco | bigscience | 🟢 Ouvert | 35,2 % | 26 août 2022 | ✅ Mesuré |
| 116 | Omartificial-Intelligence-Space/Arabic-all-nli-triplet-Matryoshka | Omartificial-Intelligence-Space | 🟢 Ouvert | 35,0 % | 14 juin 2024 | ✅ Mesuré |
| 117 | minishlab/potion-base-8M | minishlab | 🟢 Ouvert | 34,8 % | 29 octobre 2024 | ✅ Mesuré |
| 118 | aari1995/German_Semantic_STS_V2 | aari1995 | 🟢 Ouvert | 34,6 % | 17 novembre 2022 | ✅ Mesuré |
| 119 | sentence-transformers/paraphrase-multilingual-MiniLM-L12-v2 | Sentence Transformers | 🟢 Ouvert | 34,4 % | 1 novembre 2019 | ✅ Mesuré |
| 120 | Sentence Transformers: all-MiniLM-L12-v2 | Sentence Transformers | 🟢 Ouvert | 33,9 % | 18 novembre 2025 | ✅ Mesuré |
| 121 | sergeyzh/LaBSE-ru-turbo | sergeyzh | 🟢 Ouvert | 33,4 % | 27 juin 2024 | ✅ Mesuré |
| 122 | Omartificial-Intelligence-Space/Arabic-MiniLM-L12-v2-all-nli-triplet | Omartificial-Intelligence-Space | 🟢 Ouvert | 33,2 % | 25 juin 2024 | ✅ Mesuré |
| 123 | Mihaiii/Wartortle | Mihaiii | 🟢 Ouvert | 33,0 % | 30 avril 2024 | ✅ Mesuré |
| 124 | minishlab/potion-base-4M | minishlab | 🟢 Ouvert | 32,9 % | 29 octobre 2024 | ✅ Mesuré |
| 125 | ai-forever/ru-en-RoSBERTa | ai-forever | 🟢 Ouvert | 32,6 % | 29 juillet 2024 | ✅ Mesuré |
| 126 | sentence-transformers/LaBSE | Sentence Transformers | 🟢 Ouvert | 32,3 % | 1 novembre 2019 | ✅ Mesuré |
| 127 | Gameselo/STS-multilingual-mpnet-base-v2 | Gameselo | 🟢 Ouvert | 32,3 % | 7 juin 2024 | ✅ Mesuré |
| 128 | Mihaiii/Venusaur | Mihaiii | 🟢 Ouvert | 31,4 % | 29 avril 2024 | ✅ Mesuré |
| 129 | minishlab/M2V_base_output | minishlab | 🟢 Ouvert | 31,2 % | 21 septembre 2024 | ✅ Mesuré |
| 130 | minishlab/M2V_base_glove_subword | minishlab | 🟢 Ouvert | 31,0 % | 21 septembre 2024 | ✅ Mesuré |
| 131 | shibing624/text2vec-base-multilingual | shibing624 | 🟢 Ouvert | 30,6 % | 22 juin 2023 | ✅ Mesuré |
| 132 | Mihaiii/Squirtle | Mihaiii | 🟢 Ouvert | 30,1 % | 30 avril 2024 | ✅ Mesuré |
| 133 | cointegrated/LaBSE-en-ru | cointegrated | 🟢 Ouvert | 30,1 % | 10 juin 2021 | ✅ Mesuré |
| 134 | minishlab/potion-base-2M | minishlab | 🟢 Ouvert | 29,4 % | 29 octobre 2024 | ✅ Mesuré |
| 135 | consciousAI/cai-lunaris-text-embeddings | consciousAI | 🟢 Ouvert | 28,9 % | 22 juin 2023 | ✅ Mesuré |
| 136 | minishlab/M2V_base_glove | minishlab | 🟢 Ouvert | 28,3 % | 21 septembre 2024 | ✅ Mesuré |
| 137 | McGill-NLP/LLM2Vec-Mistral-7B-Instruct-v2-mntp-unsup-simcse | McGill-NLP | 🟢 Ouvert | 27,0 % | 9 avril 2024 | ✅ Mesuré |
| 138 | Mihaiii/gte-micro | Mihaiii | 🟢 Ouvert | 26,4 % | 21 avril 2024 | ✅ Mesuré |
| 139 | deepvk/USER-base | deepvk | 🟢 Ouvert | 26,2 % | 10 juin 2024 | ✅ Mesuré |
| 140 | Jaume/gemma-2b-embeddings | Jaume | 🟢 Ouvert | 23,9 % | 29 juin 2024 | ✅ Mesuré |
| 141 | sergeyzh/rubert-tiny-turbo | sergeyzh | 🟢 Ouvert | 23,0 % | 21 juin 2024 | ✅ Mesuré |
| 142 | cointegrated/rubert-tiny2 | cointegrated | 🟢 Ouvert | 22,4 % | 28 octobre 2021 | ✅ Mesuré |
| 143 | Omartificial-Intelligence-Space/Arabic-mpnet-base-all-nli-triplet | Omartificial-Intelligence-Space | 🟢 Ouvert | 20,1 % | 15 juin 2024 | ✅ Mesuré |
| 144 | cointegrated/rubert-tiny | cointegrated | 🟢 Ouvert | 20,0 % | 24 mai 2021 | ✅ Mesuré |
| 145 | silma-ai/silma-embeddding-matryoshka-v0.1 | silma-ai | 🟢 Ouvert | 19,9 % | 12 octobre 2024 | ✅ Mesuré |
| 146 | Omartificial-Intelligence-Space/Arabert-all-nli-triplet-Matryoshka | Omartificial-Intelligence-Space | 🟢 Ouvert | 16,4 % | 16 juin 2024 | ✅ Mesuré |
| 147 | jinaai/jina-embeddings-v2-small-en | Jina AI | 🟢 Ouvert | 16,0 % | 27 septembre 2023 | ✅ Mesuré |
| 148 | DeepPavlov/rubert-base-cased-sentence | DeepPavlov | 🟢 Ouvert | 15,1 % | 4 mars 2020 | ✅ Mesuré |
| 149 | DeepPavlov/distilrubert-small-cased-conversational | DeepPavlov | 🟢 Ouvert | 14,9 % | 28 juin 2022 | ✅ Mesuré |
| 150 | ai-forever/sbert_large_nlu_ru | ai-forever | 🟢 Ouvert | 11,9 % | 20 novembre 2020 | ✅ Mesuré |
| 151 | Omartificial-Intelligence-Space/Marbert-all-nli-triplet-Matryoshka | Omartificial-Intelligence-Space | 🟢 Ouvert | 11,8 % | 17 juin 2024 | ✅ Mesuré |
| 152 | DeepPavlov/rubert-base-cased | DeepPavlov | 🟢 Ouvert | 11,7 % | 4 mars 2020 | ✅ Mesuré |
| 153 | malenia1/ternary-weight-embedding | malenia1 | 🟢 Ouvert | 11,6 % | 23 octobre 2024 | ✅ Mesuré |
| 154 | ai-forever/sbert_large_mt_nlu_ru | ai-forever | 🟢 Ouvert | 10,6 % | 18 mai 2021 | ✅ Mesuré |
| 155 | deepvk/deberta-v1-base | deepvk | 🟢 Ouvert | 9,3 % | 7 février 2023 | ✅ Mesuré |
| 156 | jinaai/jina-embeddings-v2-base-en | Jina AI | 🟢 Ouvert | 3,2 % | 27 septembre 2023 | ✅ Mesuré |
Classement établi sur 156 modèles évalués, dont 4 de grands éditeurs. Score médian de l'ensemble : 39,9 %.
Notre analyse
Un score élevé sur MTEB: Legal indique qu’un modèle classe efficacement les documents juridiques pertinents parmi les premiers résultats, sur un ensemble couvrant plusieurs types de contenus et plusieurs langues. Dans la base observée, Mira190/Euler-Legal-Embedding-V1 atteint 70 %, contre une médiane de 40 % pour 155 modèles. Cet écart révèle une forte différenciation entre les systèmes évalués et ne suggère pas de saturation générale du classement.
L’interprétation doit toutefois rester prudente. Les scores sont majoritairement auto-déclarés par les éditeurs, ce qui réduit l’homogénéité de la vérification par rapport à une évaluation entièrement mesurée dans un cadre unique. L’accès public au benchmark peut aussi favoriser une adaptation spécifique aux tâches, avec un risque de contamination entre données d’évaluation et processus de développement. Enfin, sa portée reste celle de la recherche juridique, des questions-réponses et du résumé juridique multilingues. Le classement renseigne donc sur ces capacités ciblées, plutôt que sur les performances générales d’un modèle hors de ce périmètre.
Sources des scores : mteb.