Medical
MTEB: Medical est un benchmark public créé en 2025 par MTEB / MMTEB (embeddings-benchmark). Il évalue la recherche d’information médicale dans des contenus cliniques, biomédicaux et destinés au grand public.
MTEB: Medical est un benchmark public créé en 2025 par MTEB / MMTEB (embeddings-benchmark). Il évalue la recherche d’information médicale dans des contenus cliniques, biomédicaux et destinés au grand public.
Son agrégat de tâches couvre le retrieval, le reranking et le clustering, principalement sur des textes anglais, avec quelques autres langues. Il sert à comparer la capacité des modèles d’embeddings à retrouver, ordonner ou regrouper des informations médicales pertinentes dans plusieurs contextes.
Carte d'identité
| Caractéristique | Valeur |
|---|---|
| Éditeur du benchmark | MTEB / MMTEB (embeddings-benchmark) |
| Capacités mesurées | Recherche d'information medicale a travers les domaines clinique, biomedical et sante grand public (retrieval, reranking, clustering). |
| Modalité | Texte |
| Type de questions | Recherche d'information, reranking et clustering sur textes medicaux |
| Métrique d'évaluation | nDCG@10 et autres metriques selon la tache |
| Accès | Public |
| Langues | Principalement anglais (avec quelques autres langues) |
| Taille du jeu | Agregat de taches medicales (env. 12 jeux de donnees dans la section medicale de MTEB) |
| Année de publication | 2025 |
| Ressources | Site / dépôt officiel · Article scientifique |
Classement des modèles (158)
| # | Modèle | Éditeur | Licence | Score | Sortie | Fiabilité |
|---|---|---|---|---|---|---|
| 1 | Alibaba-NLP/gte-Qwen2-7B-instruct | Alibaba | 🟢 Ouvert | 66,6 % | 15 juin 2024 | ✅ Mesuré |
| 2 | codefuse-ai/F2LLM-v2-14B | codefuse-ai | 🟢 Ouvert | 65,2 % | 9 mars 2026 | ✅ Mesuré |
| 3 | Alibaba-NLP/gte-Qwen2-1.5B-instruct | Alibaba | 🟢 Ouvert | 65,0 % | 29 juillet 2024 | ✅ Mesuré |
| 4 | codefuse-ai/F2LLM-v2-8B | codefuse-ai | 🟢 Ouvert | 64,9 % | 9 mars 2026 | ✅ Mesuré |
| 5 | codefuse-ai/F2LLM-v2-4B | codefuse-ai | 🟢 Ouvert | 64,5 % | 9 mars 2026 | ✅ Mesuré |
| 6 | ICT-TIME-and-Querit/BOOM_4B_v1 | ICT-TIME-and-Querit | 🟢 Ouvert | 64,5 % | 31 janvier 2026 | ✅ Mesuré |
| 7 | Salesforce/SFR-Embedding-Mistral | Salesforce | 🟢 Ouvert | 64,5 % | 24 janvier 2024 | ✅ Mesuré |
| 8 | Salesforce/SFR-Embedding-2_R | Salesforce | 🟢 Ouvert | 64,3 % | 14 juin 2024 | ✅ Mesuré |
| 9 | ICT-TIME-and-Querit/ICT-TIME-and-Querit-embedding-v1 | ICT-TIME-and-Querit | 🟢 Ouvert | 64,3 % | 2 mai 2026 | ✅ Mesuré |
| 10 | Linq-AI-Research/Linq-Embed-Mistral | Linq-AI-Research | 🟢 Ouvert | 63,9 % | 29 mai 2024 | ✅ Mesuré |
| 11 | nvidia/NV-Embed-v2 | NVIDIA | 🟢 Ouvert | 63,6 % | 9 septembre 2024 | ✅ Mesuré |
| 12 | nvidia/NV-Embed-v1 | NVIDIA | 🟢 Ouvert | 63,0 % | 13 septembre 2024 | ✅ Mesuré |
| 13 | intfloat/e5-mistral-7b-instruct | Intfloat | 🟢 Ouvert | 62,7 % | 8 février 2024 | ✅ Mesuré |
| 14 | Alibaba-NLP/gte-Qwen1.5-7B-instruct | Alibaba | 🟢 Ouvert | 62,2 % | 20 avril 2024 | ✅ Mesuré |
| 15 | voyageai/voyage-3 | Voyage AI | 🟢 Ouvert | 61,5 % | 18 septembre 2024 | ✅ Mesuré |
| 16 | GritLM/GritLM-7B | GritLM | 🟢 Ouvert | 61,5 % | 15 février 2024 | ✅ Mesuré |
| 17 | codefuse-ai/F2LLM-v2-1.7B | codefuse-ai | 🟢 Ouvert | 61,4 % | 9 mars 2026 | ✅ Mesuré |
| 18 | NovaSearch/jasper_en_vision_language_v1 | NovaSearch | 🟢 Ouvert | 61,2 % | 11 décembre 2024 | ✅ Mesuré |
| 19 | openai/text-embedding-3-large | OpenAI | 🟢 Ouvert | 61,0 % | 25 janvier 2024 | ✅ Mesuré |
| 20 | voyageai/voyage-multilingual-2 | Voyage AI | 🟢 Ouvert | 61,0 % | 10 juin 2024 | ✅ Mesuré |
| 21 | GritLM/GritLM-8x7B | GritLM | 🟢 Ouvert | 60,9 % | 15 février 2024 | ✅ Mesuré |
| 22 | Cohere/Cohere-embed-multilingual-v3.0 | Cohere | 🟢 Ouvert | 60,8 % | 2 novembre 2023 | ✅ Mesuré |
| 23 | jinaai/jina-embeddings-v3 | Jina AI | 🟢 Ouvert | 59,5 % | 18 septembre 2024 | ✅ Mesuré |
| 24 | NovaSearch/stella_en_1.5B_v5 | NovaSearch | 🟢 Ouvert | 59,4 % | 12 juillet 2024 | ✅ Mesuré |
| 25 | Snowflake/snowflake-arctic-embed-l-v2.0 | Snowflake | 🟢 Ouvert | 59,4 % | 4 décembre 2024 | ✅ Mesuré |
| 26 | intfloat/multilingual-e5-large-instruct | Intfloat | 🟢 Ouvert | 58,5 % | 8 février 2024 | ✅ Mesuré |
| 27 | codefuse-ai/F2LLM-v2-0.6B | codefuse-ai | 🟢 Ouvert | 57,9 % | 9 mars 2026 | ✅ Mesuré |
| 28 | HIT-TMG/KaLM-embedding-multilingual-mini-instruct-v1 | HIT-TMG | 🟢 Ouvert | 57,3 % | 23 octobre 2024 | ✅ Mesuré |
| 29 | HIT-TMG/KaLM-embedding-multilingual-mini-v1 | HIT-TMG | 🟢 Ouvert | 57,3 % | 27 août 2024 | ✅ Mesuré |
| 30 | Snowflake/snowflake-arctic-embed-m-v2.0 | Snowflake | 🟢 Ouvert | 57,3 % | 4 décembre 2024 | ✅ Mesuré |
| 31 | Alibaba-NLP/gte-multilingual-base | Alibaba | 🟢 Ouvert | 57,0 % | 20 juillet 2024 | ✅ Mesuré |
| 32 | Intfloat: Multilingual-E5-Large | Intfloat | 🟢 Ouvert | 56,6 % | 18 novembre 2025 | ✅ Mesuré |
| 33 | codefuse-ai/F2LLM-v2-330M | codefuse-ai | 🟢 Ouvert | 56,4 % | 9 mars 2026 | ✅ Mesuré |
| 34 | Cohere/Cohere-embed-multilingual-light-v3.0 | Cohere | 🟢 Ouvert | 56,1 % | 2 novembre 2023 | ✅ Mesuré |
| 35 | openai/text-embedding-3-small | OpenAI | 🟢 Ouvert | 55,6 % | 25 janvier 2024 | ✅ Mesuré |
| 36 | voyageai/voyage-3-lite | Voyage AI | 🟢 Ouvert | 55,4 % | 18 septembre 2024 | ✅ Mesuré |
| 37 | intfloat/multilingual-e5-base | Intfloat | 🟢 Ouvert | 54,3 % | 8 février 2024 | ✅ Mesuré |
| 38 | OrdalieTech/Solon-embeddings-large-0.1 | OrdalieTech | 🟢 Ouvert | 54,1 % | 9 décembre 2023 | ✅ Mesuré |
| 39 | BAAI: bge-m3 | BAAI | 🟢 Ouvert | 53,6 % | 18 novembre 2025 | ✅ Mesuré |
| 40 | Lajavaness/bilingual-embedding-large | Lajavaness | 🟢 Ouvert | 53,5 % | 24 juin 2024 | ✅ Mesuré |
| 41 | intfloat/multilingual-e5-small | Intfloat | 🟢 Ouvert | 53,4 % | 8 février 2024 | ✅ Mesuré |
| 42 | codefuse-ai/F2LLM-v2-160M | codefuse-ai | 🟢 Ouvert | 52,4 % | 9 mars 2026 | ✅ Mesuré |
| 43 | ibm-granite/granite-embedding-278m-multilingual | IBM | 🟢 Ouvert | 50,9 % | 18 décembre 2024 | ✅ Mesuré |
| 44 | codefuse-ai/F2LLM-v2-80M | codefuse-ai | 🟢 Ouvert | 50,7 % | 9 mars 2026 | ✅ Mesuré |
| 45 | Lajavaness/bilingual-embedding-base | Lajavaness | 🟢 Ouvert | 50,6 % | 26 juin 2024 | ✅ Mesuré |
| 46 | manu/bge-m3-custom-fr | manu | 🟢 Ouvert | 49,0 % | 11 avril 2024 | ✅ Mesuré |
| 47 | McGill-NLP/LLM2Vec-Mistral-7B-Instruct-v2-mntp-supervised | McGill-NLP | 🟢 Ouvert | 48,3 % | 9 avril 2024 | ✅ Mesuré |
| 48 | Haon-Chen/speed-embedding-7b-instruct | Haon-Chen | 🟢 Ouvert | 47,0 % | 31 octobre 2024 | ✅ Mesuré |
| 49 | ibm-granite/granite-embedding-107m-multilingual | IBM | 🟢 Ouvert | 46,8 % | 18 décembre 2024 | ✅ Mesuré |
| 50 | Lajavaness/bilingual-embedding-small | Lajavaness | 🟢 Ouvert | 46,7 % | 17 juillet 2024 | ✅ Mesuré |
| 51 | Cohere/Cohere-embed-english-v3.0 | Cohere | 🟢 Ouvert | 45,3 % | 2 novembre 2023 | ✅ Mesuré |
| 52 | nomic-ai/nomic-embed-text-v1-ablated | Nomic AI | 🟢 Ouvert | 43,5 % | 15 janvier 2024 | ✅ Mesuré |
| 53 | NovaSearch/stella_en_400M_v5 | NovaSearch | 🟢 Ouvert | 43,5 % | 12 juillet 2024 | ✅ Mesuré |
| 54 | intfloat/e5-large | Intfloat | 🟢 Ouvert | 42,9 % | 26 décembre 2022 | ✅ Mesuré |
| 55 | deepvk/USER-bge-m3 | deepvk | 🟢 Ouvert | 42,9 % | 5 juillet 2024 | ✅ Mesuré |
| 56 | BAAI: bge-base-en-v1.5 | BAAI | 🟢 Ouvert | 42,9 % | 18 novembre 2025 | ✅ Mesuré |
| 57 | Snowflake/snowflake-arctic-embed-l | Snowflake | 🟢 Ouvert | 42,3 % | 12 avril 2024 | ✅ Mesuré |
| 58 | Thenlper: GTE-Large | Alibaba | 🟢 Ouvert | 42,1 % | 18 novembre 2025 | ✅ Mesuré |
| 59 | BAAI: bge-large-en-v1.5 | BAAI | 🟢 Ouvert | 42,1 % | 18 novembre 2025 | ✅ Mesuré |
| 60 | WhereIsAI/UAE-Large-V1 | WhereIsAI | 🟢 Ouvert | 41,4 % | 4 décembre 2023 | ✅ Mesuré |
| 61 | Cohere/Cohere-embed-english-light-v3.0 | Cohere | 🟢 Ouvert | 41,4 % | 2 novembre 2023 | ✅ Mesuré |
| 62 | Intfloat: E5-Large-v2 | Intfloat | 🟢 Ouvert | 41,2 % | 18 novembre 2025 | ✅ Mesuré |
| 63 | mixedbread-ai/mxbai-embed-large-v1 | Mixedbread AI | 🟢 Ouvert | 41,2 % | 7 mars 2024 | ✅ Mesuré |
| 64 | w601sxs/b1ade-embed | w601sxs | 🟢 Ouvert | 41,2 % | 10 mars 2025 | ✅ Mesuré |
| 65 | sdadas/mmlw-e5-large | sdadas | 🟢 Ouvert | 41,0 % | 17 novembre 2023 | ✅ Mesuré |
| 66 | nomic-ai/nomic-embed-text-v1-unsupervised | Nomic AI | 🟢 Ouvert | 40,9 % | 15 janvier 2024 | ✅ Mesuré |
| 67 | ibm-granite/granite-embedding-125m-english | IBM | 🟢 Ouvert | 40,8 % | 18 décembre 2024 | ✅ Mesuré |
| 68 | Snowflake/snowflake-arctic-embed-m | Snowflake | 🟢 Ouvert | 40,8 % | 12 avril 2024 | ✅ Mesuré |
| 69 | intfloat/e5-base | Intfloat | 🟢 Ouvert | 40,8 % | 26 décembre 2022 | ✅ Mesuré |
| 70 | nomic-ai/nomic-embed-text-v1 | Nomic AI | 🟢 Ouvert | 40,7 % | 31 janvier 2024 | ✅ Mesuré |
| 71 | Thenlper: GTE-Base | Alibaba | 🟢 Ouvert | 40,4 % | 18 novembre 2025 | ✅ Mesuré |
| 72 | Snowflake/snowflake-arctic-embed-m-v1.5 | Snowflake | 🟢 Ouvert | 40,3 % | 8 juillet 2024 | ✅ Mesuré |
| 73 | avsolatorio/GIST-large-Embedding-v0 | avsolatorio | 🟢 Ouvert | 40,2 % | 14 février 2024 | ✅ Mesuré |
| 74 | Intfloat: E5-Base-v2 | Intfloat | 🟢 Ouvert | 40,1 % | 18 novembre 2025 | ✅ Mesuré |
| 75 | avsolatorio/GIST-Embedding-v0 | avsolatorio | 🟢 Ouvert | 40,0 % | 31 janvier 2024 | ✅ Mesuré |
| 76 | sdadas/mmlw-e5-base | sdadas | 🟢 Ouvert | 39,8 % | 17 novembre 2023 | ✅ Mesuré |
| 77 | manu/sentence_croissant_alpha_v0.3 | manu | 🟢 Ouvert | 39,4 % | 26 avril 2024 | ✅ Mesuré |
| 78 | sentence-transformers/paraphrase-multilingual-mpnet-base-v2 | Sentence Transformers | 🟢 Ouvert | 39,3 % | 1 novembre 2019 | ✅ Mesuré |
| 79 | abhinand/MedEmbed-small-v0.1 | abhinand | 🟢 Ouvert | 39,2 % | 20 octobre 2024 | ✅ Mesuré |
| 80 | sdadas/mmlw-roberta-large | sdadas | 🟢 Ouvert | 39,2 % | 17 novembre 2023 | ✅ Mesuré |
| 81 | manu/sentence_croissant_alpha_v0.4 | manu | 🟢 Ouvert | 39,1 % | 27 avril 2024 | ✅ Mesuré |
| 82 | BAAI/bge-small-en-v1.5 | BAAI | 🟢 Ouvert | 39,0 % | 12 septembre 2023 | ✅ Mesuré |
| 83 | bigscience/sgpt-bloom-7b1-msmarco | bigscience | 🟢 Ouvert | 38,9 % | 26 août 2022 | ✅ Mesuré |
| 84 | avsolatorio/NoInstruct-small-Embedding-v0 | avsolatorio | 🟢 Ouvert | 38,8 % | 1 mai 2024 | ✅ Mesuré |
| 85 | sdadas/mmlw-roberta-base | sdadas | 🟢 Ouvert | 38,6 % | 17 novembre 2023 | ✅ Mesuré |
| 86 | intfloat/e5-small-v2 | Intfloat | 🟢 Ouvert | 38,6 % | 8 février 2024 | ✅ Mesuré |
| 87 | thenlper/gte-small | Alibaba | 🟢 Ouvert | 38,5 % | 27 juillet 2023 | ✅ Mesuré |
| 88 | Snowflake/snowflake-arctic-embed-m-long | Snowflake | 🟢 Ouvert | 38,4 % | 12 avril 2024 | ✅ Mesuré |
| 89 | infgrad/stella-base-en-v2 | infgrad | 🟢 Ouvert | 38,2 % | 19 octobre 2023 | ✅ Mesuré |
| 90 | Snowflake/snowflake-arctic-embed-s | Snowflake | 🟢 Ouvert | 38,1 % | 12 avril 2024 | ✅ Mesuré |
| 91 | sentence-transformers/paraphrase-multilingual-MiniLM-L12-v2 | Sentence Transformers | 🟢 Ouvert | 38,0 % | 1 novembre 2019 | ✅ Mesuré |
| 92 | intfloat/e5-small | Intfloat | 🟢 Ouvert | 37,9 % | 8 février 2024 | ✅ Mesuré |
| 93 | ibm-granite/granite-embedding-30m-english | IBM | 🟢 Ouvert | 37,3 % | 18 décembre 2024 | ✅ Mesuré |
| 94 | nomic-ai/nomic-embed-text-v1.5 | Nomic AI | 🟢 Ouvert | 37,3 % | 10 février 2024 | ✅ Mesuré |
| 95 | manu/sentence_croissant_alpha_v0.2 | manu | 🟢 Ouvert | 37,3 % | 15 mars 2024 | ✅ Mesuré |
| 96 | Omartificial-Intelligence-Space/Arabic-all-nli-triplet-Matryoshka | Omartificial-Intelligence-Space | 🟢 Ouvert | 37,2 % | 14 juin 2024 | ✅ Mesuré |
| 97 | dwzhu/e5-base-4k | dwzhu | 🟢 Ouvert | 36,9 % | 28 mars 2024 | ✅ Mesuré |
| 98 | avsolatorio/GIST-small-Embedding-v0 | avsolatorio | 🟢 Ouvert | 36,7 % | 3 février 2024 | ✅ Mesuré |
| 99 | Mihaiii/Ivysaur | Mihaiii | 🟢 Ouvert | 35,8 % | 27 avril 2024 | ✅ Mesuré |
| 100 | sdadas/mmlw-e5-small | sdadas | 🟢 Ouvert | 35,3 % | 17 novembre 2023 | ✅ Mesuré |
| 101 | Snowflake/snowflake-arctic-embed-xs | Snowflake | 🟢 Ouvert | 34,3 % | 8 juillet 2024 | ✅ Mesuré |
| 102 | McGill-NLP/LLM2Vec-Mistral-7B-Instruct-v2-mntp-unsup-simcse | McGill-NLP | 🟢 Ouvert | 34,2 % | 9 avril 2024 | ✅ Mesuré |
| 103 | Sentence Transformers: all-MiniLM-L12-v2 | Sentence Transformers | 🟢 Ouvert | 33,2 % | 18 novembre 2025 | ✅ Mesuré |
| 104 | Omartificial-Intelligence-Space/Arabic-MiniLM-L12-v2-all-nli-triplet | Omartificial-Intelligence-Space | 🟢 Ouvert | 33,2 % | 25 juin 2024 | ✅ Mesuré |
| 105 | jinaai/jina-embedding-b-en-v1 | Jina AI | 🟢 Ouvert | 33,1 % | 7 juillet 2023 | ✅ Mesuré |
| 106 | Sentence Transformers: all-mpnet-base-v2 | Sentence Transformers | 🟢 Ouvert | 32,7 % | 17 novembre 2025 | ✅ Mesuré |
| 107 | shibing624/text2vec-base-multilingual | shibing624 | 🟢 Ouvert | 32,2 % | 22 juin 2023 | ✅ Mesuré |
| 108 | avsolatorio/GIST-all-MiniLM-L6-v2 | avsolatorio | 🟢 Ouvert | 32,2 % | 3 février 2024 | ✅ Mesuré |
| 109 | sentence-transformers/static-similarity-mrl-multilingual-v1 | Sentence Transformers | 🟢 Ouvert | 32,1 % | 15 janvier 2025 | ✅ Mesuré |
| 110 | Gameselo/STS-multilingual-mpnet-base-v2 | Gameselo | 🟢 Ouvert | 31,9 % | 7 juin 2024 | ✅ Mesuré |
| 111 | Omartificial-Intelligence-Space/Arabic-labse-Matryoshka | Omartificial-Intelligence-Space | 🟢 Ouvert | 31,7 % | 16 juin 2024 | ✅ Mesuré |
| 112 | brahmairesearch/slx-v0.1 | brahmairesearch | 🟢 Ouvert | 30,9 % | 13 août 2024 | ✅ Mesuré |
| 113 | ai-forever/ru-en-RoSBERTa | ai-forever | 🟢 Ouvert | 30,7 % | 29 juillet 2024 | ✅ Mesuré |
| 114 | Mihaiii/Bulbasaur | Mihaiii | 🟢 Ouvert | 30,5 % | 27 avril 2024 | ✅ Mesuré |
| 115 | Sentence Transformers: all-MiniLM-L6-v2 | Sentence Transformers | 🟢 Ouvert | 29,8 % | 17 novembre 2025 | ✅ Mesuré |
| 116 | Mihaiii/gte-micro-v4 | Mihaiii | 🟢 Ouvert | 29,8 % | 22 avril 2024 | ✅ Mesuré |
| 117 | deepfile/embedder-100p | deepfile | 🟢 Ouvert | 29,6 % | 24 juillet 2023 | ✅ Mesuré |
| 118 | sergeyzh/LaBSE-ru-turbo | sergeyzh | 🟢 Ouvert | 29,5 % | 27 juin 2024 | ✅ Mesuré |
| 119 | sentence-transformers/LaBSE | Sentence Transformers | 🟢 Ouvert | 28,8 % | 1 novembre 2019 | ✅ Mesuré |
| 120 | jinaai/jina-embedding-s-en-v1 | Jina AI | 🟢 Ouvert | 28,4 % | 7 juillet 2023 | ✅ Mesuré |
| 121 | minishlab/potion-multilingual-128M | minishlab | 🟢 Ouvert | 27,8 % | 23 mai 2025 | ✅ Mesuré |
| 122 | NeuML/pubmedbert-base-embeddings-8M | NeuML | 🟢 Ouvert | 27,6 % | 3 janvier 2025 | ✅ Mesuré |
| 123 | minishlab/potion-base-8M | minishlab | 🟢 Ouvert | 27,5 % | 29 octobre 2024 | ✅ Mesuré |
| 124 | omarelshehy/arabic-english-sts-matryoshka | omarelshehy | 🟢 Ouvert | 26,2 % | 13 octobre 2024 | ✅ Mesuré |
| 125 | minishlab/M2V_base_glove_subword | minishlab | 🟢 Ouvert | 26,1 % | 21 septembre 2024 | ✅ Mesuré |
| 126 | minishlab/potion-base-4M | minishlab | 🟢 Ouvert | 25,8 % | 29 octobre 2024 | ✅ Mesuré |
| 127 | Mihaiii/Wartortle | Mihaiii | 🟢 Ouvert | 25,2 % | 30 avril 2024 | ✅ Mesuré |
| 128 | minishlab/M2V_base_output | minishlab | 🟢 Ouvert | 24,5 % | 21 septembre 2024 | ✅ Mesuré |
| 129 | Mihaiii/Venusaur | Mihaiii | 🟢 Ouvert | 24,2 % | 29 avril 2024 | ✅ Mesuré |
| 130 | deepvk/USER-base | deepvk | 🟢 Ouvert | 23,8 % | 10 juin 2024 | ✅ Mesuré |
| 131 | Mihaiii/Squirtle | Mihaiii | 🟢 Ouvert | 23,7 % | 30 avril 2024 | ✅ Mesuré |
| 132 | NeuML/pubmedbert-base-embeddings-2M | NeuML | 🟢 Ouvert | 23,6 % | 3 janvier 2025 | ✅ Mesuré |
| 133 | consciousAI/cai-lunaris-text-embeddings | consciousAI | 🟢 Ouvert | 23,2 % | 22 juin 2023 | ✅ Mesuré |
| 134 | NeuML/pubmedbert-base-embeddings-1M | NeuML | 🟢 Ouvert | 23,2 % | 3 janvier 2025 | ✅ Mesuré |
| 135 | minishlab/M2V_base_glove | minishlab | 🟢 Ouvert | 22,4 % | 21 septembre 2024 | ✅ Mesuré |
| 136 | minishlab/potion-base-2M | minishlab | 🟢 Ouvert | 22,2 % | 29 octobre 2024 | ✅ Mesuré |
| 137 | cointegrated/LaBSE-en-ru | cointegrated | 🟢 Ouvert | 22,0 % | 10 juin 2021 | ✅ Mesuré |
| 138 | Mihaiii/gte-micro | Mihaiii | 🟢 Ouvert | 21,3 % | 21 avril 2024 | ✅ Mesuré |
| 139 | jinaai/jina-embeddings-v2-base-en | Jina AI | 🟢 Ouvert | 21,1 % | 27 septembre 2023 | ✅ Mesuré |
| 140 | jinaai/jina-embeddings-v2-small-en | Jina AI | 🟢 Ouvert | 21,0 % | 27 septembre 2023 | ✅ Mesuré |
| 141 | NeuML/pubmedbert-base-embeddings-500K | NeuML | 🟢 Ouvert | 20,6 % | 3 janvier 2025 | ✅ Mesuré |
| 142 | sergeyzh/rubert-tiny-turbo | sergeyzh | 🟢 Ouvert | 18,6 % | 21 juin 2024 | ✅ Mesuré |
| 143 | aari1995/German_Semantic_STS_V2 | aari1995 | 🟢 Ouvert | 16,3 % | 17 novembre 2022 | ✅ Mesuré |
| 144 | Jaume/gemma-2b-embeddings | Jaume | 🟢 Ouvert | 16,2 % | 29 juin 2024 | ✅ Mesuré |
| 145 | Omartificial-Intelligence-Space/Arabic-mpnet-base-all-nli-triplet | Omartificial-Intelligence-Space | 🟢 Ouvert | 13,2 % | 15 juin 2024 | ✅ Mesuré |
| 146 | NeuML/pubmedbert-base-embeddings-100K | NeuML | 🟢 Ouvert | 12,5 % | 3 janvier 2025 | ✅ Mesuré |
| 147 | DeepPavlov/distilrubert-small-cased-conversational | DeepPavlov | 🟢 Ouvert | 12,0 % | 28 juin 2022 | ✅ Mesuré |
| 148 | cointegrated/rubert-tiny2 | cointegrated | 🟢 Ouvert | 11,4 % | 28 octobre 2021 | ✅ Mesuré |
| 149 | malenia1/ternary-weight-embedding | malenia1 | 🟢 Ouvert | 11,2 % | 23 octobre 2024 | ✅ Mesuré |
| 150 | Omartificial-Intelligence-Space/Arabert-all-nli-triplet-Matryoshka | Omartificial-Intelligence-Space | 🟢 Ouvert | 10,4 % | 16 juin 2024 | ✅ Mesuré |
| 151 | silma-ai/silma-embeddding-matryoshka-v0.1 | silma-ai | 🟢 Ouvert | 10,3 % | 12 octobre 2024 | ✅ Mesuré |
| 152 | DeepPavlov/rubert-base-cased-sentence | DeepPavlov | 🟢 Ouvert | 10,1 % | 4 mars 2020 | ✅ Mesuré |
| 153 | cointegrated/rubert-tiny | cointegrated | 🟢 Ouvert | 9,9 % | 24 mai 2021 | ✅ Mesuré |
| 154 | DeepPavlov/rubert-base-cased | DeepPavlov | 🟢 Ouvert | 9,8 % | 4 mars 2020 | ✅ Mesuré |
| 155 | deepvk/deberta-v1-base | deepvk | 🟢 Ouvert | 9,2 % | 7 février 2023 | ✅ Mesuré |
| 156 | Omartificial-Intelligence-Space/Marbert-all-nli-triplet-Matryoshka | Omartificial-Intelligence-Space | 🟢 Ouvert | 7,5 % | 17 juin 2024 | ✅ Mesuré |
| 157 | ai-forever/sbert_large_mt_nlu_ru | ai-forever | 🟢 Ouvert | 7,5 % | 18 mai 2021 | ✅ Mesuré |
| 158 | ai-forever/sbert_large_nlu_ru | ai-forever | 🟢 Ouvert | 6,9 % | 20 novembre 2020 | ✅ Mesuré |
Classement établi sur 158 modèles évalués, dont 4 de grands éditeurs. Score médian de l'ensemble : 39,2 %.
Notre analyse
Un score élevé traduit une meilleure aptitude à faire remonter les contenus médicaux pertinents, à les réordonner ou à les regrouper selon la tâche. L’interprétation dépend toutefois de la métrique employée, notamment nDCG@10 pour certaines évaluations, et de la diversité des jeux réunis dans l’agrégat. Dans la base, le score médian de 39 % et le meilleur résultat de 67 %, obtenu par Alibaba-NLP/gte-Qwen2-7B-instruct, montrent un écart important entre le niveau central et la tête du classement. Cet écart ne suggère pas une saturation générale du benchmark. La portée reste centrée sur la recherche d’information médicale, principalement en anglais, et ne représente donc pas toutes les capacités possibles d’un modèle. L’accès public impose aussi de considérer le risque de contamination lors de l’interprétation des performances. Enfin, la rigueur comparative est limitée par l’origine des résultats, majoritairement auto-déclarés par les éditeurs, plutôt que mesurés uniformément par une évaluation indépendante.
Sources des scores : mteb.