Project: Turkish Embeddings from Scratch and CPT Decoders
Infrastructure: MareNostrum 5 (BSC)
AI & ML interests
Where data finds its mind
Recent Activity
View all activity
Papers
Mecellem Models: Turkish Models Trained from Scratch and Continually Pre-trained for the Legal Domain
TurkEmbed: Turkish Embedding Model on NLI & STS Tasks
models 66
newmindai/Mecellem-Qwen3-1.7B-TR
Text Generation • Updated • 47 • 4
newmindai/Mursit-Large
Feature Extraction • Updated • 112 • 4
newmindai/Mursit-Base
Feature Extraction • Updated • 2 • 3
newmindai/Muhakim
Text Generation • Updated • 8 • 4
newmindai/Mecellem-Qwen3-4B-TR
Text Generation • 4B • Updated • 63 • 3
newmindai/Mursit-Large-TR-Retrieval
Sentence Similarity • 0.4B • Updated • 368 • 6
newmindai/Mursit-Embed-Qwen3-1.7B-TR
Sentence Similarity • 2B • Updated • 100 • 3
newmindai/Mursit-Base-TR-Retrieval
Sentence Similarity • 0.2B • Updated • 627 • 4
newmindai/Mursit-Embed-Qwen3-4B-TR
Sentence Similarity • 4B • Updated • 4 • 2
newmindai/bge-m3-stsb
Sentence Similarity • 0.6B • Updated • 18 • 3
datasets 9
newmindai/contract-retrieval
Viewer • Updated • 816 • 81 • 2
newmindai/regulation-retrieval
Viewer • Updated • 264k • 127 • 2
newmindai/caselaw-retrieval
Viewer • Updated • 4.15k • 118 • 3
newmindai/ms-marco-turkish-triplets
Viewer • Updated • 920k • 23
newmindai/stsb-deepl-tr
Viewer • Updated • 8.63k • 19
newmindai/EuroHPC-Legal
Viewer • Updated • 43k • 75 • 1
newmindai/RAGTruth-TR
Viewer • Updated • 17.8k • 124 • 6
newmindai/siu-rag-data
Viewer • Updated • 507 • 19 • 2
newmindai/mezura-eval-data
Viewer • Updated • 650 • 18 • 1