Team Ai
Datasetpublic

mteb/stsbenchmark-sts

STSBenchmark An MTEB dataset Massive Text Embedding Benchmark Semantic Textual Similarity Benchmark (STSbenchmark) dataset. Task category t2t Domains Blog, News, Written Reference https://github.com/PhilipMay/stsb-multi-mt/ How to evaluate on this task You can evaluate an embedding model on this dataset using the following code: import mteb task = mteb.get_tasks(["STSBenchmark"]) evaluator = mteb.MTEB(task) model = mteb.get_model(YOUR_MODEL)… See the full description on the dataset page: https://huggingface.co/datasets/mteb/stsbenchmark-sts.

sourceHugging Faceunknownupdated 8mo agoView on Hugging Face
19likes19kdownloads
filetest.jsonl.gz62 KBdownload
filetrain.jsonl.gz271 KBdownload
filevalidation.jsonl.gz84 KBdownload

mteb/stsbenchmark-sts · main · files are served by the source, never re-hosted here