mathematic
Mistral-portuguese-luana-7b-Mathematics-i1-GGUFMistral-portuguese-luana-7b-Mathematics-GGUFstackexchange_mathematica-GGUFQwen2.5-1.5b_Fine_Tuned_For_Mathematical_Word_Problems-ExperimentalQwen2.5-1.5B-Advanced-Mathematics-and-Modeling-Distilled-8Clusters-GGUFstorywriter-mathematician-GGUFGemMaroc-27b-it-GGUFMistral-portuguese-luana-7b-Mathematics-GGUF
math-dataset-measuring-mathematical-problem-solvingTo cite the dataset please reference it as
@article{hendrycksmath2021,
title={Measuring Mathematical Problem Solving With the MATH Dataset},
author={Dan Hendrycks and Collin Burns and Saurav Kadavath and Akul Arora and Steven Basart and Eric Tang and Dawn Song and Jacob Steinhardt},
journal={NeurIPS},
year={2021}
}
Mathematical_Modeling_Speciale_Dataset_v0.1cqadupstack-mathematica
CQADupstackMathematicaRetrieval
An MTEB dataset
Massive Text Embedding Benchmark
CQADupStack: A Benchmark Data Set for Community Question-Answering Research
Task category
t2t
Domains
Written, Academic, Non-fiction
Referencehttp://nlp.cis.unimelb.edu.au/resources/cqadupstack/
How to evaluate on this task
You can evaluate an embedding model on this dataset using the following code:
import mteb
task = mteb.get_tasks(["CQADupstackMathematicaRetrieval"])… See the full description on the dataset page: https://huggingface.co/datasets/mteb/cqadupstack-mathematica.IndustryCorpus2_mathematics_statistics
IndustryCorpus2: Mathematics & Statistics
This repository contains the IndustryCorpus2: Mathematics & Statistics domain subset of BAAI/IndustryCorpus2.
Refer to the parent dataset card for data construction, intended use, limitations,
and licensing details.
Citation
If you use this dataset in your work, please cite IndustryCorpus2:
@misc{shi2024industrycorpus2,
title = {IndustryCorpus2},
author = {Xiaofeng Shi and Lulu Zhao and Hua Zhou and Donglin Hao}… See the full description on the dataset page: https://huggingface.co/datasets/BAAI/IndustryCorpus2_mathematics_statistics.hamela_books_text_full_okdsir-pile-13m-filtered-no-github-or-dm_mathematics
DSIR Pile 13M - Filtered Version
This is a filtered version of timaeus/dsir-pile-13m.
Filtering Applied:
Excluded: All rows where metadata.pile_set_name contains 'Github' or 'DM_mathematics'
Kept: All other rows from the original dataset
Dataset Size
Original: ~13M examples
Filtered: 12,782,200 examples (99.9% of original)
Uploaded in: 64 batch files
Usage
from datasets import load_dataset
dataset =… See the full description on the dataset page: https://huggingface.co/datasets/timaeus/dsir-pile-13m-filtered-no-github-or-dm_mathematics.
