CoolFace
20 results

qwq

PrimeIntellect /NuminaMath-QwQ-CoT-5M INTELLECT-MATH: Frontier Mathematical Reasoning through Better Initializations for Reinforcement Learning INTELLECT-MATH is a 7B parameter model optimized for mathematical reasoning. It was trained in two stages, an SFT stage, in which the model was fine-tuned on verified QwQ outputs, and an RL stage, in which the model was trained using the PRIME-RL recipe. We demonstrate that the quality of our SFT data can impact the performance and training speed of the RL stage: Due to its… See the full description on the dataset page: https://huggingface.co/datasets/PrimeIntellect/NuminaMath-QwQ-CoT-5M.text1M<n<10M63 likes1.5k downloads2y agoHugging Facejonathanyin /aime_1983_2023_qwq-32b_tracestabularn<1K0 likes1.2k downloads1y agoHugging Facejonathanyin /aime_1983_2023_qwq-32b_traces_16384tabularn<1K0 likes994 downloads1y agoHugging Facew3en2g /QwQ_InfInstruct_Gen_v0use QwQ 32b preview to generate response to answer the question from Infinity-Instruct gen texttext-generation1M<n<10M0 likes771 downloads10mo agoHugging Facemlfoundations-dev /openthoughts3_code_100k_annotated_QwQ-32B_sharegpt_eval_5554 mlfoundations-dev/openthoughts3_code_100k_annotated_QwQ-32B_sharegpt_eval_5554 Precomputed model outputs for evaluation. Evaluation Results Summary Metric AIME24 AMC23 MATH500 MMLUPro JEEBench GPQADiamond LiveCodeBench CodeElo CodeForces HLE HMMT AIME25 LiveCodeBenchv5 Accuracy 34.3 74.5 79.4 49.4 51.0 44.3 53.9 21.5 23.1 12.2 17.0 22.7 40.1 AIME24 Average Accuracy: 34.33% ± 1.89% Number of Runs: 10 Run Accuracy Questions… See the full description on the dataset page: https://huggingface.co/datasets/mlfoundations-dev/openthoughts3_code_100k_annotated_QwQ-32B_sharegpt_eval_5554.tabular10K<n<100K0 likes558 downloads1y agoHugging Facemlfoundations-dev /qwq_mix_qwen3_sciencetabular100K<n<1M1 likes520 downloads1y agoHugging Face