Pre2Post-Chess Open-sourced models and datasets for training the chess reasoning models. pavelslab-nyu/pretrain_v1_54B Updated about 10 hours ago • 1.01k • 1 pavelslab-nyu/chess_puzzle_benchmark Viewer • Updated about 10 hours ago • 2.36k • 387 pre-to-post-olmo/Math-Models Text Generation • Updated about 17 hours ago pavelslab-nyu/chess_puzzle_training_datasets Viewer • Updated Jun 8 • 229k • 26
rlvr-weak-supervision Models from "When Can LLMs Learn to Reason with Weak Supervision?" — Llama-3.2-3B with continual pre-training and Thinking SFT. pavelslab-nyu/Llama-3.2-3B-ThinkSFT 3B • Updated Apr 20 • 419 pavelslab-nyu/Llama-3.2-3B-CPT-Math-ThinkSFT 3B • Updated Apr 20 • 66 pavelslab-nyu/Llama-3.2-3B-CPT-Math 3B • Updated Apr 20 • 2
Pre2Post-Chess Open-sourced models and datasets for training the chess reasoning models. pavelslab-nyu/pretrain_v1_54B Updated about 10 hours ago • 1.01k • 1 pavelslab-nyu/chess_puzzle_benchmark Viewer • Updated about 10 hours ago • 2.36k • 387 pre-to-post-olmo/Math-Models Text Generation • Updated about 17 hours ago pavelslab-nyu/chess_puzzle_training_datasets Viewer • Updated Jun 8 • 229k • 26
rlvr-weak-supervision Models from "When Can LLMs Learn to Reason with Weak Supervision?" — Llama-3.2-3B with continual pre-training and Thinking SFT. pavelslab-nyu/Llama-3.2-3B-ThinkSFT 3B • Updated Apr 20 • 419 pavelslab-nyu/Llama-3.2-3B-CPT-Math-ThinkSFT 3B • Updated Apr 20 • 66 pavelslab-nyu/Llama-3.2-3B-CPT-Math 3B • Updated Apr 20 • 2