agi-noobs/chess-sft-20k-llm-reasoning-enriched-dpo-hard-negatives-v1 Viewer • Updated Dec 30, 2025 • 1.57k • 71 • 5
ticoAg/llm-complex-reasoning-train-qwen2-72b-instruct-correct Viewer • Updated Aug 8, 2024 • 7.11k • 92 • 5
Yuhan123/vicuna-13b-self_consistency_random_var_3 Text Generation • 13B • Updated Mar 14, 2025 • 27 • 7
s-emanuilov/LLMBG-Llama-3.1-8B-BG-Reasoning-v0.1 Text Generation • 8B • Updated Feb 9, 2025 • 215 • 15
xTayyub/High-Quality-Synthetic-Python-Dataset-with-Reasoning-Traces-Chain-of-Thought-for-LLM-Fine-Tuning Updated Dec 9, 2025 • 262 • 11
Yuhan123/vicuna-13b-self_consistency_neg_exp_var_5 Text Generation • 13B • Updated Mar 14, 2025 • 127 • 8
Yuhan123/vicuna-13b-self_consistency_neg_exp_var_4 Text Generation • 13B • Updated Mar 14, 2025 • 128 • 7
open-llm-leaderboard-old/details_alexredna__Tukan-1.1B-Chat-reasoning-sft-COLA Updated Jan 22, 2024 • 290 • 4
latkes/self-consistency-correction-exp23-correlation-discrimination Viewer • Updated Apr 13 • 236 • 35 • 5