Neural Solver Synthesis Collection Models, datasets, certified evaluations, and final evidence. Interactive reports: https://wandb.ai/neural-solver-synthesis/neural-solver-synthesis • 55 items • Updated 20 days ago
IDEALLab/Qwen2.5-Coder-14B-Instruct-GRPO-SDS-NoHypothesize-step90-seed303 Reinforcement Learning • 15B • Updated Sep 6 • 17
IDEALLab/Qwen2.5-Coder-14B-Instruct-GRPO-SDS-NoHypothesize-step90-seed202 Reinforcement Learning • 15B • Updated Sep 6 • 13
IDEALLab/Qwen2.5-Coder-14B-Instruct-GRPO-SDS-NoHypothesize-step90-seed101 Reinforcement Learning • 15B • Updated Sep 6 • 32
IDEALLab/Qwen2.5-Coder-14B-Instruct-GRPO-TSP-Hero-seed303 Reinforcement Learning • 15B • Updated Sep 6 • 11
IDEALLab/Qwen2.5-Coder-14B-Instruct-GRPO-SDS-NoHypothesize-step90-seed303 Reinforcement Learning • 15B • Updated Sep 6 • 17
IDEALLab/Qwen2.5-Coder-14B-Instruct-GRPO-TSP-Hero-seed202 Reinforcement Learning • 15B • Updated Sep 6 • 50
IDEALLab/Qwen2.5-Coder-14B-Instruct-GRPO-SDS-NoHypothesize-step90-seed202 Reinforcement Learning • 15B • Updated Sep 6 • 13
IDEALLab/Qwen2.5-Coder-14B-Instruct-GRPO-SDS-NoHypothesize-step90-seed101 Reinforcement Learning • 15B • Updated Sep 6 • 32
IDEALLab/Qwen2.5-Coder-14B-Instruct-GRPO-TSP-Hero-seed303 Reinforcement Learning • 15B • Updated Sep 6 • 11
IDEALLab/Qwen2.5-Coder-14B-Instruct-GRPO-TSP-Hero-seed202 Reinforcement Learning • 15B • Updated Sep 6 • 50