Constrained Diffusion Tasks Datasets used in the paper "Constrained Decoding of Diffusion LLMs for Context-Free Grammars". eth-sri/json-mode-eval-extended Viewer • Updated Nov 15, 2025 • 272 • 1.4k eth-sri/smiles-eval Viewer • Updated Aug 15, 2025 • 167 • 1k zai-org/humaneval-x Updated Oct 25, 2022 • 2.1k • 93 eth-sri/HumanEval-MRI-Cpp Viewer • Updated Aug 15, 2025 • 473 • 29 • 1
SWT-Bench Variations of the SWT-Bench pre-formatted dataset used for the paper "SWT-Bench: Testing and Validating Real-World Bug-Fixes with Code Agents" eth-sri/SWT-bench_bm25_27k_zsb Viewer • Updated Feb 25, 2025 • 2.2k • 85 eth-sri/SWT-bench_Lite_bm25_27k_zsb Viewer • Updated Feb 25, 2025 • 299 • 31 eth-sri/SWT-bench_Verified_bm25_27k_zsp Viewer • Updated Feb 25, 2025 • 433 • 219 eth-sri/SWT-bench_bm25_27k_zsp Viewer • Updated Feb 25, 2025 • 2.52k • 31 • 2
Constrained Diffusion Tasks Datasets used in the paper "Constrained Decoding of Diffusion LLMs for Context-Free Grammars". eth-sri/json-mode-eval-extended Viewer • Updated Nov 15, 2025 • 272 • 1.4k eth-sri/smiles-eval Viewer • Updated Aug 15, 2025 • 167 • 1k zai-org/humaneval-x Updated Oct 25, 2022 • 2.1k • 93 eth-sri/HumanEval-MRI-Cpp Viewer • Updated Aug 15, 2025 • 473 • 29 • 1
SWT-Bench Variations of the SWT-Bench pre-formatted dataset used for the paper "SWT-Bench: Testing and Validating Real-World Bug-Fixes with Code Agents" eth-sri/SWT-bench_bm25_27k_zsb Viewer • Updated Feb 25, 2025 • 2.2k • 85 eth-sri/SWT-bench_Lite_bm25_27k_zsb Viewer • Updated Feb 25, 2025 • 299 • 31 eth-sri/SWT-bench_Verified_bm25_27k_zsp Viewer • Updated Feb 25, 2025 • 433 • 219 eth-sri/SWT-bench_bm25_27k_zsp Viewer • Updated Feb 25, 2025 • 2.52k • 31 • 2