SWE-rebench nebius
Dataset Summary SWE-rebench is a large-scale dataset designed to support training and evaluation of LLM-based software engineering (SWE) agents, building upon and expanding our earlier release, SWE-bench-extra. It is constructed using a fully automated pipeline that continuously extracts real-world interactive SWE tasks from GitHub repositories at scale, as detailed in our paper SWE-rebench: An Automated Pipeline for Task Collection and Decontaminated Evaluation of Software… See the full description on the dataset page: https://huggingface.co/datasets/nebius/SWE-rebench.
- Type
- dataset
- License
- cc-by-4.0
- Downloads
- 382,468
- Likes
- 71
- Access
- public
- Files
- 0
Tags
- 代码数据
- 评测基准
- Agent
- 代码生成
- 大语言模型
- 数据处理
- 深度学习
README
--- license: cc-by-4.0 task_categories: - other library_name: datasets dataset_info: features: - name: instance_id dtype: string - name: base_commit dtype: string - name: created_at dtype: string - name: environment_setup_commit dtype: string - name: hints_text dtype: string - name: patch dtype: st…