MDRSS · MARKDOWN SNAPSHOT
0.0/10

News

[  Read the Docs  ] 日本語 | 中文简体 | 中文繁體 --- Code and data for the following works: SWE-bench is a benchmark for evaluating large language models on real world software issues collected from GitHub. Given a codebase and an issue, a language model is tasked with generating a patch that resolves the described problem.

#llm-engineering / card #2249★ 0◌ 0snapshot 2026-08-04