Skip to content

Pull requests: open-compass/opencompass

Author
Filter by author
Loading
Label
Filter by label
Loading
Use alt + click/return to exclude labels
or + click/return for logical OR
Projects
Filter by project
Loading
Milestones
Filter by milestone
Loading
Reviews
Assignee
Filter by who’s assigned
Assigned to nobody Loading
Sort

Pull requests list

Fix BABILong 2k config syntax
#2578 opened Aug 3, 2026 by Danielxu0208 Draft
Add MedFailBench dataset adapter
#2560 opened Jul 18, 2026 by goktugozkanmd Loading…
Add robust HumanEval postprocess for chat outputs
#2515 opened Jul 8, 2026 by Ding-god Loading…
3 of 6 tasks
[Fix] Extract the final answer in GPQA simple-eval predictions
#2496 opened Jun 27, 2026 by Hibbert133 Contributor Loading…
4 of 6 tasks
[Fix] Combine split eval results in default summarizer
#2451 opened May 15, 2026 by yhzhu99 Contributor Loading…
feat: upgrade MiniMax default model to M3
#2418 opened Mar 20, 2026 by octo-patch Contributor Loading…
3 tasks done
[Fix] CEval ModelScope load and HF generate for causal LMs
#2416 opened Mar 19, 2026 by DeliWang Loading…
6 tasks
Add support for Azure OpenAI models and managed identity auth
#2415 opened Mar 18, 2026 by jgbradley1 Loading…
2 of 6 tasks
[Fix] Fix eval stage is extremely slow
#2409 opened Mar 5, 2026 by xming521 Loading…
6 tasks
ProTip! Mix and match filters to narrow down what you’re looking for.