Articles | Volume 30, issue 18
https://doi.org/10.5194/hess-30-5857-2026
https://doi.org/10.5194/hess-30-5857-2026
Research article
 | 
16 Sep 2026
Research article |  | 16 Sep 2026

Setting the bar: benchmarks for model performances in large-sample hydrology

Jan Seibert, Marc Vis, and Sandra Pool

Download

Interactive discussion

Status: closed

Comment types: AC – author | RC – referee | CC – community | EC – editor | CEC – chief editor | : Report abuse
  • RC1: 'Very valuable, but needs to provide details on setup and performance values', Benedikt Heudorfer, 19 Jun 2026
    • AC1: 'Quick Reply on link comment in RC1', Jan Seibert, 22 Jun 2026
    • AC2: 'Reply on RC1', Jan Seibert, 08 Jul 2026
  • RC2: 'Comment on egusphere-2026-3272', Tam Nguyen, 24 Jun 2026
    • AC3: 'Reply on RC2', Jan Seibert, 08 Jul 2026

Peer review completion

AR – Author's response | RR – Referee report | ED – Editor decision | EF – Editorial file upload
ED: Publish subject to technical corrections (30 Jul 2026) by Ralf Loritz
AR by Jan Seibert on behalf of the Authors (05 Sep 2026)  Author's response   Manuscript 
Download
Short summary
We studied how well simple bucket-type models can reproduce observed river flow using large data sets from many regions in the world. Model performance varies widely depending on local conditions, so fixed performance thresholds are misleading. To better judge model performances, we propose lower and upper benchmarks. These benchmarks help to better understand what level of model performance is achievable and, thus, enable us to compare models across different catchments.
Share