Datasets
Individual model/measurement problems in PEtab format, with problem-specific observables and noise assumptions.
The PEtab collection supports evaluation of computational methods for fitting mathematical models to observations.
No reviewed evaluations are linked here in this release. See the sources and separately identified configurations below.
Individual model/measurement problems in PEtab format, with problem-specific observables and noise assumptions.
Objectives depend on each PEtab measurement/noise model and the selected optimization-performance criterion; there is no universal prediction metric.
inapplicable
PEtab model, parameter, condition, observable and measurement tables.
Conceptual procedure. Task variants and protocol versions retain their separate scoring conditions.
Source reviewed · Automated source review, 2026-09-16. All specifications and missing details
0 evaluations · 0 results. Different protocols are not a single leaderboard.
Applied filters: All linked evaluations
No evaluations linked in this release.
Source checking is not independent reproduction. Release 2026-09-29-06401fd5b220.
The collection packages dynamical biological models with the experimental measurements and observation/noise assumptions needed for parameter estimation. A benchmark run specifies a model, solver and inference procedure against this fixed problem. Calibration fit, computational reliability and predictive validation are different assessment targets.
Reference methods help show what a model adds beyond simple controls. We track a null control and a conventional method for each protocol.
No concrete protocols are explicitly linked to this suite. Protocol identification and baseline selection are outstanding.
Protocol coverage CSV · Model evaluation matrix · Source table · Release and checksums
Coverage is derived from release 2026-09-29-06401fd5b220. Source citations describe the original records; they do not validate an unreviewed baseline proposal. No results have been generated by this audit.
The repository provides PEtab problems and links INSTALL.md. It is a benchmark-problem collection rather than a prescribed optimizer run; choose a model, solver, initialization and budget separately. Source-model/data licenses can differ from the repository license.
A maintained rewire runner has not been verified for this benchmark. Check data access, weights, licences, dependencies and hardware in the linked official documentation; requirements have not been fully extracted.
Benchmarking-Initiative/Benchmark-Models-PEtab / README.md · README.md lines 7–12 and 54–68 (Overview, license, installation and version citation)Primary paper and/or task implementation reviewed for the explicitly cited methodology claims. Scope-limited absence is recorded only after the documented source search; no model runs or independent reproduction.
Stable record: discovery-benchmark-petab-benchmark-collectionExplanatory profile: source reviewed · Automated source review, 2026-09-16. Review applies to the cited claims; unresolved fields are listed below. Numerical results retain their own review status.
| Property | Description and evidence |
|---|---|
| Datasets | Individual model/measurement problems in PEtab format, with problem-specific observables and noise assumptions.SourcesBenchmarking-Initiative/Benchmark-Models-PEtab official source · Pinned README: collection description; benchmark problem table |
| Splits | A supervised train/test split is not intrinsic to the parameter-estimation problem collection; the chosen study must define any held-out observations. · Not applicableSourcesBenchmarking-Initiative/Benchmark-Models-PEtab official source · Pinned README: collection description; benchmark problem table |
| Metrics | Objectives depend on each PEtab measurement/noise model and the selected optimization-performance criterion; there is no universal prediction metric. · Not applicableSourcesBenchmarking-Initiative/Benchmark-Models-PEtab official source · Pinned README: collection description; benchmark problem table |
| Baselines | The collection provides common problem definitions for comparing modeling/estimation methods, not a fixed universal baseline.SourcesBenchmarking-Initiative/Benchmark-Models-PEtab official source · Pinned README: collection description; benchmark problem table |
| Leakage controls | The collection provides mechanistic models with calibration data, observation functions and experimental conditions. It is intended to compare numerical inference methods; a predictive train/test exclusion policy must be defined by the study using each model. · Not applicableSourcespetab primary benchmark evidence · Sections 2.2–2.4: observations, noise models and experimental conditions |
| Uncertainty | Measurement errors can be fixed from experiments or estimated jointly through explicit noise models. Parameter uncertainty and optimizer variability are different quantities and require their own evaluation procedure.Sourcespetab primary benchmark evidence · Sections 2.2–2.4: observations, noise models and experimental conditions |
| Entity type | Collection of parameter-estimation benchmark problems.SourcesBenchmarking-Initiative/Benchmark-Models-PEtab official source · Pinned README: collection description; benchmark problem table |
| Organisms | Organism identity is problem-specific; the collection spans distinct systems. · Not applicableSourcesBenchmarking-Initiative/Benchmark-Models-PEtab official source · Pinned README: collection description; benchmark problem table |
| Assays | Problem-specific observations with explicit measurement/noise models.SourcesBenchmarking-Initiative/Benchmark-Models-PEtab official source · Pinned README: collection description; benchmark problem table |
| Allowed inputs | PEtab model, parameter, condition, observable and measurement tables.SourcesBenchmarking-Initiative/Benchmark-Models-PEtab official source · Pinned README: collection description; benchmark problem table |
| Adaptation | Numerical parameter estimation against supplied observations; optimizer settings define the tested method.SourcesBenchmarking-Initiative/Benchmark-Models-PEtab official source · Pinned README: collection description; benchmark problem table |
Source checking verifies the cited claim or transcription. It does not establish independent reproduction.
Last literature check: 2026-09-17. Primary-source discovery and table/protocol screening; source checked is not independently reproduced. Raw acquisitions not automatically numerical publication approval.
| Paper or primary resource | Version | Reference |
|---|---|---|
| Benchmark problems for dynamic modeling of intracellular processes | PMC6735869 | Read source DOI: 10.1093/bioinformatics/btz020 |
source found structured extraction pending
Trace each statement to its source and review. A context-only reference supports the record generally; it does not verify an individual field. Source checking does not reproduce an experiment.
One row per statement and cited source. Multiple citations are not independent evaluations. Shared locators are labelled explicitly.
18 evidence rows matching the loaded filters
| Property and statement | Original source and location | Review and provenance |
|---|---|---|
| Diagram caption Conceptual procedure. Task variants and protocol versions retain their separate scoring conditions. Individual claims | Benchmarking-Initiative/Benchmark-Models-PEtab official source Pinned README: collection description; benchmark problem table Version: ddaa86d13f708926c57ec8918ce75a6b50e2e562 | source checked automated source review · 2026-09-16 Audit detailsPrimary paper and/or task implementation reviewed for the explicitly cited methodology claims. Scope-limited absence is recorded only after the documented source search; no model runs or independent reproduction. Field: Source artifact SHA-256: Hash scope: Hash scope not separately documented; inspect source record |
Diagram steps
| Benchmarking-Initiative/Benchmark-Models-PEtab official source Pinned README: collection description; benchmark problem table Version: ddaa86d13f708926c57ec8918ce75a6b50e2e562 | source checked automated source review · 2026-09-16 Audit detailsPrimary paper and/or task implementation reviewed for the explicitly cited methodology claims. Scope-limited absence is recorded only after the documented source search; no model runs or independent reproduction. Field: Source artifact SHA-256: Hash scope: Hash scope not separately documented; inspect source record |
| Diagram title Evaluation procedure Individual claims | Benchmarking-Initiative/Benchmark-Models-PEtab official source Pinned README: collection description; benchmark problem table Version: ddaa86d13f708926c57ec8918ce75a6b50e2e562 | source checked automated source review · 2026-09-16 Audit detailsPrimary paper and/or task implementation reviewed for the explicitly cited methodology claims. Scope-limited absence is recorded only after the documented source search; no model runs or independent reproduction. Field: Source artifact SHA-256: Hash scope: Hash scope not separately documented; inspect source record |
| Datasets Individual model/measurement problems in PEtab format, with problem-specific observables and noise assumptions. Individual claims | Benchmarking-Initiative/Benchmark-Models-PEtab official source Pinned README: collection description; benchmark problem table Version: ddaa86d13f708926c57ec8918ce75a6b50e2e562 | source checked automated source review · 2026-09-16 Audit detailsPrimary paper and/or task implementation reviewed for the explicitly cited methodology claims. Scope-limited absence is recorded only after the documented source search; no model runs or independent reproduction. Field: Source artifact SHA-256: Hash scope: Hash scope not separately documented; inspect source record |
| Splits A supervised train/test split is not intrinsic to the parameter-estimation problem collection; the chosen study must define any held-out observations. Individual claims | Benchmarking-Initiative/Benchmark-Models-PEtab official source Pinned README: collection description; benchmark problem table Version: ddaa86d13f708926c57ec8918ce75a6b50e2e562 | inapplicable automated source review · 2026-09-16 Audit detailsPrimary paper and/or task implementation reviewed for the explicitly cited methodology claims. Scope-limited absence is recorded only after the documented source search; no model runs or independent reproduction. Field: Source artifact SHA-256: Hash scope: Hash scope not separately documented; inspect source record |
| Adaptation Numerical parameter estimation against supplied observations; optimizer settings define the tested method. Individual claims | Benchmarking-Initiative/Benchmark-Models-PEtab official source Pinned README: collection description; benchmark problem table Version: ddaa86d13f708926c57ec8918ce75a6b50e2e562 | source checked automated source review · 2026-09-16 Audit detailsPrimary paper and/or task implementation reviewed for the explicitly cited methodology claims. Scope-limited absence is recorded only after the documented source search; no model runs or independent reproduction. Field: Source artifact SHA-256: Hash scope: Hash scope not separately documented; inspect source record |
| Metrics Objectives depend on each PEtab measurement/noise model and the selected optimization-performance criterion; there is no universal prediction metric. Individual claims | Benchmarking-Initiative/Benchmark-Models-PEtab official source Pinned README: collection description; benchmark problem table Version: ddaa86d13f708926c57ec8918ce75a6b50e2e562 | inapplicable automated source review · 2026-09-16 Audit detailsPrimary paper and/or task implementation reviewed for the explicitly cited methodology claims. Scope-limited absence is recorded only after the documented source search; no model runs or independent reproduction. Field: Source artifact SHA-256: Hash scope: Hash scope not separately documented; inspect source record |
| Baselines The collection provides common problem definitions for comparing modeling/estimation methods, not a fixed universal baseline. Individual claims | Benchmarking-Initiative/Benchmark-Models-PEtab official source Pinned README: collection description; benchmark problem table Version: ddaa86d13f708926c57ec8918ce75a6b50e2e562 | source checked automated source review · 2026-09-16 Audit detailsPrimary paper and/or task implementation reviewed for the explicitly cited methodology claims. Scope-limited absence is recorded only after the documented source search; no model runs or independent reproduction. Field: Source artifact SHA-256: Hash scope: Hash scope not separately documented; inspect source record |
| Leakage controls The collection provides mechanistic models with calibration data, observation functions and experimental conditions. It is intended to compare numerical inference methods; a predictive train/test exclusion policy must be defined by the study using each model. Individual claims | petab primary benchmark evidence Sections 2.2–2.4: observations, noise models and experimental conditions Version: PMC6735869 | inapplicable automated source review · 2026-09-16 Audit detailsPrimary paper and/or task implementation reviewed for the explicitly cited methodology claims. Scope-limited absence is recorded only after the documented source search; no model runs or independent reproduction. Field: Source artifact SHA-256: Hash scope: Hash scope not separately documented; inspect source record |
| Uncertainty Measurement errors can be fixed from experiments or estimated jointly through explicit noise models. Parameter uncertainty and optimizer variability are different quantities and require their own evaluation procedure. Individual claims | petab primary benchmark evidence Sections 2.2–2.4: observations, noise models and experimental conditions Version: PMC6735869 | source checked automated source review · 2026-09-16 Audit detailsPrimary paper and/or task implementation reviewed for the explicitly cited methodology claims. Scope-limited absence is recorded only after the documented source search; no model runs or independent reproduction. Field: Source artifact SHA-256: Hash scope: Hash scope not separately documented; inspect source record |
View linked audit checks and correction history
Release 2026-09-29-06401fd5b220 · Record review: discovered
Stable ID: discovery-benchmark-petab-benchmark-collection