rewirebio.iobenchmarks
Configuration

Qwen 2.5 72B, basic prompt, temperature 0.8 (default) (Lin et al. 2025)

Configuration as run in the cited comparison.

3 evaluations · 3 results

Overview

Configuration as run in the cited comparison.

Consult the linked sources for architecture or protocol details. Missing evidence is not evidence of a missing capability.

Evaluations and results

3 evaluations · 3 results. Different protocols are not a single leaderboard.

Filter evaluations

Applied filters: All linked evaluations

Exact evaluated configurations and original reported results
Tested configurationProtocol and datasetFindingEvidence and details
Configuration: Qwen 2.5 72B, basic prompt, temperature 0.8 (default) (Lin et al. 2025)Protocol: Lin et al. 2025 CIViC evidence-level assignment
Dataset: CIViC clinical evidence summary, 4,426 variant-disease associations (accessed 2024-11-20)
0.248 top-1-accuracy
fraction · higher

Uncertainty: 95% CI 0.2477 to 0.2492

Coverage: Not reported scored / Not reported eligible

Independent external evaluation · Source checked
Methods, coverage and source

qwen-basic on civic (Lin et al. 2025)

egfrnsclc-20261009-protocol-lin2025-civic-level-assignment

Aggregation: Not reported

Benchmarking large language models GPT-4o, llama 3.1, and qwen 2.5 for cancer genetic variant classification · Table 1 (Tab1), row 'CIViC', column 'Qwen 2.5' Mean accuracy; 95% CI in the next column
Configuration: Qwen 2.5 72B, basic prompt, temperature 0.8 (default) (Lin et al. 2025)Protocol: Lin et al. 2025 FoundationOne variants: clinically relevant vs VUS
Dataset: FoundationOne CDx report variants, 612 patients (Lin et al. 2025)
0.573 top-1-accuracy
fraction · higher

Uncertainty: 95% CI 0.5725 to 0.5736

Coverage: Not reported scored / Not reported eligible

Independent external evaluation · Source checked
Methods, coverage and source

qwen-basic on foundationone (Lin et al. 2025)

egfrnsclc-20261009-protocol-lin2025-foundationone-relevant-vs-vus

Aggregation: Not reported

Benchmarking large language models GPT-4o, llama 3.1, and qwen 2.5 for cancer genetic variant classification · Table 1 (Tab1), row 'Foundationone', column 'Qwen 2.5' Mean accuracy; 95% CI in the next column
Configuration: Qwen 2.5 72B, basic prompt, temperature 0.8 (default) (Lin et al. 2025)Protocol: Lin et al. 2025 OncoKB level-of-evidence assignment
Dataset: OncoKB actionable-genes table, 625 variant-cancer-drug associations (accessed 2024-11-20)
0.333 top-1-accuracy
fraction · higher

Uncertainty: 95% CI 0.3316 to 0.334

Coverage: Not reported scored / Not reported eligible

Independent external evaluation · Source checked
Methods, coverage and source

qwen-basic on oncokb (Lin et al. 2025)

egfrnsclc-20261009-protocol-lin2025-oncokb-level-assignment

Aggregation: Not reported

Benchmarking large language models GPT-4o, llama 3.1, and qwen 2.5 for cancer genetic variant classification · Table 1 (Tab1), row 'OncoKB', column 'Qwen 2.5' Mean accuracy; 95% CI in the next column

Source checking is not independent reproduction. Release 2026-10-10-7b8f80935f90.

Use this model

How it works, versions and access
Strengths, limitations and unresolved questions

Evidence

Source checking verifies the cited claim or transcription. It does not establish independent reproduction.

Evidence table

Inspect claims, sources and review details

Trace each statement to its source and review. A context-only reference supports the record generally; it does not verify an individual field. Source checking does not reproduce an experiment.

One row per statement and cited source. Multiple citations are not independent evaluations. Shared locators are labelled explicitly.

0 evidence rows matching the loaded filters

Claims, original sources and review scope · Release 2026-10-10-7b8f80935f90
Property and statementOriginal source and locationReview and provenance

No evidence rows match these filters. Choose another scope or clear the search.

Sources and history

Release 2026-10-10-7b8f80935f90 · Record review: source checked

1 source records and release historyDownload this release (gzip)
Technical metadata and extraction receipts

Stable ID: egfrnsclc-20261009-config-lin2025-qwen-basic

areas
dna-genomes
contexts
clinical_research
method types
foundation_model
reported name
Qwen2.5
foundation model eligible
true
source locator
Lin et al. 2025 Methods P39-P43; Table 2 'Model conditions'
parameters
basic prompt; temperature 0.8 (default)
missing metadata
version: reason: unreported; note: Model size is printed (70B or 72B); no checkpoint tag or quantisation is printed
Related records

Suggest a correction