GUE TRANSCRIPTION-FACTOR-PREDICTION-MOUSE-0: Transcription factor prediction (mouse), dataset 0
Transcription factor prediction (mouse), dataset 0. Scored with MCC on GUE Transcription factor prediction (mouse), 0. Fine-tuned on the GUE training split, scored on its test split. Split sizes are in Table 12.
Overview
Transcription factor prediction (mouse), dataset 0. Scored with MCC on GUE Transcription factor prediction (mouse), 0. Fine-tuned on the GUE training split, scored on its test split. Split sizes are in Table 12.
Consult the linked sources for architecture or protocol details. Missing evidence is not evidence of a missing capability.
Results
Each comparison retains its reviewed evaluation scope, dataset and metric. Results are shown without a pooled ranking.
GUE TRANSCRIPTION-FACTOR-PREDICTION-MOUSE-0: Transcription factor prediction (mouse), dataset 0
mcc (percent) · Higher values are better.
GUE TRANSCRIPTION-FACTOR-PREDICTION-MOUSE-0: Transcription factor prediction (mouse), dataset 0 · GUE Transcription factor prediction (mouse), 0 (GUE split)
Evidence origin: Author-reported evaluation.
DNABERT-2: Efficient Foundation Model and Benchmark for Multi-Species Genomes · Table 12, row(Transcription factor prediction (mouse))- Scores are MCC, except Covid variant classification which is F1, both on a 0 to 100 scale.
- The diamond entry is DNABERT-2 with further pre-training on the GUE training sets, so it is not directly comparable to the others.
Comparison details and limitations
Every method GUE reports on Transcription factor prediction (mouse), dataset 0, scored with MCC on GUE Transcription factor prediction (mouse), 0.
- Author-reported numbers, source checked but not independently reproduced.
Automated source review: 2026-09-18. Numerical source review does not establish independent reproduction.
Dots show point estimates. Whiskers show only explicitly defined uncertainty (standard deviation, standard error or a labelled interval); their definitions remain in Table. Unresolved uncertainty is not plotted. Differences do not establish statistical significance.
Showing 10 of 10 matching rows.
Methods and evaluation design
Procedure, tasks and evaluated configurations
Evaluation design
Benchmarks bring together tasks and protocols. A task describes the biological question; a protocol defines a particular test.
Benchmarks
These source-backed links do not make different protocols or scores interchangeable.
Recorded evaluations
Each evaluation records what was tested and under which conditions.
- DNABERT-2 (further pre-trained on GUE) on GUE TRANSCRIPTION-FACTOR-PREDICTION-MOUSE-0: Transcription factor prediction (mouse), dataset 0
- DNABERT-2 on GUE TRANSCRIPTION-FACTOR-PREDICTION-MOUSE-0: Transcription factor prediction (mouse), dataset 0
- DNABERT (3-mer) on GUE TRANSCRIPTION-FACTOR-PREDICTION-MOUSE-0: Transcription factor prediction (mouse), dataset 0
- DNABERT (4-mer) on GUE TRANSCRIPTION-FACTOR-PREDICTION-MOUSE-0: Transcription factor prediction (mouse), dataset 0
- DNABERT (5-mer) on GUE TRANSCRIPTION-FACTOR-PREDICTION-MOUSE-0: Transcription factor prediction (mouse), dataset 0
- DNABERT (6-mer) on GUE TRANSCRIPTION-FACTOR-PREDICTION-MOUSE-0: Transcription factor prediction (mouse), dataset 0
- NT-2500M-1000g on GUE TRANSCRIPTION-FACTOR-PREDICTION-MOUSE-0: Transcription factor prediction (mouse), dataset 0
- NT-2500M-multi on GUE TRANSCRIPTION-FACTOR-PREDICTION-MOUSE-0: Transcription factor prediction (mouse), dataset 0
- NT-500M-1000g on GUE TRANSCRIPTION-FACTOR-PREDICTION-MOUSE-0: Transcription factor prediction (mouse), dataset 0
- NT-500M-human on GUE TRANSCRIPTION-FACTOR-PREDICTION-MOUSE-0: Transcription factor prediction (mouse), dataset 0
Run instructions
No runnable recipe has been reviewed for this task. Dataset access, model requirements, licences and compute requirements must be checked against its sources before execution.
A task describes a biological question. Choose a linked protocol to obtain concrete split and scoring instructions.
Strengths, limitations and unresolved questions
Evidence
Source checking verifies the cited claim or transcription. It does not establish independent reproduction.
Evidence table
Inspect claims, sources and review details
Trace each statement to its source and review. A context-only reference supports the record generally; it does not verify an individual field. Source checking does not reproduce an experiment.
One row per statement and cited source. Multiple citations are not independent evaluations. Shared locators are labelled explicitly.
1 evidence row matching the loaded filters
| Property and statement | Original source and location | Review and provenance |
|---|---|---|
| Relationship: part of discovery-benchmark-gue Individual claims | DNABERT-2: Efficient Foundation Model and Benchmark for Multi-Species Genomes Table 12, row(Transcription factor prediction (mouse)) Version: Primary full-text snapshot retrieved 2026-09-17; exact bytes pinned by SHA-256 | source checked automated source review · 2026-09-18 Audit detailsPrimary-source transcription with no human sign-off and no independent reproduction. Field: Claim: gue-association-transcription-factor-prediction-mouse-0 Source artifact SHA-256: Hash scope: Exact retrieved primary paper artifact bytes. |
Sources and history
View linked audit checks and correction history
Release 2026-09-29-06401fd5b220 · Record review: source checked
1 source records and release history
- DNABERT-2: Efficient Foundation Model and Benchmark for Multi-Species Genomes · Original source · Primary full-text snapshot retrieved 2026-09-17; exact bytes pinned by SHA-256
Technical metadata and extraction receipts
Stable ID: gue-task-transcription-factor-prediction-mouse-0
- areas
- dna-genomes
- tasks
- Transcription factor prediction (mouse), dataset 0
- metric
- MCC
- metric direction
- higher
- dataset
- GUE Transcription factor prediction (mouse), 0
- protocol
- Fine-tuned on the GUE training split, scored on its test split. Split sizes are in Table 12.
- source locator
- Table 12, row(Transcription factor prediction (mouse))
- comparison panels
- id: gue-panel-transcription-factor-prediction-mouse-0; title: GUE TRANSCRIPTION-FACTOR-PREDICTION-MOUSE-0: Transcription factor prediction (mouse), dataset 0; protocol id: gue-task-transcription-factor-prediction-mouse-0; dataset id: gue-dataset-gue-transcription-factor-prediction-mouse-0; metric: mcc; unit: percent; direction: higher; result ids: gue-result-dnabert-3-mer-transcription-factor-prediction-mouse-0-mcc; gue-result-dnabert-4-mer-transcription-factor-prediction-mouse-0-mcc; gue-result-dnabert-5-mer-transcription-factor-prediction-mouse-0-mcc; gue-result-dnabert-6-mer-transcription-factor-prediction-mouse-0-mcc; gue-result-nt-500m-human-transcription-factor-prediction-mouse-0-mcc; gue-result-nt-500m-1000g-transcription-factor-prediction-mouse-0-mcc; gue-result-nt-2500m-1000g-transcription-factor-prediction-mouse-0-mcc; gue-result-nt-2500m-multi-transcription-factor-prediction-mouse-0-mcc; gue-result-dnabert-2-transcription-factor-prediction-mouse-0-mcc; gue-result-dnabert-2-further-pre-trained-on-gue-transcription-factor-prediction-mouse-0-mcc; source ids: evidence-expansion-gue-49300ace; source locator: Table 12, row(Transcription factor prediction (mouse)); context: Every method GUE reports on Transcription factor prediction (mouse), dataset 0, scored with MCC on GUE Transcription factor prediction (mouse), 0.; caveats: Author-reported numbers, source checked but not independently reproduced.; Scores are MCC, except Covid variant classification which is F1, both on a 0 to 100 scale.; The diamond entry is DNABERT-2 with further pre-training on the GUE training sets, so it is not directly comparable to the others.; review: method: automated_source_review; date: 2026-09-18
Related records
- part of: GUE
- subject: GUE TRANSCRIPTION-FACTOR-PREDICTION-MOUSE-0: part of discovery-benchmark-gue
- benchmark: DNABERT-2 (further pre-trained on GUE) on GUE TRANSCRIPTION-FACTOR-PREDICTION-MOUSE-0: Transcription factor prediction (mouse), dataset 0
- benchmark: DNABERT-2 on GUE TRANSCRIPTION-FACTOR-PREDICTION-MOUSE-0: Transcription factor prediction (mouse), dataset 0
- benchmark: DNABERT (3-mer) on GUE TRANSCRIPTION-FACTOR-PREDICTION-MOUSE-0: Transcription factor prediction (mouse), dataset 0
- benchmark: DNABERT (4-mer) on GUE TRANSCRIPTION-FACTOR-PREDICTION-MOUSE-0: Transcription factor prediction (mouse), dataset 0
- benchmark: DNABERT (5-mer) on GUE TRANSCRIPTION-FACTOR-PREDICTION-MOUSE-0: Transcription factor prediction (mouse), dataset 0
- benchmark: DNABERT (6-mer) on GUE TRANSCRIPTION-FACTOR-PREDICTION-MOUSE-0: Transcription factor prediction (mouse), dataset 0
- benchmark: NT-2500M-1000g on GUE TRANSCRIPTION-FACTOR-PREDICTION-MOUSE-0: Transcription factor prediction (mouse), dataset 0
- benchmark: NT-2500M-multi on GUE TRANSCRIPTION-FACTOR-PREDICTION-MOUSE-0: Transcription factor prediction (mouse), dataset 0
- benchmark: NT-500M-1000g on GUE TRANSCRIPTION-FACTOR-PREDICTION-MOUSE-0: Transcription factor prediction (mouse), dataset 0
- benchmark: NT-500M-human on GUE TRANSCRIPTION-FACTOR-PREDICTION-MOUSE-0: Transcription factor prediction (mouse), dataset 0