BDA-Rand, llama-3-1-8b backbone (Gupta et al. 2025)
Configuration as run in the cited comparison.
Overview
Configuration as run in the cited comparison.
Consult the linked sources for architecture or protocol details. Missing evidence is not evidence of a missing capability.
Evaluations and results
5 evaluations · 5 results. Different protocols are not a single leaderboard.
Filter evaluations
Applied filters: All linked evaluations
| Tested configuration | Protocol and dataset | Finding | Evidence and details |
|---|---|---|---|
| Configuration: BDA-Rand, llama-3-1-8b backbone (Gupta et al. 2025) | Protocol: Independent replication of 1-gene perturbation design: cumulative hits after 5 rounds of 128 genes Dataset: Carnevale et al. 2022 screen, T-cell resistance to tumour-microenvironment inhibitory signals | 31.6 true-positive-count count · higher Uncertainty: Not reported by the source Coverage: Not reported scored / Not reported eligible | Independent external evaluation · Source checkedMethods, coverage and sourceBDA-Rand (Llama-3.1-8B backbone) on Carnevale (Gupta et al. 2025) tgtval-20261009-protocol-gupta2025-cumulative-hits-round5 Aggregation: Not reported LLMs for Bayesian Optimization in Scientific Domains: Are We There Yet? · Table 1, 'Llama-3.1-8B backbone' block, row 'BDA-Rand', column 'Carnevale' |
| Configuration: BDA-Rand, llama-3-1-8b backbone (Gupta et al. 2025) | Protocol: Independent replication of 1-gene perturbation design: cumulative hits after 5 rounds of 128 genes Dataset: Sanchez et al. 2021 screen, decreased endogenous tau protein level in neurons | 45 true-positive-count count · higher Uncertainty: Not reported by the source Coverage: Not reported scored / Not reported eligible | Independent external evaluation · Source checkedMethods, coverage and sourceBDA-Rand (Llama-3.1-8B backbone) on Sanchez Down (Gupta et al. 2025) tgtval-20261009-protocol-gupta2025-cumulative-hits-round5 Aggregation: Not reported LLMs for Bayesian Optimization in Scientific Domains: Are We There Yet? · Table 1, 'Llama-3.1-8B backbone' block, row 'BDA-Rand', column 'Sanchez Down' |
| Configuration: BDA-Rand, llama-3-1-8b backbone (Gupta et al. 2025) | Protocol: Independent replication of 1-gene perturbation design: cumulative hits after 5 rounds of 128 genes Dataset: Sanchez et al. 2021 screen, endogenous tau protein level in neurons | 30.8 true-positive-count count · higher Uncertainty: Not reported by the source Coverage: Not reported scored / Not reported eligible | Independent external evaluation · Source checkedMethods, coverage and sourceBDA-Rand (Llama-3.1-8B backbone) on Sanchez (Gupta et al. 2025) tgtval-20261009-protocol-gupta2025-cumulative-hits-round5 Aggregation: Not reported LLMs for Bayesian Optimization in Scientific Domains: Are We There Yet? · Table 1, 'Llama-3.1-8B backbone' block, row 'BDA-Rand', column 'Sanchez' |
| Configuration: BDA-Rand, llama-3-1-8b backbone (Gupta et al. 2025) | Protocol: Independent replication of 1-gene perturbation design: cumulative hits after 5 rounds of 128 genes Dataset: Schmidt et al. 2022 screen, interferon-gamma production in primary human T cells | 51 true-positive-count count · higher Uncertainty: Not reported by the source Coverage: Not reported scored / Not reported eligible | Independent external evaluation · Source checkedMethods, coverage and sourceBDA-Rand (Llama-3.1-8B backbone) on IFNG (Gupta et al. 2025) tgtval-20261009-protocol-gupta2025-cumulative-hits-round5 Aggregation: Not reported LLMs for Bayesian Optimization in Scientific Domains: Are We There Yet? · Table 1, 'Llama-3.1-8B backbone' block, row 'BDA-Rand', column 'IFNG' |
| Configuration: BDA-Rand, llama-3-1-8b backbone (Gupta et al. 2025) | Protocol: Independent replication of 1-gene perturbation design: cumulative hits after 5 rounds of 128 genes Dataset: Schmidt et al. 2022 screen, interleukin-2 production in primary human T cells | 37 true-positive-count count · higher Uncertainty: Not reported by the source Coverage: Not reported scored / Not reported eligible | Independent external evaluation · Source checkedMethods, coverage and sourceBDA-Rand (Llama-3.1-8B backbone) on IL2 (Gupta et al. 2025) tgtval-20261009-protocol-gupta2025-cumulative-hits-round5 Aggregation: Not reported LLMs for Bayesian Optimization in Scientific Domains: Are We There Yet? · Table 1, 'Llama-3.1-8B backbone' block, row 'BDA-Rand', column 'IL2' |
Source checking is not independent reproduction. Release 2026-10-10-7fcc3e48a123.
Use this model
How it works, versions and access
Strengths, limitations and unresolved questions
Evidence
Source checking verifies the cited claim or transcription. It does not establish independent reproduction.
Evidence table
Inspect claims, sources and review details
Trace each statement to its source and review. A context-only reference supports the record generally; it does not verify an individual field. Source checking does not reproduce an experiment.
One row per statement and cited source. Multiple citations are not independent evaluations. Shared locators are labelled explicitly.
0 evidence rows matching the loaded filters
| Property and statement | Original source and location | Review and provenance |
|---|
No evidence rows match these filters. Choose another scope or clear the search.
Sources and history
Release 2026-10-10-7fcc3e48a123 · Record review: source checked
1 source records and release history
- LLMs for Bayesian Optimization in Scientific Domains: Are We There Yet? · Original source · Findings of the Association for Computational Linguistics: EMNLP 2025, pages 15482-15510
Technical metadata and extraction receipts
Stable ID: tgtval-20261009-config-gupta2025-bda-rand-llama-3-1-8b
- areas
- cells-tissues
- contexts
- research
- method types
- foundation_model
- reported name
- BDA-Rand
- foundation model eligible
- true
- source locator
- Table 1, 'Llama-3.1-8B backbone' block
- missing metadata
- version: reason: unreported; note: No release or commit is printed for this configuration
- parameters
- No-Tool variant of BioDiscoveryAgent; experimental feedback replaced by randomly permuted outcomes
- limitations
- Receives randomly permuted outcomes instead of true feedback; a control, not a method anyone would deploy.
Related records
- configuration of: BioDiscoveryAgent
- uses model: Llama-3.1-8B
- system: BDA-Rand (Llama-3.1-8B backbone) on Carnevale (Gupta et al. 2025)
- system: BDA-Rand (Llama-3.1-8B backbone) on Sanchez (Gupta et al. 2025)
- system: BDA-Rand (Llama-3.1-8B backbone) on Sanchez Down (Gupta et al. 2025)
- system: BDA-Rand (Llama-3.1-8B backbone) on IFNG (Gupta et al. 2025)
- system: BDA-Rand (Llama-3.1-8B backbone) on IL2 (Gupta et al. 2025)