| attributes.comparison.adaptation Prompting only Context-only references | Benchmarking large language models GPT-4o, llama 3.1, and qwen 2.5 for cancer genetic variant classification Original source ↗ Table 1 row 'Foundationone', columns for Llama 3 Version: npj Precision Oncology 9:141, published 2025-05-15; PMC12078457 full-text XML Retrieved: 2026-10-09T20:43:36Z | not individually reviewed No individual claim review recorded independent paper Audit detailsField: attributes.comparison.adaptation Source artifact SHA-256: 09c67fcaf7b367d74500db5fd015968389371dcb9ee5a37ccc83241aa63c0f80 Hash scope: Hash scope not separately documented; inspect source record Inspected artifact |
|---|
| attributes.comparison.aggregation Mean over 100 iterations Context-only references | Benchmarking large language models GPT-4o, llama 3.1, and qwen 2.5 for cancer genetic variant classification Original source ↗ Table 1 row 'Foundationone', columns for Llama 3 Version: npj Precision Oncology 9:141, published 2025-05-15; PMC12078457 full-text XML Retrieved: 2026-10-09T20:43:36Z | not individually reviewed No individual claim review recorded independent paper Audit detailsField: attributes.comparison.aggregation Source artifact SHA-256: 09c67fcaf7b367d74500db5fd015968389371dcb9ee5a37ccc83241aa63c0f80 Hash scope: Hash scope not separately documented; inspect source record Inspected artifact |
|---|
| attributes.comparison.budget Not reported Context-only references | Benchmarking large language models GPT-4o, llama 3.1, and qwen 2.5 for cancer genetic variant classification Original source ↗ Table 1 row 'Foundationone', columns for Llama 3 Version: npj Precision Oncology 9:141, published 2025-05-15; PMC12078457 full-text XML Retrieved: 2026-10-09T20:43:36Z | missing or unspecified No individual claim review recorded independent paper Audit detailsField: attributes.comparison.budget Source artifact SHA-256: 09c67fcaf7b367d74500db5fd015968389371dcb9ee5a37ccc83241aa63c0f80 Hash scope: Hash scope not separately documented; inspect source record Inspected artifact |
|---|
| attributes.comparison.dataset_version egfrnsclc-20261009-data-lin2025-foundationone-variants Context-only references | Benchmarking large language models GPT-4o, llama 3.1, and qwen 2.5 for cancer genetic variant classification Original source ↗ Table 1 row 'Foundationone', columns for Llama 3 Version: npj Precision Oncology 9:141, published 2025-05-15; PMC12078457 full-text XML Retrieved: 2026-10-09T20:43:36Z | not individually reviewed No individual claim review recorded independent paper Audit detailsField: attributes.comparison.dataset_version Source artifact SHA-256: 09c67fcaf7b367d74500db5fd015968389371dcb9ee5a37ccc83241aa63c0f80 Hash scope: Hash scope not separately documented; inspect source record Inspected artifact |
|---|
| attributes.comparison.inputs Gene, alteration and tumour type as a natural-language query Context-only references | Benchmarking large language models GPT-4o, llama 3.1, and qwen 2.5 for cancer genetic variant classification Original source ↗ Table 1 row 'Foundationone', columns for Llama 3 Version: npj Precision Oncology 9:141, published 2025-05-15; PMC12078457 full-text XML Retrieved: 2026-10-09T20:43:36Z | not individually reviewed No individual claim review recorded independent paper Audit detailsField: attributes.comparison.inputs Source artifact SHA-256: 09c67fcaf7b367d74500db5fd015968389371dcb9ee5a37ccc83241aa63c0f80 Hash scope: Hash scope not separately documented; inspect source record Inspected artifact |
|---|
| attributes.comparison.metric_implementation Top-1 answer compared with the reference level Context-only references | Benchmarking large language models GPT-4o, llama 3.1, and qwen 2.5 for cancer genetic variant classification Original source ↗ Table 1 row 'Foundationone', columns for Llama 3 Version: npj Precision Oncology 9:141, published 2025-05-15; PMC12078457 full-text XML Retrieved: 2026-10-09T20:43:36Z | not individually reviewed No individual claim review recorded independent paper Audit detailsField: attributes.comparison.metric_implementation Source artifact SHA-256: 09c67fcaf7b367d74500db5fd015968389371dcb9ee5a37ccc83241aa63c0f80 Hash scope: Hash scope not separately documented; inspect source record Inspected artifact |
|---|
| attributes.comparison.population clinically relevant vs VUS via CIViC levels Context-only references | Benchmarking large language models GPT-4o, llama 3.1, and qwen 2.5 for cancer genetic variant classification Original source ↗ Table 1 row 'Foundationone', columns for Llama 3 Version: npj Precision Oncology 9:141, published 2025-05-15; PMC12078457 full-text XML Retrieved: 2026-10-09T20:43:36Z | not individually reviewed No individual claim review recorded independent paper Audit detailsField: attributes.comparison.population Source artifact SHA-256: 09c67fcaf7b367d74500db5fd015968389371dcb9ee5a37ccc83241aa63c0f80 Hash scope: Hash scope not separately documented; inspect source record Inspected artifact |
|---|
| attributes.comparison.protocol_id egfrnsclc-20261009-protocol-lin2025-foundationone-relevant-vs-vus Context-only references | Benchmarking large language models GPT-4o, llama 3.1, and qwen 2.5 for cancer genetic variant classification Original source ↗ Table 1 row 'Foundationone', columns for Llama 3 Version: npj Precision Oncology 9:141, published 2025-05-15; PMC12078457 full-text XML Retrieved: 2026-10-09T20:43:36Z | not individually reviewed No individual claim review recorded independent paper Audit detailsField: attributes.comparison.protocol_id Source artifact SHA-256: 09c67fcaf7b367d74500db5fd015968389371dcb9ee5a37ccc83241aa63c0f80 Hash scope: Hash scope not separately documented; inspect source record Inspected artifact |
|---|
| attributes.comparison.split Whole table; no training split (prompted models) Context-only references | Benchmarking large language models GPT-4o, llama 3.1, and qwen 2.5 for cancer genetic variant classification Original source ↗ Table 1 row 'Foundationone', columns for Llama 3 Version: npj Precision Oncology 9:141, published 2025-05-15; PMC12078457 full-text XML Retrieved: 2026-10-09T20:43:36Z | not individually reviewed No individual claim review recorded independent paper Audit detailsField: attributes.comparison.split Source artifact SHA-256: 09c67fcaf7b367d74500db5fd015968389371dcb9ee5a37ccc83241aa63c0f80 Hash scope: Hash scope not separately documented; inspect source record Inspected artifact |
|---|
| attributes.origin independent_paper Context-only references | Benchmarking large language models GPT-4o, llama 3.1, and qwen 2.5 for cancer genetic variant classification Original source ↗ Table 1 row 'Foundationone', columns for Llama 3 Version: npj Precision Oncology 9:141, published 2025-05-15; PMC12078457 full-text XML Retrieved: 2026-10-09T20:43:36Z | not individually reviewed No individual claim review recorded independent paper Audit detailsField: attributes.origin Source artifact SHA-256: 09c67fcaf7b367d74500db5fd015968389371dcb9ee5a37ccc83241aa63c0f80 Hash scope: Hash scope not separately documented; inspect source record Inspected artifact |
|---|