Вход на сайт

Просмотр новости

Найдите то, что Вас интересует

Building a Living, Ontology-Informed and Large Language Model-Assisted Knowledge Base for Cancer Screening Implementation: A Study Protocol [version 1; peer review: awaiting peer review]

Дата публикации: 29-07-2026 08:46:28

Background Although numerous implementation strategies have been evaluated to improve cancer screening uptake, the evidence base remains fragmented, inconsistently described, and difficult to apply to local contexts. This project will develop a living, ontology-informed, and large language model (LLM)-assisted knowledge base for cancer screening implementation. This resource will serve as a structured, continuously updated, and computable evidence infrastructure, enabling evidence mapping, synthesis, and retrieval of context -sensitive implementation strategies to support evidence-informed implementation decision-making . Methods First, a cancer screening implementation ontology will be developed, building on the Behavior Change Intervention Ontology to formally represent screening-specific pathways, modalities, implementation determinants, strategies, and outcomes. Second, guided by the Cochrane Effective Practice and Organization of Care method, we will systematically identify randomized controlled trials evaluating cancer screening implementation strategies from four databases: PubMed, EMBASE, CINAHL, and PsycINFO, from database inception to May 13, 2026. Information from eligible studies will be extracted and standardized using ontology. All the above steps will be supported by a human-LLM collaborative workflow that incorporates one or more LLMs with source-grounded verification and expert review to improve efficiency while maintaining human accountability. The outputs will include an interactive knowledge base that supports evidence visualization, synthesis, structured comparison, and future decision support. The knowledge base will be continuously updated and progressively expanded to incorporate additional forms of implementation evidence. Workflow performance and knowledge base quality will be evaluated using predefined process and outcome metrics. Discussion The proposed knowledge base will improve the standardization, comparability, and usability of implementation evidence in cancer screening by transforming heterogeneous trial evidence into structured and computable evidence units. Future prediction or decision-support applications, when developed, will contribute to emerging efforts in precision implementation science by enabling a more context-sensitive understanding of cancer screening strategies.

Основное содержимое страницы с новостью.

Introduction
Suboptimal cancer screening uptake

Cancer screening is a key strategy for reducing cancer-related morbidity and mortality. Despite the establishment of evidence-based clinical guidelines and the availability of screening tools and technologies, population-level uptake remains suboptimal across cancer types. For instance, although low-dose computed tomography reduces lung cancer mortality in randomized controlled trials,1,2 fewer than 20% of eligible individuals in the United States undergo recommended screening.3,4 In colorectal cancer screening, overall uptake was reported at only 35.9%.5 These gaps are further compounded by persistent inequities across subgroups.68 In the Cancer CHRNA study, up-to-date mammography, cervical, and colorectal cancer screening rates varied across racial and ethnic groups, ranging from 50.6% to 73.2% for mammography, 12.6% to 76.8% for Pap testing, and 26.8% to 69.0% for colonoscopy.9 Geographic disparities are also evident, with only 48.6% of rural residents and 64.0% of urban residents receiving a Pap test for cervical cancer screening.10 These observations highlight a persistent knowledge-to-practice gap and the need for effective and equitable implementation strategies.

Gaps in cancer screening implementation research

Implementation science seeks to bridge the gap between evidence and practice by developing and evaluating implementation strategies.11 In cancer screening, a wide range of implementation strategies have been evaluated to improve uptake across patients, clinicians, and system levels, such as reminders, patient navigation, outreach, and workflow redesign.1215 Across cancer types, many of these strategies target similar behavioral, organizational, and structural barriers, such as limited patient awareness, competing clinician demands, fragmented care pathways, or resource constraints.

Despite rapid growth in implementation research, the evidence base remains fragmented and difficult to translate into practice. Implementation trials are dispersed across cancer-specific, public health, and implementation science literature, making it difficult to maintain a comprehensive understanding of the field. Implementation strategies are described using inconsistent terminology and heterogeneous reporting formats, limiting comparisons across studies and cancer types.16 This lack of standardized reporting also hinders the identification of active intervention ingredients and mechanisms of action through which implementation strategies influence screening behaviors.17

In addition, most existing evidence syntheses are static and quickly become outdated, limiting their value for ongoing decision-making.18 There is no continuously updated resource that systematically organizes implementation evidence across cancer types, populations, and contexts. This gap limits the advancement of both implementation research and practice.

Toward a large language model-assisted, ontology-informed living knowledge base

Recent advances in large language models (LLMs) have created new opportunities to support labor-intensive evidence review tasks, including literature search, screening, and data extraction.19,20 These capabilities may support more efficient maintenance of a continuously updated (“living”) knowledge base. In addition, LLMs have recently been applied to ontology engineering, supporting the generation of candidate concepts, definitions, and relationships under expert oversight.21,22 An ontology is a structured framework that organizes knowledge about entities or concepts through standardized terms, definitions, and relationships, so that information can be represented in a computer-readable format. Ontologies can support consistent organization, integration, reuse, and analysis of knowledge.23 In cancer screening, an ontology-informed structure could help translate inconsistently reported trial descriptions into comparable evidence units. These advancements create new opportunities to integrate LLM-assisted workflows with ontology-based representations to address fragmentation in research evidence. To date, however, LLM and ontology have not been systematically applied to develop a comprehensive knowledge base for cancer screening implementation.

Aim

The overall aim of this project is to develop a living, LLM-assisted knowledge base for cancer screening implementation that supports continuous evidence visualization, synthesis, structured comparison, and future decision support. The knowledge base is designed to represent evidence in a standardized, computable, and provenance-aware format through a shared ontology framework.

Objective 1: Develop a cancer screening implementation ontology to structure and standardize cancer screening implementation evidence.

Objective 2: Develop and validate a reusable, ontology-informed human–LLM collaborative workflow to support literature review and its living updates.

Objective 3: Build the initial cancer screening implementation knowledge base.

Objective 4: Assess the coverage, representation, integrity, queryability, and traceability of the final release-ready knowledge base, with ongoing monitoring of workflow drift across living-update cycles.

Methods
Ontology-informed knowledge base framework

An ontology-informed framework will guide the harmonization, organization, representation, and continuous updating of evidence on cancer screening implementation. In this study, we will leverage the Behavior Change Intervention Ontology (BCIO) as the overarching conceptual framework.2426

The BCIO is a theory-informed ontology framework that comes from The Human Behaviour-Change Project, which provides a standardized and computable representation of behavior change interventions and their related components.2426 To support detailed and standardized representation of implementation evidence, several BCIO-compatible lower-level ontologies will be incorporated, including the Behavior Change Technique Ontology,27 Mechanism of Action Ontology,28 Mode of Delivery Ontology,29 Intervention Setting Ontology,30 Population Ontology31 and the Intervention Source Ontology.32

Building on this BCIO-aligned conceptual foundation, a cancer screening implementation ontology will be developed to further represent domain-specific constructs not explicitly captured in existing BCIO structures. These constructs include screening modalities, implementation determinants, implementation strategies, screening pathway stages and associated outcomes.

Cancer screening pathway is characterized by multi-step and longitudinal behaviors that require structural representation. To capture this complexity, ontology modules will be organized around a core screening pathway, which defines shared stages across cancer types, including invitation, decision-making, test completion, follow-up, and longitudinal adherence. Cancer-specific modules then apply this core structure to different screening programs.

We will use the project’s ontology-informed human–LLM workflow (see “LLM-assisted workflow”), configured with ontology-specific per-task protocols, to support ontology development. The workflow will include automated checks for duplicate or near-duplicate classes, verification of class hierarchies and relationships,21 and conflict-resolution procedures that route candidate concepts for integration, human review, or exclusion.33 Given the known limitations of LLM-assisted ontology engineering,22,34 domain experts will retain responsibility for all schema-level decisions. The ontology will be developed through four steps:

  • 1) Extract candidate concepts and their relations. One or more extraction LLMs operating under ontology-grounded per-task protocols will extract candidate concepts and their relations from literature such as randomized trials, reviews, and clinical guidelines. These protocols will constrain extraction to predefined entity types (e.g., interventions, settings, populations, mechanisms, screening stages, and outcomes) and relation types, and exclude off-scope content.21,35 Each model will generate structured candidate assertions linked to verbatim source text and document metadata. Candidate extraction identified by one or more models will then be consolidated through the source verification and expert adjudication process, with input from the study team to ensure conceptual accuracy, consistency, and completeness.

  • 2) Normalize identified concepts. The BCIO and its lower-level ontologies will be instantiated as a machine-readable class registry to support automated lookup and semantic matching. Each candidate concept will then be normalized to existing ontology classes using combined lexical, embedding-based, and hierarchical similarity matching, applying a “reuse before mint” principle: an existing class is reused whenever semantic equivalence or sufficient subsumption is established, and a new class is proposed only when no adequate match exists.31,32 All reuse and mappings decisions will be verified against the source ontologies to confirm correct alignment, proper parent-class assignment, and avoidance of redundant classes.19

  • 3) Develop modules. Aligned and newly introduced concepts will be organized hierarchically under the core screening pathway superclass and within cancer-specific modules. For each concept, we will develop a standardized specification using a retrieval-augmented workflow that draws on evidence from included studies, screening guidelines, and existing ontology resources. The specification will include a preferred label and synonyms, a clear definition, inclusion and exclusion criteria, illustrative examples, relationships to parent and related concepts, and annotation guidance to support consistent coding and application across studies. High-quality BCIO entries will be used as structural exemplars to maintain consistent modeling patterns.18,32 The core screening pathway superclass, concept definitions, boundary criteria, and hierarchical relationships will be reviewed and finalized by domain experts to ensure conceptual clarity, coherence, and non-overlap.19

  • 4) Pilot-testing and iterative refinement. The draft ontology will be evaluated through pilot annotation of a purposive sample of randomized trials representing multiple cancer screening domains (e.g., breast and colorectal cancer screening) to assess its applicability across diverse screening contexts.21,36 During pilot testing, annotation challenges, including ambiguous or poorly defined concepts, overlapping classes, missing concepts, and inconsistent mappings across LLM-models or human annotators, will be systematically documented. These findings will be used to refine concept definitions, class boundaries, hierarchical relationships, mapping rules, annotation guidance, and the associated extraction and annotation protocols. The ontology and workflow will be iteratively revised and retested until predefined criteria for stability and consistency are achieved, such as acceptable annotation agreement, reduced numbers of unresolved concepts, and consistent concept mapping across screening domains. The resulting ontology will be released as Version 1.0 under a version-controlled governance process.

As the living knowledge base evolves, newly identified concepts will be reviewed through a version-controlled ontology governance process and incorporated into subsequent releases where appropriate. Consistent with previous LLM-assisted BCIO work, in which fully automated ontology annotation achieved an F1 of approximately 0.42 compared with a human benchmark of 0.75,37 the LLM-assisted workflow in this project will be used as a decision-support and prioritization tool rather than a fully autonomous system. All flagged, uncertain, novel, or high-impact concepts will undergo expert review and adjudication before being incorporated into the ontology.

LLM-assisted workflow

The project will be implemented through a human–LLM collaborative workflow to support reproducibility, transparency, and task-specific performance. Each discrete task in this workflow is governed by a reusable, modular, and version-controlled per-task protocol. Each protocol operationalizes the task for the model (specifying task objectives, input requirements, extraction or annotation rules, the required output schema, worked examples, and prohibited inferences) and defines how the resulting outputs are verified (through quality-control checks, validation criteria, and a version history). The runtime prompt issued to a given model is derived from, and versioned against, its protocol, so that the protocol remains stable and auditable while prompts or model versions change.38 Each LLM-assisted task will generate structured candidate outputs that include source text, source location, evidence type, confidence rating where applicable, protocol version, model metadata where available, and validation status. These outputs will be treated as provisional until verified against source evidence and validated through predefined quality control procedures.

Source verification and human adjudication architecture

Figure 1 summarizes the overall workflow architecture. The workflow employs one or more LLMs operating under predefined per-task protocols and structured schemas. If more than one model is used, outputs may be compared across models as a triage signal. Only outputs that are supported by source evidence are accepted as provisional candidates. Outputs that are inconsistent or uncertain are first checked automatically and, if needed, reviewed by human experts. Human-adjudicated outputs will be the reference standard for process validation. This verification-driven design keeps final decisions human-accountable and supports a model-agnostic workflow suitable for continuous updates in a living evidence system.

83a5015e-d610-424c-8d2e-649d115d5040_figure1.gif

Figure 1. Project workflow

Pre-deployment validation criteria for LLM-assisted tasks

Before deployment in the living knowledge base, each LLM-assisted task will be validated against a human-adjudicated reference standard using an independent validation set. Validation criteria will be defined a priori based on the complexity of each task: (1) During title/abstract screening, we will prioritize sensitivity to minimize the risk of excluding eligible studies. The LLM is expected to achieve ≥90% sensitivity, with uncertain records routed to human review; (2) During full-text screening, the LLM must provide explicit eligibility evidence extracted from the source text and achieve ≥85% agreement with human reviewers before being deployed for full-scale screening; (3) Structured data extraction must include source-text evidence, pass schema validation, and achieve acceptable field-level performance (≥85% field-level accuracy), with complex or uncertain fields routed to human adjudication; (4) Ontology annotation must use valid ontology identifiers, where applicable, include source-text justification, pass ontology consistency checks, and achieve acceptable agreement with expert-generated mappings; and (5) For risk-of-bias assessment, the LLM will identify methodological evidence and generate preliminary domain-level judgments, with uncertain, unsupported, or conflicting judgments routed to human adjudication. Outputs that fail to meet validation criteria, require human adjudication, or exhibit recurrent uncertainty will inform iterative refinement of the protocols and LLM workflows.

Protocol refinement

Protocol refinement will be iterative. Human adjudication logs will be reviewed to identify recurrent error patterns, such as missed intervention components, unsupported ontology mappings, invalid ontology identifiers, inconsistent outcome extraction, and incorrect interpretation of risk-of-bias evidence. These error patterns will be used to update task instructions, extraction schemas, annotation guidance, examples, and quality-control rules. Substantive changes to protocols, prompts, model versions, ontology terms, or output schemas will be documented and re-evaluated before deployment in the living workflow.

Deployment and monitoring of LLM-assisted tasks

LLM-assisted tasks will be deployed in the living knowledge base workflow only after meeting predefined validation criteria against a human-adjudicated reference standard. Workflow performance will be monitored through periodic human review of LLM-processed records, and revalidation will be required after substantive changes to models, protocols, schemas, ontologies, or eligibility criteria. Tasks that fall below predefined thresholds will be paused, routed for human review, refined, and revalidated before redeployment.

Artificial intelligence uses and disclosure

All LLM assistance in this project will be reported transparently and documented for reproducibility. For each LLM-assisted task, we will document the model’s name and version, access mode and date of use, the protocol/prompt version and input-document version, inference settings where configurable, the output schema, and the human-adjudication decision. LLMs will be used only as tools under human supervision and will not be listed as authors. All LLM-generated outputs will remain provisional until they pass source verification and human adjudication.

Building the initial knowledge base

To build the knowledge base, we will follow the Cochrane Effective Practice and Organization of Care (EPOC) Group’s methodological guidance for systematic reviews39 and the Responsible use of AI in Evidence Synthesis (RAISE) principles proposed by Cochrane.40 All processes will be governed by the human–LLM workflow described above. LLM-assisted outputs will be incorporated into the knowledge base only after completion of the predefined verification and adjudication procedures.

Literature search

Four databases will be searched: PubMed, EMBASE, CINAHL, and PsycINFO. The search strategy will be developed and iteratively refined based on published literature and the expertise of the research team. Searches will be structured around three core concepts: cancer, screening, and uptake, with database-specific controlled vocabulary and keywords used to maximize sensitivity for implementation-related studies (see the Extended data for the PubMed search strategy). The initial search will include studies published up to May 13, 2026.

The Cochrane EPOC guideline recommends four types of empirical evidence for systematic reviews that aim to test the effectiveness of implementation strategies: randomized trials, non-randomized trials, controlled before-after studies, and interrupted time series and repeated measure studies.41 For building the initial knowledge base, randomized controlled trials will be prioritized for inclusion to support the development of a high-internal-validity knowledge base and to enable initial ontology and workflow calibration. Additional study designs will be considered in subsequent phases of the living knowledge base.

Literature screening

For the purpose of this knowledge base, cancer screening is defined as the use of evidence-based screening tests or examinations in individuals without signs or symptoms of cancer to identify cancer or precancerous lesions at an earlier, potentially more treatable stage. Diagnostic evaluation of symptomatic individuals, post-treatment surveillance, and interventions primarily focused on hereditary cancer risk assessment (e.g., germline genetic testing or genetic counselling) were considered outside the scope of the initial knowledge base.42

Records retrieved from all databases will be deduplicated and screened at the title/abstract and full-text stages. Before deployment, the LLM-assisted screening workflow will be evaluated using an independent validation sample of 500 records. These records will undergo both human screening and LLM-assisted screening, with human-adjudicated decisions serving as the reference standard (see details in Table 5). Screening performance will be assessed against the predefined validation criteria described in the previous sections. Records with discordant or uncertain classifications will be reviewed by human experts and used to refine screening protocols and decision rules where necessary. Once the validation criteria are met, remaining records will be screened using the validated LLM-assisted screening process described in the previous sections, with uncertain or conflicting cases continuing to be routed to human review. Eligibility criteria are presented in Table 1.

Table 1. Eligibility criteria.PICOS ElementInclusion CriteriaExclusion CriteriaPopulation Individuals eligible for any type of cancer screening according to established guidelines; healthcare providers; healthcare organizations or systems involved in cancer screening deliverySymptomatic individuals undergoing diagnostic workup; post-treatment surveillance populationsIntervention Any implementation strategy or implementation strategy package intended to improve the implementation of evidenced-based cancer screening, its uptake, follow-up, adherence, quality, equity, or delivery processesEfficacy trials of screening modalities themselves; interventions targeting uptake of germline/genetic testing, genetic counseling, or cascade testing for hereditary cancer risk; decision support interventions (e.g., decision aids) whose sole objective is to improve decision quality, knowledge, or preference clarification without evaluating screening implementation-related outcomes; de-implementation interventions targeting overscreeningComparator Usual care, waitlist control, active comparison implementation strategySingle-arm studies without a comparatorOutcomes Screening-related behavioral outcomes
• Screening uptake/completion
• Repeat screening adherence
• Return of screening kits (e.g., fecal immunochemical test (FIT) return)
• Diagnostic follow-up completion after abnormal results
• Screening invitation response
• Appointment scheduling completion
• Referral completion
• Time to screening completion or follow-up
Implementation outcomes: Reach, fidelity, sustainability, implementation cost, workflow integration, provider screening recommendation or referral behaviors
Equity-related outcomes: Reduction in disparities in screening participation or follow-up across racial/ethnic, socioeconomic, geographic, language, insurance, or other underserved populations• Studies reporting only biological or clinical endpoints (e.g., cancer incidence, stage distribution, mortality, survival) without screening-related behavioral or implementation outcomes
• Studies examining knowledge and attitudes without an implementation focus
• Studies reporting only decision-related outcomes (e.g., decision quality, shared decision-making participation, decisional conflict) without screening-related behavioral or implementation outcomes.Study Design Randomized controlled trials, including individual randomized trials, cluster randomized trials, and stepped-wedge randomized trials; secondary analyses evaluating the effects of the randomized intervention (e.g., subgroup analyses); pilot or feasibility RCTs reporting intervention effectsNon-randomized studies, non-empirical publications, study protocols, review articles, and conference abstracts without full reports; independent studies (e.g., surveys or questionnaires) conducted within an RCT

Data extraction

Data extraction will be guided by the Cochrane EPOC template.43 Implementation-relevant information extraction will be informed by the Expert Recommendations for Implementing Change (ERIC) taxonomy44 and Proctor’s implementation outcomes framework.45

Two levels of data will be extracted. Study-level fields will include publication details, country, study design, theoretical framework, unit of randomization, setting, cancer type, screening modality, target population, eligibility criteria, intervention and comparator conditions, sample size, follow-up period, outcomes, and effect estimates where available.

Implementation-relevant fields will include target screening behavior, screening pathway stage, target actor, intervention recipient, intervention source, mode of delivery, implementation setting, population characteristics, engagement, implementation determinants, equity-relevant context, implementation strategies, and implementation outcomes.

All extracted data will be linked to source text and provenance metadata to ensure traceability. For each reported outcome, both the effect estimates (where available) and explicit effect-direction indicators (e.g., favoring intervention, favoring control, null, or non-inferior) will be recorded to preserve directionality and support downstream synthesis of null and negative findings.46 LLM-assisted data extraction will be validated against human-adjudicated reference standards using an independent sample of approximately 50 randomized controlled trials. Detailed procedures and evaluation metrics are provided in Table 5.

Ontology annotation

Intervention descriptions will be decomposed into discrete components within each study arm. Original author-reported descriptions are retained prior to standardization to preserve traceability. Each component is mapped to the ontology framework to enable structured representation across key dimensions (see Table 2 for details).

Table 2. Cancer screening implementation ontology annotation schema.Ontology domainSource ontologyPurpose of annotationExample codingImplementation strategy Cancer screening implementation ontology (To be developed; Informed by the Expert Recommendations for Implementing Change (ERIC) taxonomy44)Represent the overarching implementation approach used to improve cancer screening implementationPatient navigation; provider reminder; implementation facilitationBehavior change technique Behavior Change Technique Ontology27Represent the active ingredients intended to influence screening-related behaviors or implementation processesPrompts/cues; social support (practical); problem solvingMechanism of action Mechanism of Action Ontology28Represent hypothesized processes through which implementation strategies influence screening-related behaviors or implementation outcomesKnowledge; beliefs about capabilities; environmental context and resourcesImplementation determinants Cancer screening implementation determinant module (To be developed; Informed by the Consolidated Framework for Implementation Research68)Represent patient-, provider-, organizational-, system-, and equity-related determinants that may influence implementation processes and outcomesTransportation barrier; limited healthcare access; workflow constraints; staffing capacity; insurance barrier; structural racismMode of delivery Mode of Delivery Ontology29Represent how implementation strategies are deliveredTelephone; SMS (short message service); postal mail; electronic health record promptIntervention source Intervention Source Ontology32Represent who delivers the implementation strategyClinician; patient navigator; automated system; community health workerImplementation setting Intervention Setting Ontology30Represent where implementation strategies are deliveredPrimary care clinic; community setting; public health programScreening pathway stage Cancer screening implementation ontology (To be developed)Represent the stage of the cancer screening continuum targeted by the strategyInvitation; scheduling; test completion; abnormal-result follow-up; repeat screening adherencePopulation characteristics Population Ontology31 + cancer screening implementation ontology (To be developed)Represent characteristics of populations receiving implementation strategies, including demographic, clinical, and screening-related characteristicsNever screened; under-screened; medically underserved population; average risk; HIV-associated riskScreening outcome Cancer screening implementation ontology (To be developed)Represent screening-specific behavioral, implementation, and pathway-related outcomesScreening completed; FIT returned; diagnostic follow-up completed; repeat screening adherence

To balance accuracy and computational cost across the cancer screening implementation ontology and its lower-level ontologies, a two-stage retrieve-and-rerank workflow will be used. In the first stage, an embedding-similarity retrieval module selects a small set of top-k candidate ontology terms from each annotation domain. In the second stage, the LLM model will select the most appropriate term from this restricted candidate set, using the source text and intervention component as context.46 For each annotation domain, the LLM workflow will also include an explicit no-applicable-term option. Records assigned to this option will be routed to the ontology governance process as candidate concepts for future ontology versions. This process will support the identification of coverage gaps within the cancer screening implementation ontology and inform subsequent ontology development.46 LLM-assisted ontology annotation will be validated against human-adjudicated reference standards using 5–10 pilot trials for calibration, followed by an independent validation sample of 20–30 randomized controlled trials. Detailed procedures and evaluation metrics are provided in Table 5.

Risk-of-bias assessment

Risk of bias in included randomized controlled trials will be assessed using the Cochrane Risk of Bias 2 (RoB 2) tool.39,47 Domains assessed will include bias arising from the randomization process, deviations from intended interventions, missing outcome data, measurement of the outcome, and selection of the reported result. As later phases incorporate non-randomized study designs, appropriate additional risk-of-bias tools (e.g., ROBINS-I) will be adopted. Risk-of-bias assessments will be informed by LLM-assisted identification of relevant methodological text, but all domain-level judgments will be finalized through human validation. The LLM-assisted risk-of-bias workflow will be validated against human-adjudicated reference standards using an independent sample of 20–30 randomized controlled trials. Detailed procedures and evaluation metrics are provided in Table 5.

Data analysis

Descriptive and exploratory analyses. We will summarize the basic characteristics of included studies and examine variation in implementation strategies across cancer types, healthcare settings, populations, and equity-relevant contexts. Results will be reported using counts, proportions, and cross-tabulations. Building on these analyses, results will be further organized into evidence tables, descriptive summaries, and implementation strategy frequency distributions. The evidence units will be integrated into a provenance-aware evidence graph that represents relationships among implementation strategies, behavior change techniques, mechanisms of action, implementation determinants, screening modalities, screening pathway stages, populations, settings, outcomes, and effect estimates, where available. These outputs will support evidence synthesis, knowledge base querying, and structured reporting.

Network and co-occurrence analysis. The ontology-informed evidence graph will support network-based evidence mapping and exploratory network analyses of relationships among implementation strategies, behavior change techniques, mechanisms of action, implementation determinants, screening pathway stages, and outcomes. Co-occurrence patterns across studies will be examined to identify frequently co-occurring implementation components, recurrent implementation configurations, and clusters of strategies used within specific screening contexts. For example, analyses may examine which implementation strategies and behavior change techniques are most frequently associated with specific screening modalities, pathway stages, implementation determinants, or target populations. The graph also supports identification of gaps in evidence across cancer types, populations, screening stages, healthcare settings, and equity-relevant contexts.

Exploratory hypothesis-generation analyses. Exploratory graph-based analyses may be conducted using the evidence graph to identify under-studied implementation strategy–context–outcome relationships. If link-ranking or neighborhood-similarity methods are used, they will be reported as research-prioritization tools rather than predictions of effectiveness and may generate candidate links between implementation strategies, populations, screening modalities, and outcomes. Candidate relationships will be reviewed alongside available effect estimates, reporting completeness, and provenance quality. Details of the exploratory analysis procedures are provided under outcome evaluation ( Table 4).

Graph-based querying and downstream applications. The ontology-informed evidence graph will be queried using structured graph-query approaches (e.g., SPARQL- or Cypher-style queries over evidence triples) to retrieve and organize evidence across screening modalities, screening pathway stages, populations, healthcare settings, and equity-relevant contexts. Query results will be used to identify implementation strategies, behavior change techniques, mechanisms of action, implementation determinants, and reported outcomes associated with specific screening contexts. All retrieved results will remain linked to their original source studies through provenance records. These query outputs will be used to generate evidence-gap maps, comparative analyses of implementation approaches, and other knowledge base outputs planned for later phases of the project.

Data synthesis

Upon establishment of the ontology-informed evidence infrastructure, a subsequent project will use systematic review and meta-analysis approaches to synthesize the effectiveness of cancer screening implementation strategies. This follow-on work will examine which strategies are effective across cancer screening contexts, which are most effective for specific screening pathways or outcomes, and which strategies are most applicable or beneficial for particular cancer screening modalities.

Living updates and incorporation of other forms of evidence

This project is designed as a living knowledge base in which newly identified studies are periodically incorporated using the established update workflow. Search updates will follow a two-tier schedule, with application-programming-interface-accessible databases (e.g., PubMed through E-utilities) queried at high frequency (e.g., weekly) and databases requiring manual interface searches (e.g., EMBASE) queried at lower frequency (e.g., quarterly to biannually).48 To keep the cumulative knowledge base reproducible across releases, evidence units that are not affected by an ontology, model, protocol, or schema change will be preserved across update cycles without re-annotation. Only records affected by the predefined recalibration trigger or routinely sampled for drift audit will be re-annotated. This update strategy follows an append-only versioning approach consistent with established living evidence systems.48 A designated team member will oversee the living-update process, confirm the search schedule, and approve each knowledge base release. The living-review component will be reported in accordance with the PRISMA-LSR extension for living systematic reviews.49

The living workflow also supports iterative refinement of the ontology and associated annotation guidelines. New or under-represented concepts identified during updates will be reviewed through the ontology governance process and incorporated into subsequent versioned releases where necessary.

The initial version of the knowledge base focuses on randomized controlled trials, with later phases expanding to incorporate non-randomized studies, controlled before-after studies, interrupted time series studies, and other forms of real-world implementation evidence where appropriate. In selected cases and where feasible, additional collaboration with trial investigators may be explored to obtain individual participant data to support more detailed representation of implementation contexts, populations, screening pathways, and outcomes.

Process and outcome evaluation

Both process and outcome evaluations will be conducted to assess the performance, reproducibility, transparency, and ongoing maintainability of the ontology-informed and LLM-assisted workflow and the resulting knowledge base.5052

Process evaluation

Process evaluation will assess workflow performance by examining the agreement between LLM-assisted outputs and human-adjudicated reference standards across literature screening, data extraction, ontology annotation, and risk-of-bias assessment ( Table 3), with ongoing monitoring of living-update currency and workflow drift. Workflow development and validation will be separated: a small pilot calibration sample will first refine eligibility criteria, protocols, prompts, output schemas, and annotation guidance (treated as development and quality assurance), and formal validation will be followed only after the protocols are finalized, using independent human-reviewed samples. Where appropriate, error rates from these validation samples may be used to estimate uncertainty in the results of the full-corpus LLM-assisted output. Evaluation of the ontology itself (e.g., ontology coverage, consistency, fitness for purpose, and expert assessment of the ontology structure) is outside the scope of this initial phase and will be assessed in subsequent dedicated studies.

Table 3. Process evaluation of the LLM-assisted workflow (see Table 5 for the detailed process-evaluation plan; the per-release living-update validation is detailed in Table 6).Workflow stagePrimary focusPrimary metricsTitle/abstract and full-text screeningReproduce human eligibility decisions; prioritize sensitivitySensitivity, specificity, Cohen’s κ (LLM vs. tiebroken human)Data extractionAccuracy, completeness, source-grounding Field-level accuracy (exact match/mean absolute error (MAE) /F1); strict record accuracy; source-text verificationOntology annotationAgreement with human annotationsAssertion-level precision/recall/F1; ontology agreement and conformance; hallucination rateRisk-of-bias assessmentReliability of LLM-assisted RoB 2 supportDomain-level agreement and Cohen’s κ; source-text verificationLiving update (per-release ontology-version validation)Per-release validation with backward compatibility and preservation of existing contentNew-content concept/relation precision, recall, F1 (release-specific reference standard); version diff (remain/add/delete/update, per category); competency-question regression; reasoner consistency; inter-annotator κ; drift-benchmark performance

Validation metrics are aligned with the operational goals of each task ( Table 3). Screening validation will prioritize sensitivity to minimize false exclusions of eligible studies for downstream review, whereas data extraction, ontology annotation, and risk-of-bias assessment will emphasize agreement with human judgments and source-text grounding to limit unsupported inference. Substantial changes to models, prompts, ontology terms, or output schemas introduced during living updates will trigger the same recalibration.48

Outcome evaluation

Outcome evaluation will assess the completeness, structure, transparency, and usability of the final release-ready knowledge base generated in this initial development phase. These evaluations will be conducted on the final human-validated or release-ready knowledge base rather than on provisional LLM-generated outputs. The primary outcome evaluation will focus on coverage of existing systematic review evidence by determining whether eligible randomized trials included in published reviews are represented in the released knowledge base. Secondary evaluations will assess evidence representation breadth and balance, release integrity and traceability, and later-stage usability and perceived relevance of the knowledge base. Detailed indicators, definitions, and evaluation metrics for each dimension are provided in Table 4.

Table 4. Outcome evaluation of the released knowledge base (see Table 6 for the detailed outcome-evaluation plan).Evaluation domainPrimary indicatorsCoverage of existing systematic review evidenceAnchor-study capture (%), missed-anchor rate; direction and significance agreement versus Cochrane/EPOC pooled effects70Evidence representation breadth and balanceNumber and proportion of studies/evidence units across cancer type, modality, pathway stage, strategy, population, and equity-relevant contextRelease integrity and traceabilityProvenance-field completeness; source-support rate; release validation pass rateQueryability of representative evidence questionsWhether representative evidence questions can be answered with source-supported graph queriesExploratory research-prioritization outputs, if implementedOptional descriptive signals for under-studied combinations; not effectiveness predictionUsability and perceived relevance*Perceived usefulness, interpretability, navigability, and traceability to source studiesLiving-update transparencyUpdate-cycle metadata; drift-benchmark performance; audit performance; incremental-search recall

Table 5. Detailed process-evaluation plan for the LLM-assisted workflow.Workflow stagePrimary focusValidation sampleHuman reference standardPrimary metricsTitle/abstract eligibility screening Whether the final LLM-assisted workflow correctly identifies potentially eligible records at title/abstract stage and confirms final inclusion at full-text stage• Approximately 500 randomly selected records from the deduplicated pool
• if few eligible records are present, add a positive-enriched seed set from known trials and prior reviews46Two independent reviewers will independently screen each record. An independent senior adjudicator will resolve conflicts for reviewer disagreements.46 Tiebroken human decisions will serve as the reference standard for both title/abstract and full-text eligibility validation.Sensitivity, specificity, and Cohen’s κ between LLM-assisted and tiebroken human decisions will additionally be reported with 95% confidence intervals computed via percentile bootstrap (50,000 resamples). Corpus-level estimates of true positive, true negative, false positive, and false negative counts in the full deduplicated search pool will be derived using a Bayesian model with a Jeffreys-prior beta posterior on prevalence, propagating sensitivity and specificity uncertainty from the validation sample.46Full-text eligibility screening71Full-text eligibility validation will be conducted on all or a purposive sample of records advanced to full-text review within the validation subset, enriched for uncertain, discordant, and borderline cases.Sensitivity, specificity, Cohen’s κ, at the full-text stage will additionally be reported with 95% confidence intervals computed via percentile bootstrap (50,000 resamples).46Data extraction Accuracy, completeness, and source-grounding of extracted study informationApproximately 50 included randomized trials. The sample should include variation by cancer type, screening modality, study design, setting, intervention complexity, and outcome type.Human verification or correction of extracted fieldsStratified by field type:
• Simple fields: exact match;
• Numeric fields: MAE or tolerance accuracy;
• Complex fields: precision/recall/F1; Completeness per field
These field-level metrics will be aggregated into record-level audits as follows. Specifically, simple fields include bibliographic items (e.g., year, journal, country, funding source) and core study design items (study design, unit of randomization, sample size, follow-up period); numeric fields include effect estimates, confidence intervals, event counts, and follow-up durations; complex fields include free-text intervention component descriptions, eligibility criteria, and outcome definitions, evaluated against human-adjudicated content units.
Strict record accuracy, the proportion of audited trials for which all required fields are jointly correct, will additionally be reported alongside field-level metrics.48 Each extracted field must be linked to a verifiable source-text span (source-text verification rate), and errors will be classified by severity (critical: changes eligibility, effect estimate, or major conclusion; major: materially changes interpretation; minor: formatting or non-substantive). Field-level accuracies, completeness rates, and Cohen’s κ where applicable will be reported with 95% confidence intervals computed via percentile bootstrap (50,000 resamples).46Ontology annotation24,72Agreement between LLM-assisted and human ontology annotations across ontology domains and hierarchy levels5–10 pilot trials for calibration, then 20–30 validation trials.Human-adjudicated evidence units, including verified intervention, components, ontology mappings, relationships, and source-text support• Assertion-level correctness
• fact/assertion-level precision, recall, and F1 where feasible
• component overlap
• missed-component rate
• unsupported-component rate
• ontology/schema conformance rate
• valid ontology identifier rate
• exact-match, parent-level, and domain-level ontology agreement
• unsupported mapping rate
• hallucinated component or relationship rateRisk-of-bias assessment Reliability of LLM-assisted support for RoB 2 assessment within a human-supervised workflow20–30 included trialsHuman RoB 2 judgments by domain• Domain-level agreement and Cohen’s κ between LLM-assisted and human RoB 2 judgments
• unclear-rate comparison calculated as differences in frequency of unclear-risk judgments
• source-text verification assessed as the proportion of judgments supported by cited methodological text

Table 6. Detailed outcome-evaluation plan for the released knowledge base.Evaluation domainPrimary focusEvaluation approachPrimary indicatorsCoverage of existing systematic review evidence70Completeness of the released knowledge base relative to benchmark systematic review evidenceFor each selected review, included randomized trials will be extracted and compared against the released knowledge base to determine whether each eligible trial is represented. Trials not captured in the knowledge base will undergo targeted review to identify reasons for non-capture, including retrieval failure, screening exclusion, duplicate linkage issues, or incomplete bibliographic information. A pre-specified missing-reason taxonomy will be applied: (1) retrieval failure (record not returned by the search strategy); (2) screening exclusion (excluded by LLM or human reviewer); (3) duplicate or linkage issue; (4) incomplete bibliographic information; and (5) captured but unmappable to the current cancer screening implementation ontology. Category (5) is treated as a coverage gap of the ontology rather than the literature corpus and feeds back to the ontology governance process.
As a second-layer validation, for meta-analyzable outcome subsets in which the same outcome is reported across multiple included trials, the pooled effect re-computed from the trials available in the released knowledge base will be compared to the corresponding Cochrane/EPOC meta-analytic effect in terms of direction (favoring intervention vs control) and statistical significance, following the Trial Bank Coverage validation framework.70anchor studies captured (%), missed-anchor rate, missing reason taxonomy. Stratified anchor capture by cancer type, screening modality, and equity-relevant context; ontology coverage gap rate (proportion of captured anchor trials unmappable to the ontology).
For meta-analyzable subsets: direction-agreement rate, statistical-significance-agreement rate, and joint direction-and-significance agreement rate between knowledge-base-derived and Cochrane/EPOC pooled effects.Evidence representation breadth and balance Distribution of evidence across cancer screening implementation dimensionsReleased studies and evidence units will be characterized across key representation dimensions, including cancer type, screening modality, screening pathway stage, implementation strategy type, target actor, healthcare setting, population group, and equity-relevant context. For each category, the number and proportion of studies or evidence units assigned to the category will be calculated.• Number and proportion of studies or evidence units across representation categories
• identification of underrepresented or concentrated evidence areasRelease integrity and traceability Completeness, transparency, auditability, and provenance tracking of released evidence unitsFor each provenance field, completeness will be calculated as the proportion of required fields with non-missing values. Source-support and unsupported assertion rates will be assessed through targeted audits comparing structured evidence units against cited source text. If implemented, release validation pass rate will be calculated as the proportion of records passing automated graph or schema validation checks.• Provenance field completeness
• overall provenance completeness
• source-support rate
• unsupported assertion rate
• release validation pass rateQueryability of representative evidence questions: internal coherence Whether the released knowledge base can be used to retrieve and answer questions about evidence it already containsAuto-generate yes/no question–answer pairs from released evidence units with high extraction confidence and substantive supporting source-text spans (~50 per cancer type), with question stems covering strategy effectiveness, contextual moderators, pathway-stage outcomes, and equity contrasts. Four retrieval-augmented generation (RAG) settings will be compared using the same answering model: baseline language model with no retrieval, a generic biomedical knowledge graph baseline, retrieval over raw article abstracts, and retrieval over the released knowledge base. Performance will be reported with bootstrap 95% confidence intervals.Exact yes/no accuracy; semantic similarity between generated and reference explanations; relative gain of knowledge-base-retrieval over baseline and over raw-article retrieval
• Exact yes/no accuracy
• semantic similarity between generated and reference explanations
• relative gain of knowledge-base-retrieval over baseline and over raw-article retrievalQueryability of representative evidence questions: external implementation-question utility Whether the released knowledge base supports independently formulated implementation questions that were not used to build the knowledge baseAn external benchmark of implementation questions will be constructed from plain-language summaries, key messages, and implications-for-practice statements of published Cochrane EPOC reviews and other implementation systematic reviews on cancer screening. Questions will be reformatted into yes/no or short-answer items and held out from knowledge base construction. The same RAG settings as in internal coherence evaluation will be compared, together with a combined knowledge base plus generic-graph setting to test complementarity.Accuracy and semantic similarity on the external benchmark; complementarity between knowledge-base retrieval and generic-graph retrieval; stratified performance by question type (strategy effectiveness, contextual moderator, equity, screening pathway stage)
• Accuracy and semantic similarity on the external benchmark
• complementarity between knowledge-base retrieval and generic-graph retrieval
• stratified performance by question type (strategy effectiveness, contextual moderator, equity, screening pathway stage)Exploratory link-ranking: time-sliced hold-out Whether graph structure can help prioritize strategy–context–outcome links for future researchAn initial knowledge base version will be frozen with a pre-specified cut-off date (e.g., 31 December of a given year). Trials published after this date will be processed through the same workflow to form a temporal hold-out set. Candidate pairs will be strategy–context–outcome links extracted from hold-out trials but absent from the frozen knowledge base; random unseen pairs may be used only for exploratory ranking evaluation. Multiple predictors will be compared, including a shortest-path topological baseline, node2vec embedding with logistic regression, and a random-forest classifier combining structural and semantic features. These structural predictors and performance metrics serve only to prioritize strategy–context–outcome links for future research and to gauge the robustness of that prioritization. The temporal split itself will require careful definition, because an undetected overlap between a hold-out trial and a trial already represented in the frozen knowledge base (for example, a companion or duplicate publication, shared trial data, or an effectively identical strategy–context–outcome link) could leak information across the cut-off and inflate apparent performance; potential overlaps will therefore be screened and removed before evaluation.Area under the ROC curve (AUC); average precision (AP); calibration by cancer type and screening pathway stage; comparison of learned models against the topological baseline as a robustness check
• Area under the ROC curve (AUC)
• average precision (AP)
• calibration by cancer type and screening pathway stage
• comparison of learned models against the topological baseline as a robustness checkExploratory gap discovery and emerging-strategy prioritization Whether the knowledge base can prioritize emerging strategy × population × modality combinations for future research planningCandidate emerging combinations will be identified from trials published after the knowledge base cut-off. For each candidate, a knowledge-base proximity score (e.g., neighborhood Jaccard similarity or learned embedding proximity) will be computed against the frozen knowledge base, and the resulting ranking will be compared with their actual emergence in the hold-out period. Cold-start cases (candidates absent from the frozen knowledge base vocabulary) will be reported as a distinct failure mode to inform future ontology coverage decisions. These proximity scores and rankings are used only to prioritize candidate combinations for future research.Top-k precision for prioritization; proportion of emerging combinations with non-zero proximity score in the frozen knowledge base; cold-start rate; comparison against a generic-knowledge-graph baseline
• Top-k precision for prioritization
• proportion of emerging combinations with non-zero proximity score in the frozen knowledge base
• cold-start rate
• comparison against a generic-knowledge-graph baselineUsability and perceived relevance* Ability of stakeholders to navigate, interpret, and use the knowledge base for implementation decision-making Structured feedback and exploratory usability assessments will be conducted after a functional prototype, dashboard, or query interface becomes available. Participants may be asked to navigate the knowledge base, interpret evidence units, trace findings back to source studies, and identify implementation strategies or evidence gaps relevant to cancer screening implementation decision-making.Perceived usefulness, interpretability, navigability, traceability to source studies, and ability to identify implementation strategies or evidence gapsLiving-update transparency Whether the living knowledge base remains current and auditableTrack update-cycle metadata, including last search date, last screening date, number of new studies added, number of new evidence units validated, and number pending review.
In addition, workflow drift will be monitored across releases through a frozen regression benchmark, defined as a fixed held-out set of human-adjudicated records that is re-evaluated under each new release of the workflow (covering model versions, protocols, schemas, and ontology terms), consistent with the ongoing-monitoring principles articulated in living systematic review methodology18 and AI governance frameworks.40,5052,71 Pre-specified recalibration triggers include substantive model-version change; protocol or prompt revision; ontology update; output-schema change; recurrent error pattern observed in adjudication logs; high-confidence audit error rate exceeding a predefined threshold; and introduction of a new evidence type or study design.
In addition to the above release-level drift monitoring, two further validation procedures will be applied across update cycles: (i) a periodic full-search reference benchmark (e.g., annually) will be conducted to evaluate the recall of the incremental search strategy used for routine living updates, comparing the records identified by the incremental approach against those identified by a complete re-run of the original search strategy73; and (ii) the consistency of large language model outputs will be tested by re-running the same protocol and input through each LLM model multiple times and quantifying agreement across runs, following the development–test–multiple-update validation pattern recently demonstrated for large language model–assisted living systematic reviews.74Last search/update date; new records per update; newly validated evidence units; pending validation count; update lag; update status by cancer type or evidence category. Drift-benchmark performance vs. prior release; audit performance trend; high-confidence audit error rate; number of ontology terms proposed, accepted, and deprecated per release; protocol/prompt/schema change-log completeness.
Incremental search recall vs. periodic full-search reference benchmark; large language model output consistency (proportion of records receiving identical LLM decisions across repeated runs of the same protocol and input).
Discussion

This project aims to develop a living, ontology-informed, and LLM-assisted knowledge base to systematically organize evidence on implementation strategies used in cancer screening. By integrating behavior change ontologies with a human–LLM collaborative evidence review workflow, the project aims to create a standardized and continuously updatable infrastructure for understanding what implementation strategies work, for whom, under which conditions, and at what stage of the cancer screening pathway.

Advancing cancer screening implementation research through an ontology-informed living knowledge base

Despite growing interest in ontologies in biomedical53 and oncology research,54,55 especially recent efforts focused on breast cancer screening56 and cancer registry data,57 their application to cancer screening implementation remains limited. Ontology-based representations, particularly when combined with knowledge graph approaches, offer a powerful way to connect implementation strategies, mechanisms, determinants, populations, and outcomes into a unified evidence structure. This may improve the ability to synthesize fragmented implementation evidence and support more consistent interpretation across studies. This project addresses these gaps through the development of an ontology-informed living knowledge base specifically designed for cancer screening implementation research.

Opportunities and challenges of human–LLM collaborative evidence review

Generative LLM offers opportunities to improve efficiency in evidence synthesis tasks such as literature identification, screening, data extraction, and risk-of-bias assessment,58,59 but current systems still exhibit substantial error rates across tasks. A recent systematic review assessed the use of generative LLM systems including ChatGPT, GPT, Claude, Bing AI, and Perplexity AI in evidence review and reported substantial limitations across multiple review stages. For literature searching, generative AI systems failed to identify between 68% and 96% of relevant studies. During study screening, the probabilities of incorrect inclusion and incorrect exclusion ranged from 0% to 29% and 1% to 83%, respectively. Errors in data extraction ranged from 4% to 31%, and inaccurate risk-of-bias assessments ranged from 10% to 56%.20 To minimize these risks of bias, this project adopts a human–LLM collaborative approach in which LLM supports high-volume literature screening, extraction, and annotation tasks, while human reviewers will provide oversight, adjudication, and interpretation for uncertain or complex cases.60 This approach aims to balance efficiency with methodological rigor and may provide practical insights into how LLM can be responsibly integrated into future living evidence review infrastructures.61

Towards precision implementation science

A major challenge in implementation science is that implementation strategies often show variable effectiveness across populations, settings, and healthcare systems. Strategies that are successful in one context may have limited impact in another because implementation outcomes are shaped by interactions among contextual factors, delivery processes, organizational conditions, and target population characteristics.62,63 Recent implementation science initiatives are increasingly moving toward LLM-enabled and ontology-informed infrastructures designed to support more context-sensitive implementation decision-making. For example, the AIM-IS program seeks to develop frameworks and reporting standards for responsible AI use across implementation workflows.64 The ImpleMATE platform proposes ontology-based knowledge representation combined with LLM-assisted reasoning to support implementation planning and learning health system cycles.65 These emerging initiatives reflect growing interest in developing computable implementation knowledge systems to support the continuous integration and updating of implementation evidence across contexts. The ontology-informed living knowledge base proposed in this project contributes to this evolving field by enabling structured representation and continuous updating of implementation knowledge, which may support future development of more precise, context-sensitive cancer screening implementation research and decision-making in the future.66

Conclusion

The proposed knowledge base may improve the standardization, comparability, and reuse of implementation evidence by converting fragmented, inconsistently reported trial evidence into standardized, computable, and provenance-aware evidence units. By linking implementation strategies, behavior change techniques, mechanisms of action, implementation determinants, populations, settings, screening pathway stages, and outcomes, the project may also support precision implementation science with a more context-sensitive and mechanism-informed account of cancer screening implementation strategies. This project aims to establish a continuously updated foundation for knowledge integration and future implementation planning and research in cancer screening.

Ethics and consent

The ontology development component is informed by existing ontologies, published literature, cancer screening guidelines, and concepts identified from included studies. This phase does not involve the collection of data from human participants and therefore does not require Institutional Review Board (IRB) review. The knowledge base construction component of this project involves the analysis of publicly available published literature and does not involve human participants or identifiable private information. Therefore, IRB review was not required for this component. Future phases involving stakeholder engagement, expert consultation, Delphi consensus procedures, usability testing, or other activities involving human participants to further develop, refine, validate, or evaluate the ontology and knowledge base will be submitted for IRB review and approval prior to commencement.

Consent for publication

Not applicable.

Declaration of generative AI and AI-assisted technologies in the writing process

No generative AI tools were used to generate scientific content in this protocol. The use of large language models described in this protocol refers to the planned research workflow and not to the generation of the manuscript content. The use of AI was limited to language refinement of author-drafted text to improve clarity and readability. No AI tools were used to create or modify any figures or images. The use of large language models described in this protocol refers to the planned research workflow and not to the generation of the manuscript content. All content was reviewed, revised, and approved by the authors, who take full responsibility for the final manuscript.

Схожие новости

#Наименование новостиТональностьИнформативностьДата публикации
1AI-guided outreach increased cancer screenings and reduced mortality, new study finds5807-07-2026
2An Ontology‑Guided Drug–Herb–Food Interaction Checker with Mechanism‑Based Knowledge Graph Reasoning and Condition‑Aware Interpretation [version 2; peer review: 1 approved with reservations]014.6217-07-2026
3Diagnostic Performance of Computed Tomography-Based Machine Learning Models in the Classification of Adnexal Masses - A Systematic Review [version 1; peer review: 2 approved]0802-04-2026
4Heath Care Providers’ Perspectives on Pregnant Women’s use of Internet Based Health Information: A Qualitative Study [version 1; peer review: awaiting peer review]09.0101-08-2026
5AI in Clinical Trial Design for Study Teams and Sponsors011.7830-07-2026
6Leveraging Artificial Intelligence (AI) Based Algorithm for Accurate Estrogen Receptor (ER) and Progesterone Receptor (PR) Analysis in Breast Cancer Diagnostics: Potential to be a Crucial Aid in Routine Workflow [version 2; peer review: 2 approved]08.4403-08-2026
7 AI Not Automatically Better for Colorectal Cancer Screening Lancet study from Bonn on Lynch syndrome0717-07-2026
8Uncertainty-aware AI and lensfree holography enable reliable automated HER2 assessment for breast cancer diagnostics5710-06-2026
9Dataset of multi-focus (Z-stack) images derived from liquid-based cervical cancer cytology specimens [version 1; peer review: 2 approved]013.8510-04-2026
10KI verbessert Darmkrebsvorsorge nicht automatisch Lancet-Studie aus Bonn zu Menschen mit Lynch-Syndrom0717-07-2026

Классификация: Наука. Схожих патентов: 0. Схожих новостей: 10. Тональность: 0. Информативность: 9.16. Источник: f1000research.com.