• FluTrackers.com Inc. does not provide medical advice. Information on this web site is collected from various internet resources, and the FluTrackers board of directors makes no warranty to the safety, efficacy, correctness or completeness of the information posted on this site by any author or poster. The information collated here is for instructional and/or discussion purposes only and is NOT intended to diagnose or treat any disease, illness, or other medical condition. Every individual reader or poster should seek advice from their personal physician/healthcare practitioner before considering or using any interventions that are discussed on this website. By continuing to access this website you agree to consult your personal physican before using any interventions posted on this website, and you agree to hold harmless FluTrackers.com Inc., the board of directors, the members, and all authors and posters for any effects from use of any medication, supplement, vitamin or other substance, device, intervention, etc. mentioned in posts on this website, or other internet venues referenced in posts on this website.
  • We are not asking for any donations. Do not donate to any entity who says they are raising funds for us.

One Health . Multi-model large-scale AI framework for avian influenza surveillance and preparedness: Harnessing large language models to enhance ris

tetano

Editor, Senior Moderator
One Health


. 2026 Feb 10:22:101357.
doi: 10.1016/j.onehlt.2026.101357. eCollection 2026 Jun.
Multi-model large-scale AI framework for avian influenza surveillance and preparedness: Harnessing large language models to enhance risk communication, real-time decision support, and public health response strategies

Jude D Kong[SUP] 1 2 3 4 5 [/SUP], Murray Gillies[SUP] 6 [/SUP], Emma Gardner[SUP] 6 [/SUP], Nicola Luigi Bragazzi[SUP] 2 3 7 [/SUP]


Affiliations
Abstract

Avian influenza remains a persistent threat to global health security, with serious consequences for food systems, trade, and pandemic preparedness. To address gaps in public health communication and stakeholder-specific decision-making, this study evaluated the capacity of large language models (LLMs) to provide accurate, context-sensitive, and ethically sound guidance in the context of avian influenza. Employing a multi-model, stakeholder-stratified evaluation framework, we tested four advanced generative AI models, namely, ChatGPT-4o (OpenAI), Grok (xAI), Gemini 1.5 Pro (Google), and DeepSeek R1 (DeepSeek), across two complementary tasks: (i) structured querying of 34 domain-specific items covering virological, epidemiological, veterinary, and global public health domains; and (ii) response generation to 16 synthetic vignettes simulating outbreak scenarios involving diverse societal roles. Results showed that Gemini 1.5 Pro demonstrated the highest factual accuracy (91.2% fully correct responses), followed by Grok (85.3%), ChatGPT-4o (82.4%), and DeepSeek R1 (82.4%). Vignette analysis further revealed model-specific communicative strengths and ethical orientations, ranging from procedural pragmatism to stakeholder-mapping and human-centered design. While Gemini excelled in blending empathetic and pedagogical reasoning, Grok offered implementation-oriented guidance, ChatGPT-4o emphasized legal-normative clarity, and DeepSeek R1 favored structural and institutional analysis. Collectively, our findings highlight the promise and limitations of current LLMs as tools for biosurveillance, risk communication, and cross-sectoral pandemic preparedness. They also emphasize the necessity of rigorous, role-aware benchmarking to ensure equitable and contextually appropriate integration of generative AI in public health infrastructures.

Keywords: Avian influenza; Generative artificial intelligence; Large language models; One health; Vignette analysis.

 
Back
Top