• Home
  • Search
  • Extracting circumstances of Covid-19 transmission from free text with large language models
  • Cite Icon1
  • https://doi.org/10.1038/s41467-025-60762-wCopy DOI Icon

Extracting circumstances of Covid-19 transmission from free text with large language models

Show More
  • Abstract
  • Literature Map
  • References
  • Citations
  • Similar Papers
Abstract

Identifying the circumstances of transmission of an emerging infectious disease rapidly is central for mitigation efforts. Here, we explore how large language models (LLMs) can automatically extract such circumstances from free-text descriptions in online surveys, in the context of Covid-19. In a nationwide study conducted online in France, we enrolled 545,958 adults with recent SARS-CoV-2 infection and inquired about the circumstances of transmission in both closed-ended and open-ended questions. First, we trained a classification model based on a pretrained LLM to predict one of seven predefined infection contexts (Work, Family, Friends, Sports, Cultural, Religious, Other) from the free text in answers to open-ended questions. We achieved an unbalanced accuracy of 75%, which increased to 91% when eliminating the 43% highest entropy responses. Second, we used topic modeling to define clusters of transmission circumstances agnostically. This led to 23 clusters, which agreed with the seven predefined infection contexts, but also provided finer details on previously undefined circumstances of transmission. Our study suggests that LLM-based analysis of free text may alleviate the need for closed-ended questions in epidemiological surveys and enable insights into previously unsuspected circumstances of transmission. This approach is poised to accelerate and enrich the acquisition of epidemiological insights in future pandemics.

Similar Papers
  • Research Article

Unlocking the Potential of Large Language Models in Education: Factors Influencing Adoption by Instructional Designers and Academics

  • Jan 01, 2026
  • Journal of Information Technology Education: Research
  • Katherine L Fourie +2
  • Research Article

Evaluating gpt-4 for zero-shot classification of bleeding and clotting events: Can large language models serve as second reviewers?

  • Nov 03, 2025
  • Blood
  • Samantha Rizzo +5
  • Supplementary Content

Large language and vision-language models for robot: safety challenges, mitigation strategies and future directions

  • Jul 29, 2025
  • Industrial Robot: the international journal of robotics research and application
  • Xiangyu Hu +1
  • Research Article
  • Citations1

CUPCase: Clinically Uncommon Patient Cases and Diagnoses Dataset

  • Apr 11, 2025
  • Proceedings of the AAAI Conference on Artificial Intelligence
  • Oriel Perets +3
  • Research Article

Neuro-symbolic AI for auditable cognitive information extraction from medical reports.

  • Nov 21, 2025
  • Open Access CRIS of the University of Bern
  • Prenosil, George A +8
  • Research Article

Research and selection of Large Learning Models for automation of ABAP-code migration

  • Sep 24, 2025
  • Management of Development of Complex Systems
  • Oleg Pozdnyakov +1
  • Research Article

Readability & quality of large language model responses in CAR-T patient education

  • Nov 03, 2025
  • Blood
  • Sridhar Balasubramanian +3
  • Research Article

Das Urheberrecht als (KI‑)Innovationsbremse in der Rechtswissenschaft?

  • Jan 01, 2025
  • Zeitschrift für geistiges Eigentum
  • Tristan Radtke
  • Research Article

Application of Large Language Models (LLMs) to Geriatric Practice and Its Evaluation at 4 VA GRECCs

  • Dec 01, 2025
  • Innovation in Aging
  • Huai Cheng +2
  • Research Article

1401 Bringing Medicine Expertise to Your Screen: A New Frontier in Curbside Sleep Consultation Leveraging Large Language Models?

  • May 19, 2025
  • SLEEP
  • Nina Kuei +4
  • Front Matter
  • Citations1

Editorial: Large language models in work and business.

  • Nov 29, 2024
  • Frontiers in artificial intelligence
  • Şadi Evren Şeker
  • Research Article
  • Citations15

Large language models in neurosurgery: a systematic review and meta-analysis.

  • Nov 23, 2024
  • Acta neurochirurgica
  • Advait Patil +5
  • Research Article

#2924 Comparison of large language models and traditional natural language processing techniques in predicting arteriovenous fistula failure

  • May 23, 2024
  • Nephrology Dialysis Transplantation
  • Suman Lama +6
  • Research Article
  • Citations1

Optimization of traditional methods for determining the similarity of project names and purchases using large language models

  • Apr 01, 2024
  • Litera
  • Aleksei Aleksandrovich Golikov +2
  • Supplementary Content
  • Citations114

Applications and Concerns of ChatGPT and Other Conversational Large Language Models in Health Care: Systematic Review

  • Nov 07, 2024
  • Journal of Medical Internet Research
  • Leyao Wang +7
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.