Already a member or subscriber? Sign in now

Use of AI in Family Medicine Publications: A Joint Editorial From Journal Editors

SARINA SCHRAGER, MD, MS
DEAN A. SEEHUSEN, MD, MPH
SUMI SEXTON, MD
CAROLINE RICHARDSON, MD
JON NEHER, MD
NICHOLAS J. PIMLOTT, MD
MARJORIE A. BOWMAN, MD
JOSÉ RODRÍGUEZ, MD
CHRISTOPHER P. MORLEY, PhD
LI LI, MD, PhD, MPH
JAMES DOMDERA, MD

FPM. 2025;32(1):6-9.

Author disclosures: no relevant financial relationships reported. Note: This article is being published simultaneously in Family Medicine, JABFM, AFP, Annals of Family Medicine, FPIN, CFP, PRiMER, FMCH, FP Essentials, and FPM.

There are multiple guidelines from publishers and organizations on the use of artificial intelligence (AI) in publishing.1–5 However, none are specific to family medicine. Most journals have some basic AI use recommendations for authors, but more explicit direction is needed, as not all AI tools are the same.

As family medicine journal editors, we want to provide a unified statement about AI in academic publishing for authors, editors, publishers, and peer reviewers based on our current understanding of the field. The technology is advancing rapidly. While text generated from early large language models (LLMs) was relatively easy to identify, text generated from newer versions is getting progressively better at imitating human language and more challenging to detect. Our goal is to develop a unified framework for managing AI in family medicine journals. As this is a rapidly evolving environment, we acknowledge that any such framework will need to continue to evolve. However, we also feel it is important to provide some guidance for where we are today.

Definitions: Artificial intelligence is a broad field where computers perform tasks that have historically been thought to require human intelligence. LLMs are a recent breakthrough in AI that allow computers to generate text that seems like it comes from a human. LLMs deal with language generation, while the broader term generative AI can also include AI generated images or figures. ChatGPT is one of the earliest and most widely used LLM models, but other companies have developed similar products. LLMs “learn” to do a multifaceted analysis of word sequences in a massive text training database and generate new sequences of words using a complex probability model. The model has a random component, so responses to the exact same prompt submitted multiple times will not be identical. LLMs can generate text that looks like a medical journal article in response to a prompt, but the article's content may or may not be accurate. LLMs may “confabulate,” generating convincing text that includes false information.6,7,8 LLMs do not search the internet for answers to questions. However, they have been paired with search engines in increasingly sophisticated ways. For the rest of this editorial, we will use the broad term AI synonymously with LLMs.

ROLE OF LARGE LANGUAGE MODELS IN ACADEMIC WRITING AND RESEARCH

As LLM tools are updated and authors and researchers become familiar with them, they will undoubtedly become more functional in assisting the research and writing process by improving efficiency and consistency. However, current research on the best use of these tools in publication is still lacking. A systematic review exploring the role of ChatGPT in literature searches found that most articles on the topic are commentaries, blog posts, and editorials, with little peer-reviewed research.9 Some studies have demonstrated benefit in narrowing the scope of literature review when AI tools were applied to large data sets of studies and prompted to evaluate them for inclusion based on the title and abstract. Another paper reported that AI had 70% accuracy in appropriately identifying relevant studies compared with human researchers and may reduce time and provide a less subjective approach to literature review.10,11,12 When used to assist with writing background sections, LLMs' writing was rated the same if not better than human researchers, but the citations were consistently false in another study.13 LLM models are frequently deficient in providing “real” papers and correctly matching authors to their own papers when generating citations and therefore are at risk of creating fictitious citations that appear convincing despite incorrect information including digital object identifier (DOI) numbers.6,14,

Dr. Schrager is Editor in Chief, Family Medicine, and is with the Department of Family Medicine and Community Health, University of Wisconsin School of Medicine and Public Health, Madison, Wisc.

Dr. Seehusen is Deputy Editor, Journal of the American Board of Family Medicine, and is with the Department of Family and Community Medicine, Medical College of Georgia, Augusta University, Augusta, Ga.

Dr. Sexton is Editor in Chief, American Family Physician, and is with Georgetown University School of Medicine, Washington, D.C.

Dr. Richardson is Editor in Chief, Annals of Family Medicine, and is with Warren Alpert Medical School, Brown University, Providence, R.I.

Dr. Neher is Editor in Chief, Family Physicians Inquiries Network, and is with Valley Family Medicine Residency Program, Renton, Wash.

Dr. Pimlott is Scientific Editor, Canadian Family Physician, and is with the Department of Family and Community Medicine, University of Toronto, Ontario, Canada.

Dr. Bowman is Editor in Chief, Journal of the American Board of Family Medicine, and is with the Veteran's Health Administration, Washington, D.C.

Dr. Rodríguez is Deputy Editor, Family Medicine, and is with the Department of Family and Preventive Medicine, Spencer Fox Eccles School of Medicine, University of Utah Health, Salt Lake City.

Dr. Morley is Editor in Chief, PRiMER, and is with the Department of Public Health and Preventive Medicine and Family Medicine, SUNY Upstate Medical University, Syracuse, N.Y.

Dr. Li is Editor in Chief, Family Medicine and Community Health, and is with the Department of Family Medicine, University of Virginia, Charlottesville, Va.

Dr. DomDera is Medical Editor, FPM, and is with Pioneer Physicians Network, Fairlawn, Ohio.

Author disclosures: no relevant financial relationships reported. Note: This article is being published simultaneously in Family Medicine, JABFM, AFP, Annals of Family Medicine, FPIN, CFP, PRiMER, FMCH, FP Essentials, and FPM.

  1. 1.Zielinski C, Winker MA, Aggarwal R, et al. Chatbots, generative AI, and scholarly manuscripts WAME recommendations on chatbots and generative artificial intelligence in relation to scholarly publications. World Association of Medical Editors. Published Jan. 20, 2023. Revised May 31, 2023. Accessed Oct. 3, 2024. https://wame.org/page3.php?id=106
  2. 2.Recommendations for the conduct, reporting, editing, and publication of scholarly work in medical journals. International Committee of Medical Journal Editors. Updated Jan. 2024. Accessed Oct. 3, 2024. https://www.icmje.org/icmje-recommendations.pdf
  3. 3.Adams L, Fontaine E, Lin S, Crowell T, Chung VCH, Gonzalez AA, eds. Artificial intelligence in health, health care, and biomedical science: an AI code of conduct principles and commitments discussion draft. National Academy of Medicine. Published April 8, 2024.
  4. 4.COPE position statement - authorship and AI tools. Committee on Publication Ethics. Published Feb 13, 2023. Accessed Oct. 3, 2024. https://publicationethics.org/cope-position-statements/ai-author
  5. 5.Author instructions New England Journal of Medicine, Editorial Policies. Accessed Oct. 3, 2024. https://www.nejm.org/about-nejm/editorial-policies
  6. 6.Haider J, Söderström K, Ekström B, Rödl M. GPT-fabricated scientific papers on Google Scholar: Key features, spread, and implications for preempting evidence manipulation. Harvard Kennedy School Misinformation Review. 2024;5(5).
  7. 7.Ramoni D, Sgura C, Liberale L, Montecucco F, Ioannidis JPA, Carbone F. Artificial intelligence in scientific medical writing: legitimate and deceptive uses and ethical concerns. Eur J Intern Med. 2024;127:31-35.
  8. 8.Brender TD. Chatbot confabulations are not hallucinations-reply. JAMA Intern Med. 2023;183(10):1177-1178.
  9. 9.Parisi V, Sutton A. The role of ChatGPT in developing systematic literature searches: an evidence summary. J Eur Assoc Health Inf Libr. 2024;20(2):30-34.
  10. 10.Zimmermann R, Staab M, Nasseri M, Brandtner P. Leveraging large language models for literature review tasks - a case study using ChatGPT. In: Guarda T, Portela F, Diaz-Nafria JM, eds. Advanced Research in Technologies, Information, Innovation and Sustainability. ARTIIS 2023. Communications in Computer and Information Science. Vol 1935. Springer; 2024.
  11. 11.Dennstädt F, Zink J, Putora PM, Hastings J, Cihoric N. Title and abstract screening for literature reviews using large language models: an exploratory study in the biomedical domain. Syst Rev. 2024;13(1):158.
  12. 12.Guo E, Gupta M, Deng J, Park YJ, Paget M, Naugler C. Automated paper screening for clinical reviews using large language models: data analysis study. J Med Internet Res. 2024;26:e48996.
  13. 13.Huespe IA, Echeverri J, Khalid A, et al. Clinical research with large language models generated writing-clinical research with AI-assisted writing (CRAW) Study. Crit Care Explor. 2023;5(10):e0975.
  14. 14.Byun C, Vasicek P, Seppi K. This Reference Does Not Exist: An Exploration of LLM Citation Accuracy and Relevance. In: Proceedings of the Third Workshop on Bridging Human–Computer Interaction and Natural Language Processing. Association for Computational Linguistics; 2024:28–39.
  15. 15.Chemaya N, Martin D. Perceptions and detection of AI use in manuscript preparation for academic journals. PLoS One. 2024;19(7):e0304807.
  16. 16.Gödde D, Nöhl S, Wolf C, et al. A SWOT (strengths, weaknesses, opportunities, and threats) analysis of ChatGPT in the medical literature: concise review. J Med Internet Res. 2023;25:e49368.
  17. 17.Directorate-General for Research and Innovation. Living guidelines on the responsible use of generative AI in research - from the European Commission. European Commission; March 2024. Accessed Oct. 3, 2024. https://research-and-innovation.ec.europa.eu/document/download/2b6cf7e5-36ac-41cb-aab5-0d32050143dc_en?filename=ec_rtd_ai-guidelines.pdf
  18. 18.Kaebnick GE, Magnus DC, Kao A, et al. Editors' statement on the responsible use of generative AI technologies in scholarly journal publishing. Med Health Care Philos. 2023;26(4):499-503.
  19. 19.Saguil A. Chatbots and large language models in family medicine. Am Fam Physician. 2024;109(6):501-502.
  20. 20.Andrew A. Potential applications and implications of large language models in primary care. Fam Med Community Health. 2024;12(Suppl 1):e002602.
  21. 21.Our mission is to work with companies, policy makers and experts to reduce bias in our AI. EqualAI. Accessed Oct. 3, 2024. https://www.equalai.org/about-us/mission/
  22. 22.Manyika J, Silberg J, Presten B. What do we do about the biases in AI? Harvard Business Review; 2019. Accessed Oct. 3, 2024. https://hbr.org/2019/10/what-do-we-do-about-the-biases-in-ai
  23. 23.About. Algorithmic Justice League. Accessed Oct. 3, 2024. https://www.ajl.org/about
  24. 24.Elkhatat AM, Elsaid K, Almeer S. Evaluating the efficacy of AI content detection tools in differentiating between human and AI-generated text. Int J Educ Integr. 2023;19(1):17.
  25. 25.Ranjbari D, Abbasgholizadeh Rahimi S. Implications of conscious AI in primary healthcare. Fam Med Community Health. 2024;12(Suppl 1):e002625.
  26. 26.Parente DJ. Generative artificial intelligence and large language models in primary care medical education. Fam Med. 2024;56(9):534-540.
  27. 27.Waldren SE. The promise and pitfalls of AI in primary care. Fam Pract Manag. 2024;31(2):27-31.
  28. 28.Hanna K, Chartash D, Liaw W, et al. Family medicine must prepare for artificial intelligence. J Am Board Fam Med. 2024;37(4):520-524.
  29. 29.Hanna K. Exploring the applications of ChatGPT in family medicine education: five innovative ways for faculty integration. PRiMER. 2023;7:26.
  30. 30.Kueper JK, Terry AL, Zwarenstein M, Lizotte DJ. Artificial intelligence and primary care research: a scoping review. Ann Fam Med. 2020;18(3):250-258.
  31. 31.Hake J, Crowley M, Coy A, et al. Quality, accuracy, and bias in ChatGPT-based summarization of medical abstracts. Ann Fam Med. 2024;22(2):113-120.
  32. 32.Kueper JK, Emu M, Banbury M, et al. Artificial intelligence for family medicine research in Canada: current state and future directions: report of the CFPC AI Working Group. Can Fam Physician. 2024;70(3):161-168.
  33. 33.Sheng E, Chang K-W, Natarajan P, Peng N. The woman worked as a babysitter: on biases in language generation. In: Proceedings of the 2019 Conference on Empirical Methods in Natural Language Processing and the 9th International Joint Conference on Natural Language Processing (EMNLP-IJCNLP). Association for Computational Linguistics; 2019:3407-3412.

WE WANT TO HEAR FROM YOU

The opinions expressed here do not necessarily represent those of FPM or our publisher, the American Academy of Family Physicians. We encourage you to share your views. Send comments to fpmedit@aafp.org, or add your comments below.

Copyright © 2026 by the American Academy of Family Physicians.

This content is owned by the AAFP. A person viewing it online may make one printout of the material and may use that printout only for his or her personal, non-commercial reference. This material may not otherwise be downloaded, copied, printed, stored, transmitted or reproduced in any medium, whether now known or later invented, except as authorized in writing by the AAFP. See permissions for copyright questions and/or permission requests.