Doctor&Cols — colectivo de mentoría dental
Producción científica2025· International Journal of Dentistry

Comparing Manual and ChatGPT Deep Research on Systematic Search and Selection in the PubMed Database on the Topic of Dental Implantology

Q3Impact Factor · 1.458

AutoresBulcsú Bencze, Alwin Sokolowski, Jae-Hyun Lee, Péter Hermann, Tamás Hegedüs, Wataru Kozuma, Reo Ikumi, Michael Payer, Ángel-Orión Salgado-Peralvo, Dániel Végh

Por qué importa

Este estudio compara la eficacia de la búsqueda e investigación profunda realizada por ChatGPT 4.1 frente a la metodología manual tradicional en la base de datos PubMed, centrándose en la implantología dental. Se evaluaron innovaciones recientes, incluyendo diseño de implantes y osteointegración. Los resultados indican que, si bien las herramientas de IA como ChatGPT pueden asistir en la síntesis preliminar y la mejora de la legibilidad, su fiabilidad actual es insuficiente para reemplazar la investigación manual en revisiones sistemáticas. La generación de referencias inexistentes y la omisión de estudios relevantes subrayan la necesidad de supervisión humana y una cautela estricta en su aplicación clínica, especialmente para mantenerse al día con los avances tecnológicos en el campo.

Temas tratados

Abstract original

Introduction: Dental implantology has seen rapid technological advancements, with artificial intelligence (AI) increasingly integrated into diagnostic, planning, and surgical processes. The release of chat-generative pretrained transformer (ChatGPT) and its subsequent updates, including the deep research function, presents opportunities for AI-assisted systematic reviews. However, its efficacy compared to traditional manual research has not been researched. Materials and Methods: A systematic review was conducted on May 6, 2025, to evaluate recent innovations in dental implantology and AI. Two parallel searches were performed: one using ChatGPT 4.1's deep research tool in the PubMed database and another manual PubMed search by two independent researchers. Both searches used identical keywords and Boolean operators targeting studies from 2020 to 2025. Inclusion criteria were peer-reviewed studies related to implant design, osseointegration, guided placement, and other predefined outcomes. Results: The manual search identified 124 articles, of which 23 met the inclusion criteria. ChatGPT retrieved 114 articles, selected 13 for inclusion, yet only included 11 in its analyses. Two cited articles by the AI software were nonexistent, and numerous relevant studies were not retrieved, whereas the remaining articles were correct and found by manual search as well. ChatGPT had high specificity (98%) and low sensitivity (47.8%), with a statistically significant difference compared to manual search and selection. Discussion: AI tools like ChatGPT show promise in literature search, synthesis, and assistance, especially in improving readability and identifying trending topics in science. Nevertheless, the current state of deep research function lacks the reliability required for conducting systematic reviews due to issues such as made-up references and missed articles. The results highlight the need for human supervision and improved safeguards. Conclusions: ChatGPT's deep research function can support, but not replace manual systematic search and selection. It offers substantial benefits in writing support and preliminary synthesis due to acceptable accuracy, but limitations in reliability and low sensitivity (47.8%) require cautious use and transparent reporting of any AI involvement in scientific research.

Autor Doctor&Cols

Cómo citar

Bencze B, Sokolowski A, Lee J, Hermann P, Hegedüs T, Kozuma W, et al. Comparing Manual and ChatGPT Deep Research on Systematic Search and Selection in the PubMed Database on the Topic of Dental Implantology. International Journal of Dentistry. 2025. doi:10.1155/ijod/2677641

Casos clínicos relacionados

Más producción científica del autor

Ver toda la producción científica →
Educational Partners