Episode 82: Interview of Dr.Marcello Di Martino

Episode 82: Interview of Dr.Marcello Di Martino

Humanities, Social Sciences & Thought Medicine & Health MBMedicineMBFMedical and health informatics
🎙 Artificial Intelligence Surgery 👥 55 📅 May 12, 2026 ⏱ 30 min 👁 32 📄 interview 🧭 2026-08-16
Available in: English (current) Français

Keywords

AIsurgeryHPBChatGPTtransplant

Summary

In this podcast episode, hosts Andrew Gums and Vincent Grasso interview Dr. Marcello Di Martino, an HPB surgeon from Italy. Dr. Di Martino shares his career journey, including training in Spain, fellowships in the UK and US, and current academic position in Novara. The discussion covers his involvement with IHPBA and the early career group. The conversation shifts to AI in surgery, with Dr. Di Martino expressing cautious optimism about AI as a support tool but highlighting risks like hallucinated references and the importance of prompt engineering. They discuss a study on ChatGPT underestimating emergencies and the dangers of patients using AI for diagnosis. Dr. Di Martino mentions his research on acute pancreatitis bundles and a systematic review comparing human vs. AI performance. The episode also touches on AI applications in transplant, such as using photos to assess liver quality, and the potential of merging imaging with functional data. The hosts emphasize the need for domain expertise when using AI and the importance of communication in surgical care.

168 words

Critical Evaluation

Value of the Information & Strength of the Argument

The value of the information lies in the firsthand experience of a practicing HPB surgeon with AI tools. Dr. Di Martino provides practical insights into how AI can support clinical and research tasks, such as using ChatGPT for grammar review or creating bundles for acute pancreatitis. He also highlights limitations, including the unreliability of AI-generated references and the need for careful prompt engineering. The argumentation is anecdotal and based on personal experience rather than systematic evidence. The hosts contribute to the discussion by citing a study on ChatGPT’s underestimation of emergencies and the risks of patient self-diagnosis. The conversation is engaging but lacks depth in scientific detail, with no specific data or studies cited beyond general references.

Scientific Rigor, Source Quality, Title Accuracy

The scientific rigor is moderate. The discussion is based on personal experience and general knowledge rather than cited studies. The hosts mention a study from The Guardian but do not provide specific details or a direct link. Dr. Di Martino mentions a paper on acute pancreatitis bundles published in Annals, but no citation is given. The title accurately reflects the content, as it is an interview with Dr. Di Martino. The podcast is associated with the journal ‘Artificial Intelligence Surgery’, which adds some credibility, but the episode itself lacks formal references. The conversation is more of a casual discussion than a rigorous scientific review.

236 words

Title / Content Match

The title accurately reflects the content: an interview with Dr. Marcello Di Martino.

Quality & Reliability

6/10

The interview features an experienced HPB surgeon discussing AI applications in surgery, but it is largely anecdotal and lacks detailed scientific evidence or citations.

Key Moments

Cited Sources

Concurring Sources

Contribution & Novelties

The interview provides a personal perspective on the integration of AI in HPB surgery, highlighting practical uses and limitations. It emphasizes the importance of domain expertise in prompt engineering and the risks of AI-generated misinformation. The discussion on using smartphone photos to assess liver quality during retrieval is a novel practical application.

Pour aller plus loin :

  • Artificial Intelligence in Surgery — Overview of AI applications in surgery.
  • Large language models in medicine — Review of LLMs in healthcare.
  • Radiomics — Concept related to extracting quantitative features from medical images.

90 words

Radar Profile

The radar profile shows moderate scores across all dimensions, indicating a balanced but not deeply technical discussion. The highest score is in fiabilite_globale, reflecting the credibility of the speakers, while quantite_information is lower due to the conversational nature.

Reliability 6/10

💬 No comments were provided for analysis.