AI chatbots exhibit heterogeneous reference integrity, with risks of hallucinations and biases underscoring the need for prompt engineering, model refinements and ongoing evaluation. Our findings suggest ongoing caution is required in surgical contexts to ensure safe, equitable information dissemination.
Sidhu et al. (Wed,) studied this question.