Evaluation of the accuracy and reproducibility of large language models (ChatGPT, DeepSeek, Gemini) in responding to patient-centered lipedema questions.
LLMs can provide generally accurate and consistent responses to patient-centered questions about lipedema, particularly in areas related to general information and diagnosis, however, reduced accuracy and reproducibility in complex clinical domains suggest that expert oversight is essential when using these tools for p...