
Health & Medicine
Moderate evidenceReview
Study compares two AI systems’ accuracy and usability for anesthesiology crisis scenarios
This review-style evaluation compares two large language models—OpenAI o1 and DeepSeek R1—in English and Chinese for supporting junior anesthesiologists. Using 30 Delphi-created anesthesia crisis scenarios, experts rated accuracy and logic, while junior physicians rated clarity and usefulness. Results differed: OpenAI o1 scored higher for accuracy, while the DeepSeek R1 Chinese version scored higher for practical, step-by-step guidance.
·19 views