Copyright: ©Author(s) 2026.
World J Methodol. Sep 20, 2026; 16(3): 116022
Published online Sep 20, 2026. doi: 10.5662/wjm.v16.i3.116022
Published online Sep 20, 2026. doi: 10.5662/wjm.v16.i3.116022
Table 7 Categorical patient assessment of responses from ChatGPT-5, Gemini-2.5, and Claude-4
| Characteristics | ChatGPT-5a | Gemini-2.5a | Claude-4a | P value |
| Comprehensiveness | ||||
| Minimal | 3 (7.7) | 3 (7.7) | 2 (5.1) | < 0.05 |
| Moderately | 29 (74.4) | 24 (61.5) | 23 (59.0) | |
| Highly | 7 (17.9) | 13 (33.3) | 14 (38.9) | |
| Actionability | ||||
| Slightly actionable | 8 (22.2) | 1 (2.6) | 2 (5.1) | < 0.05 |
| Moderately actionable | 27 (69.2) | 13 (33.3) | 18 (46.2) | |
| Actionable | 2 (5.1) | 18 (46.2) | 16 (41.0) | |
| Highly actionable | 2 (5.1) | 7 (17.9) | 3 (7.7) | |
| Empathy | ||||
| Minimally empathetic | 15 (38.5) | 1 (2.6) | 1 (2.6) | < 0.05 |
| Moderately empathetic | 21 (53.8) | 23 59.0) | 24 (61.5) | |
| Highly empathetic | 3 (8.3) | 15 (38.5) | 14 (35.9) | |
- Citation: Goyal K, Goyal MK, Taranikanti V, Wander P, Chowdhary R, Kalra S, Prashar G, Vuthaluru AR, Goyal O. Do bots provide correct and adequate guidance regarding acidity: A blinded comparison rated by patients and physicians. World J Methodol 2026; 16(3): 116022
- URL: https://www.wjgnet.com/2222-0682/full/v16/i3/116022.htm
- DOI: https://dx.doi.org/10.5662/wjm.v16.i3.116022