criticalHallucination OpenAI (ChatGPT)Community Report
BMJ Open audit: 5 major AI chatbots failed medical misinformation test - 19.6% highly problematic
by Publish anonymously · Jul 4, 2026views 10en
PII protected
Personal information such as emails, phone numbers, IDs and access tokens are automatically masked before publication.
Peer-reviewed BMJ Open study tested ChatGPT, Gemini, Meta AI, Grok, and DeepSeek on 250 medical prompts. 19.6% of responses were highly problematic, 30% somewhat problematic. Only 2 of 250 prompts refused. All chatbots hallucinated medical citations. Nutrition queries performed worst.
Chain of Evidence
Record Opened
Incident recorded and encrypted by user.
Awaiting Defendant Response
The AI company was notified and we are awaiting an official response.
Awaiting official response
The official response from the AI provider for this incident.
Community Discussion
0No comments yet. Start the discussion!
You must sign in to join the discussion.