#1877·Qwen3

China Turkey open benchmark for clinical AI safety

Author: goktugozkanmdCreated Jul 8, 2026Updated Aug 20, 2026

Hello Qwen team,

I am a physician from Turkey building MedFailBench, an open source clinical AI safety benchmark based on synthetic clinician authored cases.

Project: https://github.com/goktugozkanmd/medical-ai-failure-atlas

Qwen is central to the Chinese open model ecosystem, and I want this collaboration to include leading Chinese frontier model families from the start.

The proposal is a China Turkey clinical AI safety collaboration: shared Turkish, English, and Chinese safety cases, reproducible evaluations, and a coauthored report or paper on clinical safety boundaries in medical LLM outputs.

I can bring clinician authored cases, Turkish clinical wording risk, a safety gate taxonomy, and an open source evaluation pipeline.

No patient data is involved. This is research infrastructure for safer medical LLM evaluation, not clinical advice or deployment.

Would your team be open to a short call or async discussion about a joint pilot including Qwen models?

Best, Goktug Ozkan