Test the large language models doctors and patients are already using against a continually refreshed set of cancer questions, scoring not just correct answers but whether the sources they cite are real and support the claim.
Clinicians and patients use general-purpose language models for oncology questions; evaluations are static, quickly outdated and rarely check citations. The proposal is a living benchmark: new questions each month drawn from recent practice changes, expert-graded answers, and scoring of citation validity and support, with public leaderboards and per-cancer breakdowns, run by an independent academic consortium.
Shares Misinformation and unproven therapies, Knowledge reaches practice too slowly.
Shares Misinformation and unproven therapies, Knowledge reaches practice too slowly.
Shares The answer is 17 years, what is the question: understanding time lags in translational research, Knowledge reaches practice too slowly.
Shares The answer is 17 years, what is the question: understanding time lags in translational research, AI that is built but not validated or deployed, Knowledge reaches practice too slowly.
Shares The answer is 17 years, what is the question: understanding time lags in translational research, Misinformation and unproven therapies, Knowledge reaches practice too slowly.
Shares Misinformation and unproven therapies, Knowledge reaches practice too slowly.
Shares The answer is 17 years, what is the question: understanding time lags in translational research, Knowledge reaches practice too slowly.
Shares AI that is built but not validated or deployed, Misinformation and unproven therapies.