Status: question. Rule: atlas-profile-2.0.
5 linked publications; not a count of independent studies.
Missing appraisal does not mean no evidence exists. Inspect the full profile and source relationships in the interactive Atlas.
Historical grade: WEAK (retained from 22 September 2026; not a completed profile 2.0 assessment).
These models are entering clinical view with fluent output but unquantified safety; premature trust is a patient-safety risk.
Generalist-AI promise vs GPT-4V scoring 47.8% on image-based radiology questions; failure characterization for neuro is essentially absent.
Adversarial and prospective evaluation on curated neuro-oncology cases; report calibration, abstention, and error taxonomy, not just accuracy.
This question comes from selected literature. Confirm that the gap remains open with a current search and mentor review.
This is an educational research resource from Resonant Labs, not clinical advice. Atlas assessments are heuristic, not probabilities or formal GRADE ratings. Findings are summarized from the literature and may change as the field evolves.