One paper Let large language models judge each other: multi-agent peer-reviewed reasoning for medical question answering just got accepted at Journal of the American Medical Informatics Association