Since 2024, Large Language Models have become the standard tool used for automated marking of physics exams, especially for hand-written exams and questions which involve diagrams. Nobody has tested them on astronomy questions where students annotate a projection of the night sky. This project benchmarks LLMs against classical computer-vision methods on marking constellation and object identification tasks.
Mr Lachlan McGinness