Back to Research papers
Research paper index

Assessing AI in Introductory Physics Problem Solving

Amir Bralin, N. Sanjay Rebello

arXiv:2607.14303Published July 15, 20260 citations
  • physics.ed-ph
  • cs.AI

Abstract

Reasoning or inference-scaling models are the new generation of Large Language Models (LLMs) capable of complex problem solving. To investigate their problem-solving capability in physics, we evaluated model o4-mini by OpenAI on solving traditional, end-of-chapter problems from Halliday and Resnick's "Fundamentals of Physics," spanning core topics in the undergraduate physics curriculum. Performance was analyzed across modality and problem difficulty. The model solved the problems with overall accuracy of about 90%, but performance depended strongly on representation: accuracy was much higher on text-only problems (96%) than on problems requiring coordinated interpretation of text and images (79%). Accuracy also declined significantly as the problem difficulty increased from low to medium to high. These results show that state-of-the-art LLMs can solve much of the standard introductory physics problems, but that their performance remains uneven and constrained by problem modality and problem difficulty.

Read the original paper

This page indexes public paper metadata. The manuscript remains with its original publisher and authors.