The key shift seems to be from "can you produce mathematics?" to "do you understand mathematics?"... Those used to be correlated strongly enough that a thesis or paper could serve as evidence of both. If AI breaks that correlation, evaluating people through defenses, discussion and live problem solving suddenly makes a lot more sense.