A Princeton-led team posted a paper today describing the first AI evaluation designed around a deceptively simple idea: give an AI agent a genuinely open research question — one whose answer is ...