Everyone Is Judging AI by These Tests. But Experts Say They’re Close to Meaningless.

ModerateImprovement@sh.itjust.works · 4 months ago

Everyone Is Judging AI by These Tests. But Experts Say They’re Close to Meaningless.

A_A@lemmy.world · 4 months ago

Looks quite satisfying to me, otherwise, we can still create new tests … :

The tests cover an astounding range of knowledge, such as eighth-grade math, world history, and pop culture. Many are multiple choice, others take free-form answers. Some purport to measure knowledge of advanced fields like law, medicine and science. Others are more abstract, asking AI systems to choose the next logical step in a sequence of events, or to review “moral scenarios” and decide what actions would be considered acceptable behavior in society today.

Everyone Is Judging AI by These Tests. But Experts Say They’re Close to Meaningless.

Everyone Is Judging AI by These Tests. But Experts Say They’re Close to Meaningless.

Everyone Is Judging AI by These Tests. But Experts Say They’re Close to Meaningless – The Markup