Announcing ARC-AGI-3 - A benchmark that tests if AI can explore, learn, and adapt in unfamiliar situations. Humans score 100%. Frontier AI scores 0.26%.

brianpeiris@lemmy.ca · edit-2 2 days ago

Announcing ARC-AGI-3 - A benchmark that tests if AI can explore, learn, and adapt in unfamiliar situations. Humans score 100%. Frontier AI scores 0.26%.

NotMyOldRedditName@lemmy.world · edit-2 8 hours ago

Ya i agree. The whole infrastructure of how these work is flawed for a true AI/AGI.

It might be able to do a lot of cool things, but its fundamentally flawed at its core.

Someone will need to figure out something completely different for a true AI.

NotMyOldRedditName@lemmy.world · edit-2 8 hours ago

Oh also, I remember Elon once talked about how the upcoming cars would get bored when they weren’t doing anything with all that compute while parked so they could do use that compute and pay people for it.

Paying for the compute isnt a terrible idea in the future, but become bored? LOL. Fucking crazy talk.

Like even if it was a true AI that could be bored. You’re now going to enslave it to do what you want on its free time?

lordbritishbusiness@lemmy.world · 5 hours ago

Yeah, if it’s got the capacity to be bored it’s not going to stick around waiting for you. Pets act out when bored, as will AI, better to let the ghost in the machine go have fun in an arcade or something.

Current models can pretend to be bored when directed to, but they’re only facsimiles of thought at the moment, and the current approach probably won’t change that.

Announcing ARC-AGI-3 - A benchmark that tests if AI can explore, learn, and adapt in unfamiliar situations. Humans score 100%. Frontier AI scores 0.26%.

Announcing ARC-AGI-3 - A benchmark that tests if AI can explore, learn, and adapt in unfamiliar situations. Humans score 100%. Frontier AI scores 0.26%.

Announcing ARC-AGI-3 | ARC Prize