Announcing ARC-AGI-3 - A benchmark that tests if AI can explore, learn, and adapt in unfamiliar situations. Humans score 100%. Frontier AI scores 0.26%.

brianpeiris@lemmy.ca · edit-2 2 days ago

Announcing ARC-AGI-3 - A benchmark that tests if AI can explore, learn, and adapt in unfamiliar situations. Humans score 100%. Frontier AI scores 0.26%.

hitmyspot@aussie.zone · 2 hours ago

You’re discounting the fact that a human reading Wikipedia will attribute intonation and tone to the text to give further context and meaning. I think the analogy is good. Its not precise but it is the same thing.

I do think AI has a useful purpose and is here to stay. I don’t think it’s groundbreaking like the AI companies want us to think. The bubble will burst and then we’ll see where the cards lie.

OpenAI has lost their lead and I expect they will start to struggle with further funding. There are quite a few warning signs. The price of oil is likely to increase power prices generally and cause construction delays and cost rises. Both will hamper their plans. They still don’t have a viable model for profit.

mechoman444@lemmy.world · 54 minutes ago

The analogy is terrible and is not at all, once again, what llms do.

This is an objective fact I have provided evidence to support this.

How are you saying the analogy is good?

hitmyspot@aussie.zone · 19 minutes ago

Ana analogy does not need to be precise. It expresses a comparison for easier understanding. It is not what LLMs do. However what you’ve expressed is simplified also. So by your standard, it is not useful for the discussion.

So maybe get your head out of your ass and try to understand what people are trying to express instead of correcting them when they are not incorrect.

If precision was of that much importance to you, you would have a different opinion of LLMs.

Announcing ARC-AGI-3 - A benchmark that tests if AI can explore, learn, and adapt in unfamiliar situations. Humans score 100%. Frontier AI scores 0.26%.

Announcing ARC-AGI-3 - A benchmark that tests if AI can explore, learn, and adapt in unfamiliar situations. Humans score 100%. Frontier AI scores 0.26%.

Announcing ARC-AGI-3 | ARC Prize