Elektrine lite

← Feed

@gankouskhan@bookwyr.me

2026-09-04 02:52 UTC

For as long as it is an LLM I will not call it intelligence for it does not think. It’s an approximation or an estimate at best. While capable this is a major hurdle these companies will face to replace skilled workers. We are a generation at least and a ways off from general intelligence, and not algorithmic output for a given input (idempotency). I do however agree that we would remain informed about the ever changing world around us; this however, is until reviewed by an unbiased third party in masse a marketing document. I will be taking it with a grain of salt, but will take it as an improvement over the prior versions. Now with that out of the way… interesting. I will be watching this progress, but am far less optimistic in this being as capable as they are suggesting. It’s more than likely another Mythos hype attempt that is not revolutionary, but a nice addition to existing systems or augments to staff. Now if we assume its 100% as capable as this suggests then we are in for a ride.

Replies (1)

  • @brianpeiris@lemmy.ca 2026-09-04 04:39

    I agree with most of this. I’ve also said in the past that LLMs cannot think, and I think that’s still true for most models. The reason ARC-AGI-3 is interesting is that it was specifically designed to test reasoning, adaptability, novel problem solving, planning, memory, etc. So it was a surprise to me that Astra was able to defeat it so effectively, and that Astra invents algebras for each novel task. But I agree we can’t trust OpenAI if these results are self-reported, and we may not be able to trust the ARC Prize Foundation fully either. Extraordinary claims require extraordinary evidence, so we need replication, transparency, and proper open science to confirm things. I also agree with ARC Prize’s conclusion, that there are still capabilities any AI system would need to demonstrate before we can claim a full general intelligence.

    Open ##4668802