linux-nerds.org

Your browser does not seem to support JavaScript. As a result, your viewing experience will be diminished, and you have been placed in read-only mode.

Please download a browser that supports JavaScript, or enable it if it's disabled (i.e. NoScript).

Apple just proved AI "reasoning" models like Claude, DeepSeek-R1, and o3-mini don't actually reason at all. They just memorize patterns really well.

Technology

350 Beiträge 149 Kommentatoren 23 Aufrufe

A allah@lemm.ee

LOOK MAA I AM ON FRONT PAGE
B This user is from outside of this forum
B This user is from outside of this forum
burgerpocalyse@lemmy.world

schrieb zuletzt editiert von

#338

hey I cant recognize patterns so theyre smarter than me at least
1 Antwort Letzte Antwort

2
K knock_knock_lemmy_in@lemmy.world

Other way around. The claimed meaningful change (reasoning) has not occurred.
C This user is from outside of this forum
C This user is from outside of this forum
communist@lemmy.frozeninferno.xyz

schrieb zuletzt editiert von

#339

Meaningful change is not happening because of this paper, either, I don't know why you're playing semantic games with me though.
K 1 Antwort Letzte Antwort

0
M mouldycat@feddit.uk

I think it's an easy mistake to confuse sentience and intelligence. It happens in Hollywood all the time - "Skynet began learning at a geometric rate, on July 23 2004 it became self-aware" yadda yadda

But that's not how sentience works. We don't have to be as intelligent as Skynet supposedly was in order to be sentient. We don't start our lives as unthinking robots, and then one day - once we've finally got a handle on calculus or a deep enough understanding of the causes of the fall of the Roman empire - we suddenly blink into consciousness. On the contrary, even the stupidest humans are accepted as being sentient. Even a young child, not yet able to walk or do anything more than vomit on their parents' new sofa, is considered as a conscious individual.

So there is no reason to think that AI - whenever it should be achieved, if ever - will be conscious any more than the dumb computers that precede it.
S This user is from outside of this forum
S This user is from outside of this forum
saturdaymorning@lemmy.ca

schrieb zuletzt editiert von

#340

Good point.
1 Antwort Letzte Antwort

0
C communist@lemmy.frozeninferno.xyz

Meaningful change is not happening because of this paper, either, I don't know why you're playing semantic games with me though.
K This user is from outside of this forum
K This user is from outside of this forum
knock_knock_lemmy_in@lemmy.world

schrieb zuletzt editiert von

#341

I don't know why you're playing semantic games

I'm trying to highlight the goal of this paper.

This is a knock them down paper by Apple justifying (to their shareholders) their non investment in LLMs. It is not a build them up paper trying for meaningful change and to create a better AI.
C 1 Antwort Letzte Antwort

0
K knock_knock_lemmy_in@lemmy.world

I don't know why you're playing semantic games

I'm trying to highlight the goal of this paper.

This is a knock them down paper by Apple justifying (to their shareholders) their non investment in LLMs. It is not a build them up paper trying for meaningful change and to create a better AI.
C This user is from outside of this forum
C This user is from outside of this forum
communist@lemmy.frozeninferno.xyz

schrieb zuletzt editiert von

#342

That's not the only way to make meaningful change, getting people to give up on llms would also be meaningful change. This does very little for anyone who isn't apple.
1 Antwort Letzte Antwort

0
S skisnow@lemmy.ca

I hate this analogy. As a throwaway whimsical quip it'd be fine, but it's specious enough that I keep seeing it used earnestly by people who think that LLMs are in any way sentient or conscious, so it's lowered my tolerance for it as a topic even if you did intend it flippantly.
G This user is from outside of this forum
G This user is from outside of this forum
gamechld@lemmy.world

schrieb zuletzt editiert von

#343

I don't mean it to extol LLM's but rather to denigrate humans. How many of us are self imprisoned in echo chambers so we can have our feelings validated to avoid the uncomfortable feeling of thinking critically and perhaps changing viewpoints?

Humans have the ability to actually think, unlike LLM's. But it's frightening how far we'll go to make sure we don't.
1 Antwort Letzte Antwort

0
V vrighter@discuss.tchncs.de

the fact that it is a fixed function, that only depends on the context AND there are a finite number of discrete inputs possible does make it equivalent to a huge, finite table. You really don't want this to be true. And again, you are describing training. Once training finishes anything you said does not apply anymore and you are left with fixed, unchanging matrices, which in turn means that it is a mathematical function of the context (by the mathematical definition of "function". stateless, and deterministic) which also has the property that the set of all possible inputs is finite. So the set of possible outputs is also finite and strictly smaller or equal to the size of the set of possible inputs. This makes the actual function that the tokens are passed through CAN be precomputed in full (in theory) making it equivalent to a conventional state transition table.

This is true whether you'd like it to or not. The training process builds a markov chain.
A This user is from outside of this forum
A This user is from outside of this forum
auraithx@lemmy.dbzer0.com

schrieb zuletzt editiert von

#344

You’re absolutely right that inference in an LLM is a fixed, deterministic function after training, and that the input space is finite due to the discrete token vocabulary and finite context length. So yes, in theory, you could precompute every possible input-output mapping and store them in a giant table. That much is mathematically valid. But where your argument breaks down is in claiming that this makes an LLM equivalent to a conventional Markov chain in function or behavior.

A Markov chain is not simply defined as “a function from finite context to next-token distribution.” It is defined by a specific type of process where the next state depends on the current state via fixed transition probabilities between discrete states. The model operates over symbolic states with no internal computation. LLMs, even during inference, compute outputs via multi-layered continuous transformations, with attention mixing, learned positional embeddings, and non-linear activations. These mechanisms mean that while the function is fixed, its structure does not resemble a state machine—it resembles a hierarchical pattern recognizer and function approximator.

Your claim is essentially that “any deterministic function over a finite input space is equivalent to a table.” This is true in a computational sense but misleading in a representational and behavioral sense. If I gave you a function that maps 4096-bit inputs to 50257-dimensional probability vectors and said, “This is equivalent to a transition table,” you could technically agree, but the structure and generative capacity of that function is not Markovian. That function may simulate reasoning, abstraction, and composition. A Markov chain never does.

You are collapsing implementation equivalence (yes, the function could be stored in a table) with model equivalence (no, it does not behave like a Markov chain). The fact that you could freeze the output behavior into a lookup structure doesn’t change that the lookup structure is derived from a fundamentally different class of computation.

The training process doesn’t “build a Markov chain.” It builds a function that estimates conditional token probabilities via optimization over a non-Markov architecture. The inference process then applies that function. That makes it a stateless function, yes—but not a Markov chain. Determinism plus finiteness does not imply Markovian behavior.
V 1 Antwort Letzte Antwort

0
M minoscopede@lemmy.world

I'd encourage you to research more about this space and learn more.

As it is, the statement "Markov chains are still the basis of inference" doesn't make sense, because markov chains are a separate thing. You might be thinking of Markov decision processes, which is used in training RL agents, but that's also unrelated because these models are not RL agents, they're supervised learning agents. And even if they were RL agents, the MDP describes the training environment, not the model itself, so it's not really used for inference.

I mean this just as an invitation to learn more, and not pushback for raising concerns. Many in the research community would be more than happy to welcome you into it. The world needs more people who are skeptical of AI doing research in this field.
T This user is from outside of this forum
T This user is from outside of this forum
tobberone@lemm.ee

schrieb zuletzt editiert von

#345

Which method, then, is the inference built upon, if not the embeddings? And the question still stands, how does "AI" escape the inherent limits of statistical inference?
1 Antwort Letzte Antwort

0
A auraithx@lemmy.dbzer0.com

You’re absolutely right that inference in an LLM is a fixed, deterministic function after training, and that the input space is finite due to the discrete token vocabulary and finite context length. So yes, in theory, you could precompute every possible input-output mapping and store them in a giant table. That much is mathematically valid. But where your argument breaks down is in claiming that this makes an LLM equivalent to a conventional Markov chain in function or behavior.

A Markov chain is not simply defined as “a function from finite context to next-token distribution.” It is defined by a specific type of process where the next state depends on the current state via fixed transition probabilities between discrete states. The model operates over symbolic states with no internal computation. LLMs, even during inference, compute outputs via multi-layered continuous transformations, with attention mixing, learned positional embeddings, and non-linear activations. These mechanisms mean that while the function is fixed, its structure does not resemble a state machine—it resembles a hierarchical pattern recognizer and function approximator.

Your claim is essentially that “any deterministic function over a finite input space is equivalent to a table.” This is true in a computational sense but misleading in a representational and behavioral sense. If I gave you a function that maps 4096-bit inputs to 50257-dimensional probability vectors and said, “This is equivalent to a transition table,” you could technically agree, but the structure and generative capacity of that function is not Markovian. That function may simulate reasoning, abstraction, and composition. A Markov chain never does.

You are collapsing implementation equivalence (yes, the function could be stored in a table) with model equivalence (no, it does not behave like a Markov chain). The fact that you could freeze the output behavior into a lookup structure doesn’t change that the lookup structure is derived from a fundamentally different class of computation.

The training process doesn’t “build a Markov chain.” It builds a function that estimates conditional token probabilities via optimization over a non-Markov architecture. The inference process then applies that function. That makes it a stateless function, yes—but not a Markov chain. Determinism plus finiteness does not imply Markovian behavior.
V This user is from outside of this forum
V This user is from outside of this forum
vrighter@discuss.tchncs.de

schrieb zuletzt editiert von

#346

you wouldn't be "freezing" anything. Each possible combination of input tokens maps to one output probability distribution. Those values are fixed and they are what they are whether you compute them or not, or when, or how many times.

Now you can either precompute the whole table (theory), or somehow compute each cell value every time you need it (practice). In either case, the resulting function (table lookup vs matrix multiplications) takes in only the context, and produces a probability distribution. And the mapping they generate is the same for all possible inputs. So they are the same function. A function can be implemented in multiple ways, but the implementation is not the function itself. The only difference between the two in this case is the implementation, or more specifically, whether you precompute a table or not. But the function itself is the same.

You are somehow saying that your choice of implementation for that function will somehow change the function. Which means that according to you, if you do precompute (or possibly cache, full precomputation is just an infinite cache size) individual mappings it somehow magically makes some magic happen that gains some deep insight. It does not. We have already established that it is the same function.
1 Antwort Letzte Antwort

0
A allah@lemm.ee

LOOK MAA I AM ON FRONT PAGE
F This user is from outside of this forum
F This user is from outside of this forum
fourwaveforms@lemm.ee

schrieb zuletzt editiert von fourwaveforms@lemm.ee

#347

WTF does the author think reasoning is
1 Antwort Letzte Antwort

1
A antonim@lemmy.dbzer0.com

That depends on your assumption that the left would have anything relevant to gain by embracing AI (whatever that's actually supposed to mean).
M This user is from outside of this forum
M This user is from outside of this forum
melvin_ferd@lemmy.world

schrieb zuletzt editiert von

#348

Saw this earlier in the week and thought of you. These short, funny videos are popping up more and more and they're only getting better. They’re sharp, engaging, and they spread like wildfire.

You strike me as someone who gets it what it means when one side embraces the latest tools while the other rejects them.

The left is still holed up on Lemmy, clinging to “Fuck AI” groups. But why? Go back to the beginning. Look at the early coverage of AI it was overwhelmingly targeted at left-leaning spaces, full of panic and doom. Compare that to how the right talks about immigration. The headlines are cut and pasted from each other. Same playbook, different topic. The media set out to alienate the left from these tools.

71 reactions · 11 shares | I’m just tryna KICK it bro 🤷🏾‍♂️👟 Share for the Marines in Los Angeles ICE riots ! Follow @themarinerapper for more ai military and music videos. | The Marine Rapper

I’m just tryna KICK it bro 🤷🏾‍♂️👟 Share for the Marines in Los Angeles ICE riots ! Follow @themarinerapper for more ai military and music videos..

(www.facebook.com)
A 1 Antwort Letzte Antwort

0
M melvin_ferd@lemmy.world

Saw this earlier in the week and thought of you. These short, funny videos are popping up more and more and they're only getting better. They’re sharp, engaging, and they spread like wildfire.

You strike me as someone who gets it what it means when one side embraces the latest tools while the other rejects them.

The left is still holed up on Lemmy, clinging to “Fuck AI” groups. But why? Go back to the beginning. Look at the early coverage of AI it was overwhelmingly targeted at left-leaning spaces, full of panic and doom. Compare that to how the right talks about immigration. The headlines are cut and pasted from each other. Same playbook, different topic. The media set out to alienate the left from these tools.

71 reactions · 11 shares | I’m just tryna KICK it bro 🤷🏾‍♂️👟 Share for the Marines in Los Angeles ICE riots ! Follow @themarinerapper for more ai military and music videos. | The Marine Rapper

I’m just tryna KICK it bro 🤷🏾‍♂️👟 Share for the Marines in Los Angeles ICE riots ! Follow @themarinerapper for more ai military and music videos..

(www.facebook.com)
A This user is from outside of this forum
A This user is from outside of this forum
antonim@lemmy.dbzer0.com

schrieb zuletzt editiert von

#349

I don't have even the slightest idea what that video is supposed to mean. (Happy cake day tho.)
M 1 Antwort Letzte Antwort

0
A antonim@lemmy.dbzer0.com

I don't have even the slightest idea what that video is supposed to mean. (Happy cake day tho.)
M This user is from outside of this forum
M This user is from outside of this forum
melvin_ferd@lemmy.world

schrieb zuletzt editiert von

#350

Come on, you know what I’m talking about. It’s a channel that started with AI content and is now pivoting to videos about the riots. You can see where this is going. Sooner or later, it’ll expand into targeting protestors and other left-leaning causes.

It’s a novelty now, but it’s spreading fast, and more channels like it are popping up every day.

Meanwhile, the left is losing ground. Losing cultural capture. Because as a group, they’re being manipulated into isolating themselves from the very tools and platforms that shape public opinion. Social media. AI. All of it. They're walking away from the battlefield while the other side builds momentum.
1 Antwort Letzte Antwort

0

Anmelden zum Antworten

L

AOSP isn't dead, but Google just landed a huge blow to custom ROM developers
Beobachtet Ignoriert Geplant Angeheftet Gesperrt Verschoben Technology technology
22

123 Stimmen

22 Beiträge

0 Aufrufe

O

The absolutely criminal dark patterns that they pull on people via Google photos auto backup is insane. Just in my own orbit 2 of my friends wives, my parents, and my in-laws all wound up paying Google because they thought they had to or lose all their photos. We helped most of them disconnect the autobackup (that they didn't even know was activated) and move it to offline safely. But that was the most downright evil shit Google has ever done and literally a fire in me for manipulating the elderly and less tech savvy so blatantly.
P

Welcome to the web we lost
Beobachtet Ignoriert Geplant Angeheftet Gesperrt Verschoben Technology technology
22

1

181 Stimmen

22 Beiträge

4 Aufrufe

C

Is it though? Its always far easier to be loud and obnoxious than do something constructive, even with the internet and LLMs, in fact those things are amplifiers which if anything make the attention imbalance even more drastic and unrepresentative of actual human behaviour. In the time it takes me to write this comment some troll can write a dozen hateful ones, or a bot can write a thousand. Doesn't mean humans are shitty in a 1000/1 ratio, just means shitty people can now be a thousand times louder.
G

Have LLMs Finally Mastered Geolocation? - bellingcat
Beobachtet Ignoriert Geplant Angeheftet Gesperrt Verschoben Technology technology
3

1

50 Stimmen

3 Beiträge

2 Aufrufe

R

Depends on who programed the AI - and no, it is not Kyoto
P

Google quietly released an app that lets you download and run AI models locally
Beobachtet Ignoriert Geplant Angeheftet Gesperrt Verschoben Technology technology
49

1

100 Stimmen

49 Beiträge

5 Aufrufe

A

Okay man.
A

US government is using AI for unprecedented social media surveillance
Beobachtet Ignoriert Geplant Angeheftet Gesperrt Verschoben Technology technology
18

1

185 Stimmen

18 Beiträge

2 Aufrufe

N

Part of the reason for my use of "might".
P

Mozilla is shutting down Pocket, their read-it-later and content discovery app, and Fakespot, their browser extension that analyzes the authenticity of online product reviews.
Beobachtet Ignoriert Geplant Angeheftet Gesperrt Verschoben Technology technology
12

1

3 Stimmen

12 Beiträge

2 Aufrufe

G

Yeah, I don’t know how they’re doing it. They’re using some “zero trust” system. It’s beyond me.
?

[paper] Evidence of a social evaluation penalty for using AI
Beobachtet Ignoriert Geplant Angeheftet Gesperrt Verschoben Technology technology
10

28 Stimmen

10 Beiträge

9 Aufrufe

V

I'm specifically talking about toil when it comes to my job as a software developer. I already know I need an if statement and a for loop all wrapped in a try catch. Rather then spending a couple minutes coding that I have cursor do it for me instantly then fill out the actual code. Or, ive written something in python and it needs to be converted to JavaScript. I can ask Claude to convert it one to one for me and test it, which comes back with either no errors or a very simple error I need to fix. It takes a minute. Instead I could have taken 15min to rewrite it myself and maybe make more mistakes that take longer.
P

DOGE bro Kyle Schutt's computer infected by malware, credentials found in stealer logs
Beobachtet Ignoriert Geplant Angeheftet Gesperrt Verschoben Technology technology
4

1

142 Stimmen

4 Beiträge

2 Aufrufe

P

The topic is more nuanced, all the logs indicate email/password combos that were compromised. While it is possible this is due to a malware infection, it could be something as simple as a phishing website. In this case, credentials are entered but no "malware" was installed. The point being it doesn't look great that someone has ANY compromises... But again, anyone who's used the Internet a bit has some compromised. For example, in a password manager (especially the one on iPhone), you'll often be notified of all your potentially compromised accounts. [image: 7a5e8350-e47e-4d67-b096-e6e470ec7050.jpeg]