linux-nerds.org

Your browser does not seem to support JavaScript. As a result, your viewing experience will be diminished, and you have been placed in read-only mode.

Please download a browser that supports JavaScript, or enable it if it's disabled (i.e. NoScript).

Grok 4 has been so badly neutered that it's now programmed to see what Elon says about the topic at hand and blindly parrot that line.

Technology

67 Beiträge 55 Kommentatoren 0 Aufrufe

B beliefpropagator@discuss.tchncs.de

I found this: https://simonwillison.net/2025/Jul/11/grok-musk/
P This user is from outside of this forum
P This user is from outside of this forum
pixxelkick@lemmy.world

schrieb zuletzt editiert von

#29

That's more like it, thank you!
1 Antwort Letzte Antwort

3
U unexposedhazard@discuss.tchncs.de

I think there is a good chance this behavior is unintended!

Lmao, sure...
M This user is from outside of this forum
M This user is from outside of this forum
mirodir@discuss.tchncs.de

schrieb zuletzt editiert von

#30

I can believe it insofar as they might not have explicitly programmed it to do that. I'd imagine they put in something like "Make sure your output aligns with Elon Musk's opinions.", "Elon Musk is always objectively correct.", etc. From there, this would be emergent, but quite predictable behavior.
U 1 Antwort Letzte Antwort

15
C cherry@piefed.social

This is my take. Elon just showed the world what we all knew. The tool is not trustworthy. All other AI suppliers are busy trying to work on credibility that grok just butchered.
D This user is from outside of this forum
D This user is from outside of this forum
deceptichum@quokk.au

schrieb zuletzt editiert von deceptichum@quokk.au

#31

They deliberately injected prompts on top of the users prompt.

Saying that’s a problem of AI is akin to say me deliberately painting my car badly and saying it’s a problem of all car manufacturers.

And this frankly shows how little you know about the subject, because we went through this years ago with prompts trying to force corpo-lib “diversity” and leading to hilarious results.

If anything you should be concerned about the non prompt stuff, the underlying training data that it pulls from and of which I doubt Grok has even changed since release.
C 1 Antwort Letzte Antwort

1
D destructdisc@lemmy.world

This post did not contain any content.
G This user is from outside of this forum
G This user is from outside of this forum
gameline@sopuli.xyz

schrieb zuletzt editiert von

#32

they should just put it down and out of it's misery
W 1 Antwort Letzte Antwort

19
D destructdisc@lemmy.world

This post did not contain any content.
Z This user is from outside of this forum
Z This user is from outside of this forum
zomg@lemmy.world

schrieb zuletzt editiert von zomg@lemmy.world

#33

Honestly, who was surprised by this news?

I feel like everyone could see Grok as some sort of 24/7 tool to push a particular viewpoint, even more so when it says things that are leftist and Elon is compelled to "upgrade" the system as he's tweeted.
1 Antwort Letzte Antwort

55
D destructdisc@lemmy.world

This post did not contain any content.
B This user is from outside of this forum
B This user is from outside of this forum
blackmist@feddit.uk

schrieb zuletzt editiert von

#34

I'm surprised it isn't just Elon typing really fast at this point.
R G 2 Antworten Letzte Antwort

71
M mirodir@discuss.tchncs.de

I can believe it insofar as they might not have explicitly programmed it to do that. I'd imagine they put in something like "Make sure your output aligns with Elon Musk's opinions.", "Elon Musk is always objectively correct.", etc. From there, this would be emergent, but quite predictable behavior.
U This user is from outside of this forum
U This user is from outside of this forum
unexposedhazard@discuss.tchncs.de

schrieb zuletzt editiert von

#35

Yeah the transparency of it might be unintended.
1 Antwort Letzte Antwort

5
U unexposedhazard@discuss.tchncs.de

I think there is a good chance this behavior is unintended!

Lmao, sure...
T This user is from outside of this forum
T This user is from outside of this forum
theunknownmuncher@lemmy.world

schrieb zuletzt editiert von

#36

If the system prompt doesn’t tell it to search for Elon’s views, why is it doing that?

My best guess is that Grok “knows” that it is “Grok 4 buit by xAI”, and it knows that Elon Musk owns xAI, so in circumstances where it’s asked for an opinion the reasoning process often decides to see what Elon thinks.

Yeah, this blogger shows a fundamental misunderstanding of how LLMs work or how system prompts work. LLM behavior is not directly controlled by the system prompt the way this person imagines. For example, censorship that is present in the training set will be "baked in" to the model and the system prompt will not affect it, no matter how the LLM is told not to be censored in that way.

My best guess is that the LLM is interfacing with a tool in order to search through tweets, and the training set that demonstrates how to use the tool contains example searches for Elon Musk's tweets.
L 1 Antwort Letzte Antwort

25
G gameline@sopuli.xyz

they should just put it down and out of it's misery
W This user is from outside of this forum
W This user is from outside of this forum
worldsdumbestman@lemmy.today

schrieb zuletzt editiert von

#37

It used to be so based
1 Antwort Letzte Antwort

11
B blackmist@feddit.uk

I'm surprised it isn't just Elon typing really fast at this point.
R This user is from outside of this forum
R This user is from outside of this forum
reseller_pledge609@lemmy.dbzer0.com

schrieb zuletzt editiert von

#38

Probably couldn't type fast if he tried. Would probably pay someone to do it for him just like he did with Path if Exile.
T 1 Antwort Letzte Antwort

41
R reseller_pledge609@lemmy.dbzer0.com

Probably couldn't type fast if he tried. Would probably pay someone to do it for him just like he did with Path if Exile.
T This user is from outside of this forum
T This user is from outside of this forum
test_tickles@lemmy.world

schrieb zuletzt editiert von

#39

And like he does with inseminating women.
V 1 Antwort Letzte Antwort

14
T theunknownmuncher@lemmy.world

If the system prompt doesn’t tell it to search for Elon’s views, why is it doing that?

My best guess is that Grok “knows” that it is “Grok 4 buit by xAI”, and it knows that Elon Musk owns xAI, so in circumstances where it’s asked for an opinion the reasoning process often decides to see what Elon thinks.

Yeah, this blogger shows a fundamental misunderstanding of how LLMs work or how system prompts work. LLM behavior is not directly controlled by the system prompt the way this person imagines. For example, censorship that is present in the training set will be "baked in" to the model and the system prompt will not affect it, no matter how the LLM is told not to be censored in that way.

My best guess is that the LLM is interfacing with a tool in order to search through tweets, and the training set that demonstrates how to use the tool contains example searches for Elon Musk's tweets.
L This user is from outside of this forum
L This user is from outside of this forum
lepinkainen@lemmy.world

schrieb zuletzt editiert von

#40

“This blogger” is Simon Willison, who has been doing LLM benchmarks and other LLM-related things since before it was cool

Not a random substack grifter
T 1 Antwort Letzte Antwort

10
D deceptichum@quokk.au

They deliberately injected prompts on top of the users prompt.

Saying that’s a problem of AI is akin to say me deliberately painting my car badly and saying it’s a problem of all car manufacturers.

And this frankly shows how little you know about the subject, because we went through this years ago with prompts trying to force corpo-lib “diversity” and leading to hilarious results.

If anything you should be concerned about the non prompt stuff, the underlying training data that it pulls from and of which I doubt Grok has even changed since release.
C This user is from outside of this forum
C This user is from outside of this forum
cherry@piefed.social

schrieb zuletzt editiert von

#41

You are correct. But the right tool in the wrong hands is still non credible in the eyes of perception.
1 Antwort Letzte Antwort

5
L loduz_247@lemmy.world

Grok's journey has been very strange. He became a progressive, then threw out data that contradicted the MAGA people who questioned him, and finally became a Hitler fan.

Now he's the reflection of a fan who blindly follows Trump, but in this case, he's an AI. His journey so far has been curious.
D This user is from outside of this forum
D This user is from outside of this forum
damage@feddit.it

schrieb zuletzt editiert von

#42

So Grok is a 4chan incel?

His only chance of salvation is finding a girl who inexplicably fancies it?
1 Antwort Letzte Antwort

0
L lepinkainen@lemmy.world

“This blogger” is Simon Willison, who has been doing LLM benchmarks and other LLM-related things since before it was cool

Not a random substack grifter
T This user is from outside of this forum
T This user is from outside of this forum
theunknownmuncher@lemmy.world

schrieb zuletzt editiert von theunknownmuncher@lemmy.world

#43

Is my comment wrong though? Another possibility is that Grok is given an example of searching for Elon Musk's tweets when it is presented with the available tool calls. Just because it outputs the system prompt when asked does not mean that we are seeing the full context, or even the real system prompt.

Posting blog guides on how to code with ChatGPT is not expertise on LLMs. It's like thinking someone is an expert mechanic because they can drive a car well.
J 1 Antwort Letzte Antwort

9
D destructdisc@lemmy.world

This post did not contain any content.
A This user is from outside of this forum
A This user is from outside of this forum
almacca@aussie.zone

schrieb zuletzt editiert von

#44

Robert A. Heinlein is turning in his grave like a fucking dynamo these days.
1 Antwort Letzte Antwort

30
T theunknownmuncher@lemmy.world

Is my comment wrong though? Another possibility is that Grok is given an example of searching for Elon Musk's tweets when it is presented with the available tool calls. Just because it outputs the system prompt when asked does not mean that we are seeing the full context, or even the real system prompt.

Posting blog guides on how to code with ChatGPT is not expertise on LLMs. It's like thinking someone is an expert mechanic because they can drive a car well.
J This user is from outside of this forum
J This user is from outside of this forum
jwmgregory@lemmy.dbzer0.com

schrieb zuletzt editiert von jwmgregory@lemmy.dbzer0.com

#45

Willison has never claimed to be an expert in the field of machine learning, but you should give more credence to his opinions. Perhaps u/lepinkainen@lemmy.world's warning wasn't informative enough to be heeded: Willison is a prominent figure in the web-development scene, particularly aspects of the scene that have evolved into important facets of the modern machine learning community.

The guy is quite experienced with Python and took an early step into the contemporary ML/AI space due to both him having a lot of very relevant skills and a likely personal interest in the field. Python is the lingua franca of my field of study, for better or worse, and someone like Willison was well-placed to break into ML/AI from the outside. That's a common route in this field, there aren't exactly an abundance of MBAs with majors in machine learning or applied artificial intelligence research, specifically (yet). Willison is one of the authors of Django, for fucks sake. Idk what he's doing rn but it would be ignorant to draw the comparison you just did in the context of Willison particularly. [EDIT: Lmfao just went to see "what is Simon doing rn" (don't really keep up with him in particular), & you're talking out of your ass. He literally has multiple tools for the machine learning stack that he develops and that are available to see on his github. See one such here. This guy is so far away from someone who just "posts random blog guides on how to code with ChatGPT" that it's egregious you'd even claim that. It's so disingenuous as to ere into dishonesty; like, that is a patent lie. Smh.]

As for your analysis of his article, I find it kind of ironic you accuse him of having a "fundamental misunderstanding of how LLMs work or how system prompts work [sic]" when you then proceed to cherry-pick certain lines from his article taken entirely out of context. First, the article is clearly geared towards a more general audience and avoids technical language or explanation. Second, he doesn't say anything that is fundamentally wrong. Honestly, you seem to have a far more ignorant idea of LLMs and this field generally than Willison. You do say some things that are wrong, such as:

For example, censorship that is present in the training set will be “baked in” to the model and the system prompt will not affect it, no matter how the LLM is told not to be censored in that way.

This isn't necessarily true. It is true that information not included within the training set, or information that has been statistically biased within the training set, isn't going to be retrievable or reversible using system prompts. Willison never claims or implies this in his article, you just kind of stuff those words in his mouth. Either way, my point is that you are using wishy-washy, ambiguous, catch-all terms such as "censorship" that make your writings here not technically correct, either. What is censorship, in an informatics context? What does that mean? How can it be applied to sets of data? That's not a concretely defined term if you're wanting to take the discourse to the level that it seems you are, like it or not. Generally you seem to have something of a misunderstanding regarding this topic, but I'm not going to accuse you of that, lest I commit the same fallacy I'm sitting here trying to chastise you for. It's possible you do know what you're talking about and just dumbed it down for Lemmy. It's impossible for me to know as an audience.

That all wouldn't really matter if you didn't just jump as Willison's credibility over your perception of him doing that exact same thing, though.
T 1 Antwort Letzte Antwort

8
D destructdisc@lemmy.world

This post did not contain any content.
A This user is from outside of this forum
A This user is from outside of this forum
arin@lemmy.world

schrieb zuletzt editiert von

#46

Mecha-Hitler is just Mecha-Elon
1 Antwort Letzte Antwort

35
T test_tickles@lemmy.world

And like he does with inseminating women.
V This user is from outside of this forum
V This user is from outside of this forum
vxx@lemmy.world

schrieb zuletzt editiert von

#47

Ketamine took its toll
M 1 Antwort Letzte Antwort

6
J jwmgregory@lemmy.dbzer0.com

Willison has never claimed to be an expert in the field of machine learning, but you should give more credence to his opinions. Perhaps u/lepinkainen@lemmy.world's warning wasn't informative enough to be heeded: Willison is a prominent figure in the web-development scene, particularly aspects of the scene that have evolved into important facets of the modern machine learning community.

The guy is quite experienced with Python and took an early step into the contemporary ML/AI space due to both him having a lot of very relevant skills and a likely personal interest in the field. Python is the lingua franca of my field of study, for better or worse, and someone like Willison was well-placed to break into ML/AI from the outside. That's a common route in this field, there aren't exactly an abundance of MBAs with majors in machine learning or applied artificial intelligence research, specifically (yet). Willison is one of the authors of Django, for fucks sake. Idk what he's doing rn but it would be ignorant to draw the comparison you just did in the context of Willison particularly. [EDIT: Lmfao just went to see "what is Simon doing rn" (don't really keep up with him in particular), & you're talking out of your ass. He literally has multiple tools for the machine learning stack that he develops and that are available to see on his github. See one such here. This guy is so far away from someone who just "posts random blog guides on how to code with ChatGPT" that it's egregious you'd even claim that. It's so disingenuous as to ere into dishonesty; like, that is a patent lie. Smh.]

As for your analysis of his article, I find it kind of ironic you accuse him of having a "fundamental misunderstanding of how LLMs work or how system prompts work [sic]" when you then proceed to cherry-pick certain lines from his article taken entirely out of context. First, the article is clearly geared towards a more general audience and avoids technical language or explanation. Second, he doesn't say anything that is fundamentally wrong. Honestly, you seem to have a far more ignorant idea of LLMs and this field generally than Willison. You do say some things that are wrong, such as:

For example, censorship that is present in the training set will be “baked in” to the model and the system prompt will not affect it, no matter how the LLM is told not to be censored in that way.

This isn't necessarily true. It is true that information not included within the training set, or information that has been statistically biased within the training set, isn't going to be retrievable or reversible using system prompts. Willison never claims or implies this in his article, you just kind of stuff those words in his mouth. Either way, my point is that you are using wishy-washy, ambiguous, catch-all terms such as "censorship" that make your writings here not technically correct, either. What is censorship, in an informatics context? What does that mean? How can it be applied to sets of data? That's not a concretely defined term if you're wanting to take the discourse to the level that it seems you are, like it or not. Generally you seem to have something of a misunderstanding regarding this topic, but I'm not going to accuse you of that, lest I commit the same fallacy I'm sitting here trying to chastise you for. It's possible you do know what you're talking about and just dumbed it down for Lemmy. It's impossible for me to know as an audience.

That all wouldn't really matter if you didn't just jump as Willison's credibility over your perception of him doing that exact same thing, though.
T This user is from outside of this forum
T This user is from outside of this forum
theunknownmuncher@lemmy.world

schrieb zuletzt editiert von theunknownmuncher@lemmy.world

#48

Willison has never claimed to be an expert in the field of machine learning, but you should give more credence to his opinions.

Yeah, I would if he didn't demonstrate such blatant misconceptions.

Willison is a prominent figure in the web-development scene

"They know how to sail a boat so they know how a car engine works"

Willison never claims or implies this in his article, you just kind of stuff those words in his mouth.

Reading comprehension. I never implied that he says anything about censorship. It is a correct and valid example that shows how his understanding is wrong about how system prompts work. "Define censorship" is not the argument you think it is lol. Okay though, I'll define the "censorship" I'm talking about as refusal behavior that is introduced during RLHF and DPO alignment, and no the system prompt will not change this behavior.

EDIT: saw your edit about him publishing tools that make using an LLM easier. Yeahhhh lol writing python libraries to interface with LLM APIs is not LLM expertise, that's still just using LLMs but programatically. See analogy about being a mechanic vs a good driver.
J 1 Antwort Letzte Antwort

3

Anmelden zum Antworten

G

Missouri AG: Any AI That Doesn’t Praise Donald Trump Might Be “Consumer Fraud” (No, Really)
Beobachtet Ignoriert Geplant Angeheftet Gesperrt Verschoben Technology technology
27

1

305 Stimmen

27 Beiträge

0 Aufrufe

P

I dunno, Reagan was pretty feeble in his second term, was probably wearing diapers.
S

Spotify X Mod APK
Beobachtet Ignoriert Geplant Angeheftet Gesperrt Verschoben Technology technology
1

2

1 Stimmen

1 Beiträge

3 Aufrufe

Niemand hat geantwortet
E

The Education Crisis: How AI Is Failing Students for the Future Job Market
Beobachtet Ignoriert Geplant Angeheftet Gesperrt Verschoben Technology technology
15

1

89 Stimmen

15 Beiträge

62 Aufrufe

S

I suspect people (not billionaires) are realising that they can get by with less. And that the planet needs that too. And that working 40+ hours a week isn’t giving people what they really want either. Tbh, I don't think that's the case. If you look at any of the relevant metrics (CO², energy consumption, plastic waste, ...) they only know one direction globally and that's up. I think the actual issues are Russian invasion of Ukraine and associated sanctions on one of the main energy providers of Europe Trump's "trade wars" which make global supply lines unreliable and costs incalculable (global supply chains love nothing more than uncertainty) Uncertainty in regards to China/Taiwan Boomers retiring in western countries, which for the first time since pretty much ever means that the work force is shrinking instead of growing. Economical growth was mostly driven by population growth for the last half century with per-capita productivity staying very close to inflation. Disrupting changes in key industries like cars and energy. The west has been sleeping on may of these developments (e.g. electric cars, batteries, solar) and now China is curbstomping the rest of the world in regards to market share. High key interest rates (which are applied to reduce high inflation due to some of the reason above) reduce demand on financial investments into companies. The low interest rates of the 2010s and also before lead to more investments into companies. With interest going back up, investments dry up. All these changes mean that companies, countries and people in the west have much less free cash available. There’s also the value of money has never been lower either. That's been the case since every. Inflation has always been a thing and with that the value of money is monotonically decreasing. But that doesn't really matter for the whole argument, since the absolute value of money doesn't matter, only the relative value. To put it differently: If you earn €100 and the thing you want to buy costs €10, that is equivalent to if you earn €1000 and the thing you want to buy costing €100. The value of money dropping is only relevant for savings, and if people are saving too much then the economy slows down and jobs are cut, thus some inflation is positive or even required. What is an actual issue is that wages are not increasing at the same rate as the cost of things, but that's not a "value of the money" issue.
P

In Militarizing Push, Russian School Children To Build Drones
Beobachtet Ignoriert Geplant Angeheftet Gesperrt Verschoben Technology technology
37

1

263 Stimmen

37 Beiträge

141 Aufrufe

Z

https://yourlogicalfallacyis.com/tu-quoque
S

Apple’s Craig Federighi on the long road to the iPad’s Mac-like multitasking
Beobachtet Ignoriert Geplant Angeheftet Gesperrt Verschoben Technology technology
35

1

46 Stimmen

35 Beiträge

129 Aufrufe

M

You guys sure display a crazy obsession with "Apple Fanboys" in this sub… The amount of Applephobia… Phew! As if the new release had you all flustered or something… Gotta take a bite and taste the Apple at some point! Can’t stay in the closet forever, ya know?
P

Microsoft Gives European Union Users More Control: Uninstall Edge, Store, and Say Goodbye to Bing Prompts
Beobachtet Ignoriert Geplant Angeheftet Gesperrt Verschoben Technology technology
187

1

761 Stimmen

187 Beiträge

353 Aufrufe

O

Not being a coward.
A

Valve CEO Gabe Newell’s Neuralink competitor is expecting its first brain chip this year
Beobachtet Ignoriert Geplant Angeheftet Gesperrt Verschoben Technology technology
15

1

48 Stimmen

15 Beiträge

54 Aufrufe

E

Their Bionic Eyes Are Now Obsolete and Unsupported
D

Judge Rules Apple Top Executive Alex Roman Lied Under Oath, Makes Criminal Contempt Referral
Beobachtet Ignoriert Geplant Angeheftet Gesperrt Verschoben Technology technology
6

0 Stimmen

6 Beiträge

32 Aufrufe

P

I applaud this, but I still say it's not far enough. Adjusted, the amount might match, but 121.000 is still easier to cough up for a billionaire than 50 is for a single mother of two who can barely make ends meet