▲ 145 ▼ ChatGPT 'got absolutely wrecked' by Atari 2600 in beginner's chess match — OpenAI's newest model bamboozled by 1970s logic (www.tomshardware.com) submitted 1 year ago by misk@sopuli.xyz to c/technology@beehaw.org 31 comments fedilink hide all child comments
[–] 30p87@feddit.org 107 points 1 year ago (3 children) Anyone even believing that a generic word auto completer would beat classic algorithms wherever possible probably belongs into a psychiatry. permalink fedilink source hideshow 6 child comments replies: [–] realitista@lemm.ee 56 points 1 year ago (3 children) There are a lot of people out there that think LLM's are somehow reasoning. Even reasoning models aren't really doing it. It important to do demonstrations like this in the hopes that the general public will understand the limitations of this tech. permalink fedilink source parent hideshow 6 child comments replies: [–] theangriestbird@beehaw.org 16 points 1 year ago (1 child) It is important to do demonstrations like this in the hopes that the general public will understand the limitations of this tech. THIS is the thing. The general public's perception of ChatGPT is basically whatever OpenAI's marketing department tells them to believe, plus their single memory of that one time they tested out ChatGPT and it was pretty impressive. Right now, OpenAI is telling everyone that they are a few years away from Artificial General Intelligence. Tests like this one demonstrate how wrong OpenAI is in that assertion. permalink fedilink source parent hideshow 2 child comments replies: [–] p03locke@lemmy.dbzer0.com 2 points 1 year ago It's almost as bad as the opposition's comparison of it to Skynet. People are never going to understand technology without applying some fucking nuance. Stop hyping new technology... in either direction. permalink fedilink source parent [–] ByteSorcerer@beehaw.org 9 points 1 year ago I think the problem is that, while the model isn't actually reasoning, it's very good at convincing people it actually is. I see current LLMs kinda like an RPG character build with all ability points put into Charisma. It's actually not that good at most tasks, but it's so good at convincing people that they start to think it's actually doing a great job. permalink fedilink source parent [+] Photuris@lemmy.ml 0 points 1 year ago* (last edited 10 months ago) (2 children) [deleted] permalink fedilink source parent hideshow 4 child comments replies: [–] realitista@lemm.ee 11 points 1 year ago (1 child) This is definitely part of the issue, not sure why people are downvoting this. That's also why tests like this are important, to illustrate that thinking in the way we know it isn't happening in these models. permalink fedilink source parent hideshow 2 child comments replies: [–] theangriestbird@beehaw.org 1 point 1 year ago (1 child) not sure why people are downvoting this downvotes are not allowed on beehaw fyi permalink fedilink source parent hideshow 2 child comments replies: [–] smeg@feddit.uk 1 point 1 year ago (1 child) Downvotes aren't federated but you still see all the downvotes sent from just your own instance permalink fedilink source parent hideshow 2 child comments replies: [–] theangriestbird@beehaw.org 1 point 1 year ago Interesting. I figured since this post is in a Beehaw community they would be invisible to everyone, but good to know. permalink fedilink source parent [–] jmcs@discuss.tchncs.de 4 points 1 year ago We understand reasoning enough to know humans (and other animals with complex brains) reason in a way that LLMs cannot. While our reasoning also works with pattern matching it incorporates immeasurably more signals than language - language is almost peripheric to it even in humans. And more importantly we experience things, everything we do acts as a small training round not just in language but on every aspect of the task we are performing, and gives us a miriad of patterns to match later. Until AI can match a fragment of this we are not going to have an AGI. And for the experience aspect there's no economic incentive under capitalism to achieve, if it happens it will come out of an underfunded university. permalink fedilink source parent [–] jjjalljs@ttrpg.network 23 points 1 year ago (1 child) I think I remember some doge goon asking online about using an LLM to parse JSON. Many people don't understand things. permalink fedilink source parent hideshow 2 child comments replies: [+] Photuris@lemmy.ml 12 points 1 year ago* (last edited 10 months ago) (1 child) [deleted] permalink fedilink source parent hideshow 2 child comments replies: [–] 30p87@feddit.org 5 points 1 year ago For us? Not as much, luckily most have the sentiment of rejecting anything LLM made and supported. But externals still have a lot of impact unfortunately, just ask @bagder@mastodon.social permalink fedilink source parent [–] MadMadBunny@lemmy.ca 20 points 1 year ago That’s too much critical thinking for most people permalink fedilink source parent
[–] realitista@lemm.ee 56 points 1 year ago (3 children) There are a lot of people out there that think LLM's are somehow reasoning. Even reasoning models aren't really doing it. It important to do demonstrations like this in the hopes that the general public will understand the limitations of this tech. permalink fedilink source parent hideshow 6 child comments replies: [–] theangriestbird@beehaw.org 16 points 1 year ago (1 child) It is important to do demonstrations like this in the hopes that the general public will understand the limitations of this tech. THIS is the thing. The general public's perception of ChatGPT is basically whatever OpenAI's marketing department tells them to believe, plus their single memory of that one time they tested out ChatGPT and it was pretty impressive. Right now, OpenAI is telling everyone that they are a few years away from Artificial General Intelligence. Tests like this one demonstrate how wrong OpenAI is in that assertion. permalink fedilink source parent hideshow 2 child comments replies: [–] p03locke@lemmy.dbzer0.com 2 points 1 year ago It's almost as bad as the opposition's comparison of it to Skynet. People are never going to understand technology without applying some fucking nuance. Stop hyping new technology... in either direction. permalink fedilink source parent [–] ByteSorcerer@beehaw.org 9 points 1 year ago I think the problem is that, while the model isn't actually reasoning, it's very good at convincing people it actually is. I see current LLMs kinda like an RPG character build with all ability points put into Charisma. It's actually not that good at most tasks, but it's so good at convincing people that they start to think it's actually doing a great job. permalink fedilink source parent [+] Photuris@lemmy.ml 0 points 1 year ago* (last edited 10 months ago) (2 children) [deleted] permalink fedilink source parent hideshow 4 child comments replies: [–] realitista@lemm.ee 11 points 1 year ago (1 child) This is definitely part of the issue, not sure why people are downvoting this. That's also why tests like this are important, to illustrate that thinking in the way we know it isn't happening in these models. permalink fedilink source parent hideshow 2 child comments replies: [–] theangriestbird@beehaw.org 1 point 1 year ago (1 child) not sure why people are downvoting this downvotes are not allowed on beehaw fyi permalink fedilink source parent hideshow 2 child comments replies: [–] smeg@feddit.uk 1 point 1 year ago (1 child) Downvotes aren't federated but you still see all the downvotes sent from just your own instance permalink fedilink source parent hideshow 2 child comments replies: [–] theangriestbird@beehaw.org 1 point 1 year ago Interesting. I figured since this post is in a Beehaw community they would be invisible to everyone, but good to know. permalink fedilink source parent [–] jmcs@discuss.tchncs.de 4 points 1 year ago We understand reasoning enough to know humans (and other animals with complex brains) reason in a way that LLMs cannot. While our reasoning also works with pattern matching it incorporates immeasurably more signals than language - language is almost peripheric to it even in humans. And more importantly we experience things, everything we do acts as a small training round not just in language but on every aspect of the task we are performing, and gives us a miriad of patterns to match later. Until AI can match a fragment of this we are not going to have an AGI. And for the experience aspect there's no economic incentive under capitalism to achieve, if it happens it will come out of an underfunded university. permalink fedilink source parent
[–] theangriestbird@beehaw.org 16 points 1 year ago (1 child) It is important to do demonstrations like this in the hopes that the general public will understand the limitations of this tech. THIS is the thing. The general public's perception of ChatGPT is basically whatever OpenAI's marketing department tells them to believe, plus their single memory of that one time they tested out ChatGPT and it was pretty impressive. Right now, OpenAI is telling everyone that they are a few years away from Artificial General Intelligence. Tests like this one demonstrate how wrong OpenAI is in that assertion. permalink fedilink source parent hideshow 2 child comments replies: [–] p03locke@lemmy.dbzer0.com 2 points 1 year ago It's almost as bad as the opposition's comparison of it to Skynet. People are never going to understand technology without applying some fucking nuance. Stop hyping new technology... in either direction. permalink fedilink source parent
[–] p03locke@lemmy.dbzer0.com 2 points 1 year ago It's almost as bad as the opposition's comparison of it to Skynet. People are never going to understand technology without applying some fucking nuance. Stop hyping new technology... in either direction. permalink fedilink source parent
[–] ByteSorcerer@beehaw.org 9 points 1 year ago I think the problem is that, while the model isn't actually reasoning, it's very good at convincing people it actually is. I see current LLMs kinda like an RPG character build with all ability points put into Charisma. It's actually not that good at most tasks, but it's so good at convincing people that they start to think it's actually doing a great job. permalink fedilink source parent
[+] Photuris@lemmy.ml 0 points 1 year ago* (last edited 10 months ago) (2 children) [deleted] permalink fedilink source parent hideshow 4 child comments replies: [–] realitista@lemm.ee 11 points 1 year ago (1 child) This is definitely part of the issue, not sure why people are downvoting this. That's also why tests like this are important, to illustrate that thinking in the way we know it isn't happening in these models. permalink fedilink source parent hideshow 2 child comments replies: [–] theangriestbird@beehaw.org 1 point 1 year ago (1 child) not sure why people are downvoting this downvotes are not allowed on beehaw fyi permalink fedilink source parent hideshow 2 child comments replies: [–] smeg@feddit.uk 1 point 1 year ago (1 child) Downvotes aren't federated but you still see all the downvotes sent from just your own instance permalink fedilink source parent hideshow 2 child comments replies: [–] theangriestbird@beehaw.org 1 point 1 year ago Interesting. I figured since this post is in a Beehaw community they would be invisible to everyone, but good to know. permalink fedilink source parent [–] jmcs@discuss.tchncs.de 4 points 1 year ago We understand reasoning enough to know humans (and other animals with complex brains) reason in a way that LLMs cannot. While our reasoning also works with pattern matching it incorporates immeasurably more signals than language - language is almost peripheric to it even in humans. And more importantly we experience things, everything we do acts as a small training round not just in language but on every aspect of the task we are performing, and gives us a miriad of patterns to match later. Until AI can match a fragment of this we are not going to have an AGI. And for the experience aspect there's no economic incentive under capitalism to achieve, if it happens it will come out of an underfunded university. permalink fedilink source parent
[–] realitista@lemm.ee 11 points 1 year ago (1 child) This is definitely part of the issue, not sure why people are downvoting this. That's also why tests like this are important, to illustrate that thinking in the way we know it isn't happening in these models. permalink fedilink source parent hideshow 2 child comments replies: [–] theangriestbird@beehaw.org 1 point 1 year ago (1 child) not sure why people are downvoting this downvotes are not allowed on beehaw fyi permalink fedilink source parent hideshow 2 child comments replies: [–] smeg@feddit.uk 1 point 1 year ago (1 child) Downvotes aren't federated but you still see all the downvotes sent from just your own instance permalink fedilink source parent hideshow 2 child comments replies: [–] theangriestbird@beehaw.org 1 point 1 year ago Interesting. I figured since this post is in a Beehaw community they would be invisible to everyone, but good to know. permalink fedilink source parent
[–] theangriestbird@beehaw.org 1 point 1 year ago (1 child) not sure why people are downvoting this downvotes are not allowed on beehaw fyi permalink fedilink source parent hideshow 2 child comments replies: [–] smeg@feddit.uk 1 point 1 year ago (1 child) Downvotes aren't federated but you still see all the downvotes sent from just your own instance permalink fedilink source parent hideshow 2 child comments replies: [–] theangriestbird@beehaw.org 1 point 1 year ago Interesting. I figured since this post is in a Beehaw community they would be invisible to everyone, but good to know. permalink fedilink source parent
[–] smeg@feddit.uk 1 point 1 year ago (1 child) Downvotes aren't federated but you still see all the downvotes sent from just your own instance permalink fedilink source parent hideshow 2 child comments replies: [–] theangriestbird@beehaw.org 1 point 1 year ago Interesting. I figured since this post is in a Beehaw community they would be invisible to everyone, but good to know. permalink fedilink source parent
[–] theangriestbird@beehaw.org 1 point 1 year ago Interesting. I figured since this post is in a Beehaw community they would be invisible to everyone, but good to know. permalink fedilink source parent
[–] jmcs@discuss.tchncs.de 4 points 1 year ago We understand reasoning enough to know humans (and other animals with complex brains) reason in a way that LLMs cannot. While our reasoning also works with pattern matching it incorporates immeasurably more signals than language - language is almost peripheric to it even in humans. And more importantly we experience things, everything we do acts as a small training round not just in language but on every aspect of the task we are performing, and gives us a miriad of patterns to match later. Until AI can match a fragment of this we are not going to have an AGI. And for the experience aspect there's no economic incentive under capitalism to achieve, if it happens it will come out of an underfunded university. permalink fedilink source parent
[–] jjjalljs@ttrpg.network 23 points 1 year ago (1 child) I think I remember some doge goon asking online about using an LLM to parse JSON. Many people don't understand things. permalink fedilink source parent hideshow 2 child comments replies: [+] Photuris@lemmy.ml 12 points 1 year ago* (last edited 10 months ago) (1 child) [deleted] permalink fedilink source parent hideshow 2 child comments replies: [–] 30p87@feddit.org 5 points 1 year ago For us? Not as much, luckily most have the sentiment of rejecting anything LLM made and supported. But externals still have a lot of impact unfortunately, just ask @bagder@mastodon.social permalink fedilink source parent
[+] Photuris@lemmy.ml 12 points 1 year ago* (last edited 10 months ago) (1 child) [deleted] permalink fedilink source parent hideshow 2 child comments replies: [–] 30p87@feddit.org 5 points 1 year ago For us? Not as much, luckily most have the sentiment of rejecting anything LLM made and supported. But externals still have a lot of impact unfortunately, just ask @bagder@mastodon.social permalink fedilink source parent
[–] 30p87@feddit.org 5 points 1 year ago For us? Not as much, luckily most have the sentiment of rejecting anything LLM made and supported. But externals still have a lot of impact unfortunately, just ask @bagder@mastodon.social permalink fedilink source parent
[–] MadMadBunny@lemmy.ca 20 points 1 year ago That’s too much critical thinking for most people permalink fedilink source parent