# The next page in the book of AI evolution is here, powered by GPT 3.5, and I am very, nay, extremely impressed

**URL:** <https://boards.straightdope.com/t/the-next-page-in-the-book-of-ai-evolution-is-here-powered-by-gpt-3-5-and-i-am-very-nay-extremely-impressed/975945>\
**Category:** Cafe Society\
**Tags:** ai\
**Created:** [December 2, 2022, 10:54pm UTC](https://boards.straightdope.com/t/the-next-page-in-the-book-of-ai-evolution-is-here-powered-by-gpt-3-5-and-i-am-very-nay-extremely-impressed/975945 "2022-12-02T22:54:31Z")\
**Posts on this page:** 20\
**Page:** 69

<div class="post-metadata">

**Author:** ![Crane](https://sea3.discourse-cdn.com/straightdope/user_avatar/boards.straightdope.com/crane/32/3495_2.png) [@Crane](https://boards.straightdope.com/u/Crane)\
**Post date:** [March 29, 2023, 8:01pm UTC](https://boards.straightdope.com/t/the-next-page-in-the-book-of-ai-evolution-is-here-powered-by-gpt-3-5-and-i-am-very-nay-extremely-impressed/975945/1361 "2023-03-29T20:01:57Z")

</div>

> [@Sam\_Stone](#):
>
> In contrast, program synthesis with few-shot learning using Codex fine-tuned on code generates programs that automatically solve 81% of these questions. Our approach improves the previous state-of-the-art automatic solution accuracy on the benchmark topics from 8.8 to 81.1%.

It is a product they developed. It was not unexpected. It was hard work by skilled folks.

---

<div class="post-metadata">

**Author:** ![Sam\_Stone](https://avatars.discourse-cdn.com/v4/letter/s/ecccb3/32.png) [@Sam\_Stone](https://boards.straightdope.com/u/Sam_Stone)\
**Post date:** [March 29, 2023, 8:20pm UTC](https://boards.straightdope.com/t/the-next-page-in-the-book-of-ai-evolution-is-here-powered-by-gpt-3-5-and-i-am-very-nay-extremely-impressed/975945/1362 "2023-03-29T20:20:21Z")

</div>

It was a product they fine-tuned for coding. They did not expect that fine-tuning for coding would result in greatly improved math skills.

Btw, I believe ChatGPT is fine-tuned for coding, so it might have just been that. Or GPT-4.

You just won’t accept the huge amount of evidence that there is more going on here than being a ‘stochastic parrot’. People claim these things don’t understand the meaning of the tokens, just their relationships. Then we find ‘neurons’ in the model that associate objects like spiders or anything else with other objects. That’s not ‘relationships between tokens’, that’s understanding that the words actually mean things in the real world. People say they are just doing word prediction, then we find generalized algorithms inside them that allow them so solve math problems.

People said image models were just doing ‘diffusion’, and the models didn’t understand what any of the pixels actually represent. Except that you can show a picture of a man off balance and ask Palm-E what’s happening, and it will respond, “That man is about to fall over.” Or you can show it a picture of a DB-15 RS-232 plugged into an iphone, and it will say, “That’s a joke, because the RS-232 is a big connector and the iPhone is a small phone. And the RS-232 dosn’t work on an iPhone.”

Clearly there’s a lot more going on in that model than just pixel manipulation, even if the output stage of the thing is ‘just diffusion’. And clearly there’s a lot more going on in these language models, even if the output stage is just ‘next word prediction’.

---

<div class="post-metadata">

**Author:** ![Babale](https://sea3.discourse-cdn.com/straightdope/user_avatar/boards.straightdope.com/babale/32/15666_2.png) [@Babale](https://boards.straightdope.com/u/Babale)\
**Post date:** [March 29, 2023, 8:20pm UTC](https://boards.straightdope.com/t/the-next-page-in-the-book-of-ai-evolution-is-here-powered-by-gpt-3-5-and-i-am-very-nay-extremely-impressed/975945/1363 "2023-03-29T20:20:30Z")

</div>

> [@Crane](#):
>
> It is a product they developed

Right.

> [@Crane](#):
>
> It was not unexpected

Wrong.

> [@Crane](#):
>
> It was hard work by skilled folks.

Right.

---

<div class="post-metadata">

**Author:** ![Chronos](https://sea3.discourse-cdn.com/straightdope/user_avatar/boards.straightdope.com/chronos/32/134_2.png) [@Chronos](https://boards.straightdope.com/u/Chronos)\
**Post date:** [March 29, 2023, 10:17pm UTC](https://boards.straightdope.com/t/the-next-page-in-the-book-of-ai-evolution-is-here-powered-by-gpt-3-5-and-i-am-very-nay-extremely-impressed/975945/1364 "2023-03-29T22:17:48Z")

</div>

> [@Half\_Man\_Half\_Wit](#):
>
> Of course. But even a full-fledged implementation of AIXI would be fallible—it uses a formalized version of induction (Solomonoff inference), and inductive conclusions are always defeasible. AIXI might produce the most parsimonious continuation of a given sequence, but that can still be just wrong.

Well, yes, but humans are also fallible well beyond those limits. In particular, it’s very difficult to predict the behavior of another prediction-engine that’s actively trying to fool you.

One might even envision a game of sorts, where two parties are simultaneously attempting to both predict the other’s behavior, and to render their own behavior unpredictable. Or one might not even need to envision the game, because it already exists and is quite familiar to everyone.

Everyone knows, of course, that even a half-decent random-number generator can’t be beaten at rock-paper-scissors by any human (though it also can’t beat any human). But beyond that, there are also computers that can consistently beat humans at rock-paper-scissors. Computers are already better than humans at both predicting and foiling predictions (which are really the same task), at least in that limited context, and have been for ages.

---

<div class="post-metadata">

**Author:** ![Darren\_Garrison](https://sea3.discourse-cdn.com/straightdope/user_avatar/boards.straightdope.com/darren_garrison/32/92_2.png) [@Darren\_Garrison](https://boards.straightdope.com/u/Darren_Garrison)\
**Post date:** [March 29, 2023, 10:26pm UTC](https://boards.straightdope.com/t/the-next-page-in-the-book-of-ai-evolution-is-here-powered-by-gpt-3-5-and-i-am-very-nay-extremely-impressed/975945/1365 "2023-03-29T22:26:22Z")

</div>

> **[Elon Musk and others urge AI pause, citing 'risks to society'](https://www.reuters.com/technology/musk-experts-urge-pause-training-ai-systems-that-can-outperform-gpt-4-2023-03-29/)**
>
> Elon Musk and a group of artificial intelligence experts and industry executives are calling for a six-month pause in developing systems more powerful than OpenAI's newly launched GPT-4, in an open letter citing potential risks to society.

---

<div class="post-metadata">

**Author:** ![Sam\_Stone](https://avatars.discourse-cdn.com/v4/letter/s/ecccb3/32.png) [@Sam\_Stone](https://boards.straightdope.com/u/Sam_Stone)\
**Post date:** [March 30, 2023, 12:48am UTC](https://boards.straightdope.com/t/the-next-page-in-the-book-of-ai-evolution-is-here-powered-by-gpt-3-5-and-i-am-very-nay-extremely-impressed/975945/1366 "2023-03-30T00:48:38Z")

</div>

Yeah, that ship has sailed. It’s not going to happen. Things are moving so fast that everyone is worried about being left behind. No one is going to stop enhancing AIs, because they know that everyone else is going to keep building new ones.

I would worry more about driving development underground where we have no visibiity into what’s going on at all. Because no one is stopping now, IMO.

---

<div class="post-metadata">

**Author:** ![Half\_Man\_Half\_Wit](https://sea3.discourse-cdn.com/straightdope/user_avatar/boards.straightdope.com/half_man_half_wit/32/21766_2.png) [@Half\_Man\_Half\_Wit](https://boards.straightdope.com/u/Half_Man_Half_Wit)\
**Post date:** [March 30, 2023, 3:36am UTC](https://boards.straightdope.com/t/the-next-page-in-the-book-of-ai-evolution-is-here-powered-by-gpt-3-5-and-i-am-very-nay-extremely-impressed/975945/1367 "2023-03-30T03:36:30Z")

</div>

> [@Sam\_Stone](#):
>
> People claim these things don’t understand the meaning of the tokens, just their relationships.

Because the information about the meaning of the tokens can’t be derived from the training data, barring magic. It’s just not there; you can’t spin gold from straw, no matter the sophistication of your loom.

> [@Sam\_Stone](#):
>
> Then we find ‘neurons’ in the model that associate objects like spiders or anything else with other objects. That’s not ‘relationships between tokens’, that’s understanding that the words actually mean things in the real world.

Well, things asserted without evidence can be dismissed without argument, but still: I don’t see why this should be surprising at all, or why you think this isn’t just a relation between tokens. A token is encoded into a vector of neuron activations in such a way that tokens that are close in the sense of being able to replace one another in many contexts—they ‘keep the same company’ in the training data—will also be close to another in the resulting vector space. Thus, if one token leads to a high activation in a particular neuron, _due to this relation between tokens_, so one would expect related ones to do.

---

<div class="post-metadata">

**Author:** ![Snarky\_Kong](https://avatars.discourse-cdn.com/v4/letter/s/a183cd/32.png) [@Snarky\_Kong](https://boards.straightdope.com/u/Snarky_Kong)\
**Post date:** [March 30, 2023, 3:52am UTC](https://boards.straightdope.com/t/the-next-page-in-the-book-of-ai-evolution-is-here-powered-by-gpt-3-5-and-i-am-very-nay-extremely-impressed/975945/1368 "2023-03-30T03:52:46Z")

</div>

If you associate visual or textual tokens with tokens generated via real world sensor data, is that understanding? And if not, what would constitute understanding?

---

<div class="post-metadata">

**Author:** ![Half\_Man\_Half\_Wit](https://sea3.discourse-cdn.com/straightdope/user_avatar/boards.straightdope.com/half_man_half_wit/32/21766_2.png) [@Half\_Man\_Half\_Wit](https://boards.straightdope.com/u/Half_Man_Half_Wit)\
**Post date:** [March 30, 2023, 4:20am UTC](https://boards.straightdope.com/t/the-next-page-in-the-book-of-ai-evolution-is-here-powered-by-gpt-3-5-and-i-am-very-nay-extremely-impressed/975945/1369 "2023-03-30T04:20:38Z")

</div>

> [@Snarky\_Kong](#):
>
> If you associate visual or textual tokens with tokens generated via real world sensor data, is that understanding?

Mere association can’t really be understanding. Consider a device that has a series of cue cards of different images it can display, a mechanics to call each one up, and a set of keys such that if one of them is inserted into its keyhole and turned, one of the cards pops up: does the device understand that a given key ‘means’ a certain picture? I don’t think so: all we have is a mechanism triggered by the ‘shape’ of the key—its syntactic properties, if you will. So any system that just calls up pictures based on syntactic properties—as it seems current AI systems are doing—also don’t have any understanding of what they’re doing.

> [@Snarky\_Kong](#):
>
> And if not, what would constitute understanding?

Essentially, having access to the semantic properties of a token: to know what a symbol refers to, to have the tokens _act_ as proper symbols.

---

<div class="post-metadata">

**Author:** ![Snarky\_Kong](https://avatars.discourse-cdn.com/v4/letter/s/a183cd/32.png) [@Snarky\_Kong](https://boards.straightdope.com/u/Snarky_Kong)\
**Post date:** [March 30, 2023, 4:28am UTC](https://boards.straightdope.com/t/the-next-page-in-the-book-of-ai-evolution-is-here-powered-by-gpt-3-5-and-i-am-very-nay-extremely-impressed/975945/1370 "2023-03-30T04:28:17Z")

</div>

> [@Half\_Man\_Half\_Wit](#):
>
> does the device understand that a given key ‘means’ a certain picture?

If you ask it for a certain picture and it reliably produces the right key I would think so? Unless I don’t know what you mean.

> [@Half\_Man\_Half\_Wit](#):
>
> Essentially, having access to the semantic properties of a token: to know what a symbol refers to, to have the tokens _act_ as proper symbols.

Latent space embedding have semantic meaning. A famous example of King - Man + Woman ~= Queen.

---

<div class="post-metadata">

**Author:** ![Babale](https://sea3.discourse-cdn.com/straightdope/user_avatar/boards.straightdope.com/babale/32/15666_2.png) [@Babale](https://boards.straightdope.com/u/Babale)\
**Post date:** [March 30, 2023, 4:32am UTC](https://boards.straightdope.com/t/the-next-page-in-the-book-of-ai-evolution-is-here-powered-by-gpt-3-5-and-i-am-very-nay-extremely-impressed/975945/1371 "2023-03-30T04:32:45Z")

</div>

> [@Half\_Man\_Half\_Wit](#):
>
> > [@Snarky\_Kong](#):
> >
> > If you associate visual or textual tokens with tokens generated via real world sensor data, is that understanding?
> 
> Mere association can’t really be understanding. Consider a device that has a series of cue cards of different images it can display, a mechanics to call each one up, and a set of keys such that if one of them is inserted into its keyhole and turned, one of the cards pops up: does the device understand that a given key ‘means’ a certain picture? I don’t think so: all we have is a mechanism triggered by the ‘shape’ of the key—its syntactic properties, if you will. So any system that just calls up pictures based on syntactic properties—as it seems current AI systems are doing—also don’t have any understanding of what they’re doing.

What do we have access to that an AI with access to vision, audio, and language doesn’t have?

---

<div class="post-metadata">

**Author:** ![Half\_Man\_Half\_Wit](https://sea3.discourse-cdn.com/straightdope/user_avatar/boards.straightdope.com/half_man_half_wit/32/21766_2.png) [@Half\_Man\_Half\_Wit](https://boards.straightdope.com/u/Half_Man_Half_Wit)\
**Post date:** [March 30, 2023, 4:33am UTC](https://boards.straightdope.com/t/the-next-page-in-the-book-of-ai-evolution-is-here-powered-by-gpt-3-5-and-i-am-very-nay-extremely-impressed/975945/1372 "2023-03-30T04:33:53Z")

</div>

Yes, that’s a very unfortunate example of a technical term in one field being the same as a different one in another. To disambiguate between the two, a distinction is sometimes made between ‘inferential’ semantics, which is the sort of thing you get from the word-vector embeddings (‘the company a word keeps’), and ‘referential’ semantics, which is the sort of thing usually thought of as the ‘meaning’ of a symbol—what it refers to. It’s the latter that’s the interesting part.

---

<div class="post-metadata">

**Author:** ![Half\_Man\_Half\_Wit](https://sea3.discourse-cdn.com/straightdope/user_avatar/boards.straightdope.com/half_man_half_wit/32/21766_2.png) [@Half\_Man\_Half\_Wit](https://boards.straightdope.com/u/Half_Man_Half_Wit)\
**Post date:** [March 30, 2023, 4:36am UTC](https://boards.straightdope.com/t/the-next-page-in-the-book-of-ai-evolution-is-here-powered-by-gpt-3-5-and-i-am-very-nay-extremely-impressed/975945/1373 "2023-03-30T04:36:25Z")

</div>

> [@Babale](#):
>
> What do we have access to that an AI with access to vision, audio, and language doesn’t have?

I don’t know for sure (noone does), but I believe it is direct access to non-structural features of cognitive processing, mediated by a self-referential modeling process—this is just the topic of [my theory of conscious experience](https://link.springer.com/article/10.1007/s10670-021-00467-w).

---

<div class="post-metadata">

**Author:** ![Snarky\_Kong](https://avatars.discourse-cdn.com/v4/letter/s/a183cd/32.png) [@Snarky\_Kong](https://boards.straightdope.com/u/Snarky_Kong)\
**Post date:** [March 30, 2023, 4:36am UTC](https://boards.straightdope.com/t/the-next-page-in-the-book-of-ai-evolution-is-here-powered-by-gpt-3-5-and-i-am-very-nay-extremely-impressed/975945/1374 "2023-03-30T04:36:27Z")

</div>

Prove those are different.

Word-vector embedding aren’t defined by the company they keep. They’re learned that way.

---

<div class="post-metadata">

**Author:** ![Half\_Man\_Half\_Wit](https://sea3.discourse-cdn.com/straightdope/user_avatar/boards.straightdope.com/half_man_half_wit/32/21766_2.png) [@Half\_Man\_Half\_Wit](https://boards.straightdope.com/u/Half_Man_Half_Wit)\
**Post date:** [March 30, 2023, 4:37am UTC](https://boards.straightdope.com/t/the-next-page-in-the-book-of-ai-evolution-is-here-powered-by-gpt-3-5-and-i-am-very-nay-extremely-impressed/975945/1375 "2023-03-30T04:37:50Z")

</div>

> [@Snarky\_Kong](#):
>
> Prove those are different.

[I already have.](https://boards.straightdope.com/t/why-chatgpt-doesnt-understand/980404)

But this isn’t really a contentious assertion. Take the [discussion here](https://www.sciencedirect.com/science/article/abs/pii/S0911604416301063):

> In philosophy of language, a distinction has been proposed by Diego Marconi between two aspects of lexical semantic competence, i.e. inferential and referential competence (Marconi, 1997). One aspect of lexical competence, i.e. inferential competence, is the «ability to deal with the network of semantic relations among lexical units, underlying such performances as semantic inference, paraphrase, definition, retrieval of a word from its definition, finding a synonym, and so forth» (Marconi, 1997, p. 59). For instance, we know that a _cat_ is an _animal_, we can verbally describe the differences between a _cat_ and a _dog_, we can recover the word _cat_ from a definition such as _The animal that meows_, and so on. Such “intralinguistic” abilities are semantic because, in order to exercise them, a speaker must possess an internalized network specifying semantic connections between a given word (e.g., _cat_) and other words of a natural language (e.g., _animal_, _meow_).
> 
> The second aspect of lexical competence, i.e. «referential competence», cognitively mediates the relation between words and entities of the world. For example, we have the ability to classify a given perceived object as a _cat_ or to distinguish it from a _dog_, to recognize and name a picture of a cat, and so on. Clearly, we can speak of referential competence only relative to words that refer to objects, properties or events we can perceive (e.g., _cat_, _red_, _hot_).

---

<div class="post-metadata">

**Author:** ![Snarky\_Kong](https://avatars.discourse-cdn.com/v4/letter/s/a183cd/32.png) [@Snarky\_Kong](https://boards.straightdope.com/u/Snarky_Kong)\
**Post date:** [March 30, 2023, 4:45am UTC](https://boards.straightdope.com/t/the-next-page-in-the-book-of-ai-evolution-is-here-powered-by-gpt-3-5-and-i-am-very-nay-extremely-impressed/975945/1376 "2023-03-30T04:45:48Z")

</div>

Pretty poor example since neural networks pass both of those aspects using the same embeddings.

---

<div class="post-metadata">

**Author:** ![Half\_Man\_Half\_Wit](https://sea3.discourse-cdn.com/straightdope/user_avatar/boards.straightdope.com/half_man_half_wit/32/21766_2.png) [@Half\_Man\_Half\_Wit](https://boards.straightdope.com/u/Half_Man_Half_Wit)\
**Post date:** [March 30, 2023, 4:47am UTC](https://boards.straightdope.com/t/the-next-page-in-the-book-of-ai-evolution-is-here-powered-by-gpt-3-5-and-i-am-very-nay-extremely-impressed/975945/1377 "2023-03-30T04:47:55Z")

</div>

Really? You know what is being cognitively mediated in a neural network, and how? That’s a trick you’ll have to explain to me!

---

<div class="post-metadata">

**Author:** ![Snarky\_Kong](https://avatars.discourse-cdn.com/v4/letter/s/a183cd/32.png) [@Snarky\_Kong](https://boards.straightdope.com/u/Snarky_Kong)\
**Post date:** [March 30, 2023, 4:50am UTC](https://boards.straightdope.com/t/the-next-page-in-the-book-of-ai-evolution-is-here-powered-by-gpt-3-5-and-i-am-very-nay-extremely-impressed/975945/1378 "2023-03-30T04:50:07Z")

</div>

> [@Half\_Man\_Half\_Wit](#):
>
> For example, we have the ability to classify a given perceived object as a _cat_ or to distinguish it from a _dog_, to recognize and name a picture of a cat, and so on.

Hmm.

> [@Half\_Man\_Half\_Wit](#):
>
> Really? You know what is being cognitively mediated in a neural network, and how? That’s a trick you’ll have to explain to me!

Really? You know what is being cognitively mediated in my head, and how? That’s a trick you’ll have to explain to me!

Take your faux- surprise to a different forum.

---

<div class="post-metadata">

**Author:** ![Half\_Man\_Half\_Wit](https://sea3.discourse-cdn.com/straightdope/user_avatar/boards.straightdope.com/half_man_half_wit/32/21766_2.png) [@Half\_Man\_Half\_Wit](https://boards.straightdope.com/u/Half_Man_Half_Wit)\
**Post date:** [March 30, 2023, 4:52am UTC](https://boards.straightdope.com/t/the-next-page-in-the-book-of-ai-evolution-is-here-powered-by-gpt-3-5-and-i-am-very-nay-extremely-impressed/975945/1379 "2023-03-30T04:52:37Z")

</div>

Please don’t misattribute quotes to me.

---

<div class="post-metadata">

**Author:** ![Ponderoid](https://sea3.discourse-cdn.com/straightdope/user_avatar/boards.straightdope.com/ponderoid/32/19221_2.png) [@Ponderoid](https://boards.straightdope.com/u/Ponderoid)\
**Post date:** [March 30, 2023, 7:35am UTC](https://boards.straightdope.com/t/the-next-page-in-the-book-of-ai-evolution-is-here-powered-by-gpt-3-5-and-i-am-very-nay-extremely-impressed/975945/1380 "2023-03-30T07:35:15Z")

</div>

> [@Sam\_Stone](#):
>
> if you have a big enough context window you can dump a research paper into it

Just a nitpick, but you’re confusing the input box limit with the context window. In ChatGPT, the context window is a LOT bigger than the amount you can stuff into a single prompt. I haven’t tested this lately to see if anything’s changed, but early on I noticed that it was possible to stuff a lot more text into the input window than would actually be processed when submitted. That amount processed was still small compared to the total context window.

And in other news, I came in here to tell everyone that OpenAI just put a Bing-like conversation length limit into ChatGPT. I think I just watched it happen right now; I got 2 generic “error in body stream” errors and then on the 3rd try for the same prompt at the same depth in the conversation it returned “conversation too long”.It was not a very long conversation compared to others I’ve had.

ETA: this might be some other stupid error that’s being misidentified. I just tried backing up and extending a shorter branch of the same tree and got the same error.

[Previous page](https://boards.straightdope.com/t/the-next-page-in-the-book-of-ai-evolution-is-here-powered-by-gpt-3-5-and-i-am-very-nay-extremely-impressed/975945.md?page=68)

[Next page](https://boards.straightdope.com/t/the-next-page-in-the-book-of-ai-evolution-is-here-powered-by-gpt-3-5-and-i-am-very-nay-extremely-impressed/975945.md?page=70)
