<?xml version="1.0" encoding="UTF-8"?><rss version="2.0"
	xmlns:content="http://purl.org/rss/1.0/modules/content/"
	xmlns:dc="http://purl.org/dc/elements/1.1/"
	xmlns:atom="http://www.w3.org/2005/Atom"
	xmlns:sy="http://purl.org/rss/1.0/modules/syndication/"
	>
<channel>
	<title>Comments for The Tar Pit</title>
	<atom:link href="http://thetarpit.org/comments/feed" rel="self" type="application/rss+xml" />
	<link>http://thetarpit.org</link>
	<description>"Now I feel like I know less about what that blog is about than I did before."</description>
	<pubDate>Wed, 05 Aug 2026 01:58:46 +0000</pubDate>
	<generator>http://thetarpit.org</generator>
	<sy:updatePeriod>hourly</sy:updatePeriod>
	<sy:updateFrequency>1</sy:updateFrequency>
		<item>
		<title>Comment on LLMs, the useful parts by spyked</title>
		<link>http://thetarpit.org/2026/llms-the-useful-parts#comment-7652</link>
		<dc:creator>spyked</dc:creator>
		<pubDate>Thu, 30 Jul 2026 18:07:00 +0000</pubDate>
		<guid isPermaLink="false">http://thetarpit.org/?p=593#comment-7652</guid>
		<description>&lt;blockquote&gt;a large (tera/petabyte-sized) corpus of so-called "training data"; think: literature, academic databases, videos, emails, Facebook comments, in other words, all the data that a large "service provider" such as Google or Meta could have gathered in a few decades;&lt;/blockquote&gt;

One of the problems with this approach -- which is quite obvious, although it might not be obvious to you, so let me spell it out clearly over here -- is that, in attempting to be an efficient mind-reader, the LLM agent will often assume the wrong thing. This, by the way, is in my humble opinion one of the most life-like characteristics of LLMs.

Take any life form -- say, a grapevine. Its characteristics are a function of both genetics, in the potential they provide via reproduction, and of the environment. Thus a quality or another of the grapes it produces will stem both from its heritage *and* the soil it inhabits, the altitude where it grew, the wind, sun and so on and so forth. And it works the same with intelligent life forms: say, if you take a bunch of &lt;a href="http://thetarpit.org/2018/lacul-morii?b=University&#038;e=Bucharest#select" rel="nofollow"&gt;UPB&lt;/a&gt; graduates, there's a high statistical likelihood that they will bear similar linguistic characteristics, which, for example, will help a small group of graduates who joined the local IT company communicate quite efficiently.

On the flip side, that's the problem with language, i.e. that it's largely contextual: if you were to bring in someone from Cluj, Iași, or even from Bucharest's own Universitate, they would have some trouble fitting in, even though supposedly engineering language works the same everywhere in the world (yeah, right). So, getting back to LLMs, they only exacerbate this problem, because of their huge training corpus.

In practice, this means that the user will have to go through great lengths to specify the problem context, but try as they might, that context cannot possibly be *complete*, due to the sheer fact that natural language is based on implicit context by its very... well, nature. So then the LLM agent will fill in the contextual gaps with whatever's more statistically likely, as derived from its training corpus. It often happens that that guessing will indeed resemble some form of mind-reading; but the flip side to that is that when it doesn't, it can lead to particularly nasty failure modes. Which makes the LLM the perfect programmable slot machine.</description>
		<content:encoded><![CDATA[<blockquote><p>a large (tera/petabyte-sized) corpus of so-called "training data"; think: literature, academic databases, videos, emails, Facebook comments, in other words, all the data that a large "service provider" such as Google or Meta could have gathered in a few decades;</p></blockquote>
<p>One of the problems with this approach -- which is quite obvious, although it might not be obvious to you, so let me spell it out clearly over here -- is that, in attempting to be an efficient mind-reader, the LLM agent will often assume the wrong thing. This, by the way, is in my humble opinion one of the most life-like characteristics of LLMs.</p>
<p>Take any life form -- say, a grapevine. Its characteristics are a function of both genetics, in the potential they provide via reproduction, and of the environment. Thus a quality or another of the grapes it produces will stem both from its heritage *and* the soil it inhabits, the altitude where it grew, the wind, sun and so on and so forth. And it works the same with intelligent life forms: say, if you take a bunch of <a href="http://thetarpit.org/2018/lacul-morii?b=University&#038;e=Bucharest#select" rel="nofollow">UPB</a> graduates, there's a high statistical likelihood that they will bear similar linguistic characteristics, which, for example, will help a small group of graduates who joined the local IT company communicate quite efficiently.</p>
<p>On the flip side, that's the problem with language, i.e. that it's largely contextual: if you were to bring in someone from Cluj, Iași, or even from Bucharest's own Universitate, they would have some trouble fitting in, even though supposedly engineering language works the same everywhere in the world (yeah, right). So, getting back to LLMs, they only exacerbate this problem, because of their huge training corpus.</p>
<p>In practice, this means that the user will have to go through great lengths to specify the problem context, but try as they might, that context cannot possibly be *complete*, due to the sheer fact that natural language is based on implicit context by its very... well, nature. So then the LLM agent will fill in the contextual gaps with whatever's more statistically likely, as derived from its training corpus. It often happens that that guessing will indeed resemble some form of mind-reading; but the flip side to that is that when it doesn't, it can lead to particularly nasty failure modes. Which makes the LLM the perfect programmable slot machine.</p>
]]></content:encoded>
	</item>
	<item>
		<title>Comment on Trilemabot and Feedbot V patches, the winter 2020 session by LLMs, the useful parts &#171; The Tar Pit</title>
		<link>http://thetarpit.org/2020/botworks-ix#comment-7621</link>
		<dc:creator>LLMs, the useful parts &#171; The Tar Pit</dc:creator>
		<pubDate>Sun, 26 Jul 2026 18:54:32 +0000</pubDate>
		<guid isPermaLink="false">http://thetarpit.org/?p=352#comment-7621</guid>
		<description>[...] work on them independently in separate sessions (y'know, just like in real life); or if it's not sufficiently clearly specified, the result will be a mess that you'll have a very hard time recovering from (also just [...]</description>
		<content:encoded><![CDATA[<p>[...] work on them independently in separate sessions (y'know, just like in real life); or if it's not sufficiently clearly specified, the result will be a mess that you'll have a very hard time recovering from (also just [...]</p>
]]></content:encoded>
	</item>
	<item>
		<title>Comment on Re: Ivanovna et al., 2020 (DOI: 10.7554/eLife.58906) by LLMs, the useful parts &#171; The Tar Pit</title>
		<link>http://thetarpit.org/2021/re-ivanovna-et-al-2020#comment-7620</link>
		<dc:creator>LLMs, the useful parts &#171; The Tar Pit</dc:creator>
		<pubDate>Sun, 26 Jul 2026 18:54:17 +0000</pubDate>
		<guid isPermaLink="false">http://thetarpit.org/?p=401#comment-7620</guid>
		<description>[...] of ways, among which: a lossy compression scheme; or a statistical pattern matching scheme; or a geometrical processing function; in any case, an algorithm which reduces the training data to a so-called [...]</description>
		<content:encoded><![CDATA[<p>[...] of ways, among which: a lossy compression scheme; or a statistical pattern matching scheme; or a geometrical processing function; in any case, an algorithm which reduces the training data to a so-called [...]</p>
]]></content:encoded>
	</item>
	<item>
		<title>Comment on Heidegger's prophecy, or: Europe is quite badly fucked by LLMs, the useful parts &#171; The Tar Pit</title>
		<link>http://thetarpit.org/2022/heideggers-prophecy-or-europe-is-quite-badly-fucked#comment-7619</link>
		<dc:creator>LLMs, the useful parts &#171; The Tar Pit</dc:creator>
		<pubDate>Sun, 26 Jul 2026 18:53:54 +0000</pubDate>
		<guid isPermaLink="false">http://thetarpit.org/?p=483#comment-7619</guid>
		<description>[...] say, try to discuss Heidegger with a LLM, and at worst you'll get some regurgitated crap; while at best you'll get references to [...]</description>
		<content:encoded><![CDATA[<p>[...] say, try to discuss Heidegger with a LLM, and at worst you'll get some regurgitated crap; while at best you'll get references to [...]</p>
]]></content:encoded>
	</item>
	<item>
		<title>Comment on On perversion by LLMs, the useful parts &#171; The Tar Pit</title>
		<link>http://thetarpit.org/2026/on-perversion#comment-7618</link>
		<dc:creator>LLMs, the useful parts &#171; The Tar Pit</dc:creator>
		<pubDate>Sun, 26 Jul 2026 18:53:39 +0000</pubDate>
		<guid isPermaLink="false">http://thetarpit.org/?p=582#comment-7618</guid>
		<description>[...] that LLMs are able to think in the proper sense, and that their association with "thought" is a perversion, I don't claim to have any sort of ideological beef with them. Quite the contrary, in fact: all the [...]</description>
		<content:encoded><![CDATA[<p>[...] that LLMs are able to think in the proper sense, and that their association with "thought" is a perversion, I don't claim to have any sort of ideological beef with them. Quite the contrary, in fact: all the [...]</p>
]]></content:encoded>
	</item>
	<item>
		<title>Comment on Artificial intelligence by LLMs, the useful parts &#171; The Tar Pit</title>
		<link>http://thetarpit.org/2024/artificial-intelligence#comment-7617</link>
		<dc:creator>LLMs, the useful parts &#171; The Tar Pit</dc:creator>
		<pubDate>Sun, 26 Jul 2026 18:53:20 +0000</pubDate>
		<guid isPermaLink="false">http://thetarpit.org/?p=540#comment-7617</guid>
		<description>[...] already been more than two years since I last shat on the whole subject of AI, and also more than a year since my last brief commentary on the matter. [...]</description>
		<content:encoded><![CDATA[<p>[...] already been more than two years since I last shat on the whole subject of AI, and also more than a year since my last brief commentary on the matter. [...]</p>
]]></content:encoded>
	</item>
	<item>
		<title>Comment on Phonopolis by spyked</title>
		<link>http://thetarpit.org/2026/phonopolis#comment-7525</link>
		<dc:creator>spyked</dc:creator>
		<pubDate>Thu, 18 Jun 2026 18:19:58 +0000</pubDate>
		<guid isPermaLink="false">http://thetarpit.org/?p=592#comment-7525</guid>
		<description>Heh. I used to know someone named Tchaikovsky.</description>
		<content:encoded><![CDATA[<p>Heh. I used to know someone named Tchaikovsky.</p>
]]></content:encoded>
	</item>
	<item>
		<title>Comment on Phonopolis by Cel Mihanie</title>
		<link>http://thetarpit.org/2026/phonopolis#comment-7524</link>
		<dc:creator>Cel Mihanie</dc:creator>
		<pubDate>Thu, 18 Jun 2026 18:15:29 +0000</pubDate>
		<guid isPermaLink="false">http://thetarpit.org/?p=592#comment-7524</guid>
		<description>Loom is based. It was my first contact, I think, with Tchaikovsky's genius music. You better download it and store it safely in the Templar Archives. Maybe they'll ban it someday. Can't have music composed by a russkie, that's bad, dontcha know :D</description>
		<content:encoded><![CDATA[<p>Loom is based. It was my first contact, I think, with Tchaikovsky's genius music. You better download it and store it safely in the Templar Archives. Maybe they'll ban it someday. Can't have music composed by a russkie, that's bad, dontcha know :D</p>
]]></content:encoded>
	</item>
	<item>
		<title>Comment on Phonopolis by spyked</title>
		<link>http://thetarpit.org/2026/phonopolis#comment-7523</link>
		<dc:creator>spyked</dc:creator>
		<pubDate>Thu, 18 Jun 2026 17:08:04 +0000</pubDate>
		<guid isPermaLink="false">http://thetarpit.org/?p=592#comment-7523</guid>
		<description>Hm, now that you mentioned it, I need to review Loom (tm) sometime, it's been quite a while since I played it.</description>
		<content:encoded><![CDATA[<p>Hm, now that you mentioned it, I need to review Loom (tm) sometime, it's been quite a while since I played it.</p>
]]></content:encoded>
	</item>
	<item>
		<title>Comment on Phonopolis by Cel Mihanie</title>
		<link>http://thetarpit.org/2026/phonopolis#comment-7521</link>
		<dc:creator>Cel Mihanie</dc:creator>
		<pubDate>Thu, 18 Jun 2026 16:13:37 +0000</pubDate>
		<guid isPermaLink="false">http://thetarpit.org/?p=592#comment-7521</guid>
		<description>And thus more games are added to my, ah, monotonically increasing to-play list. Although since you mentioned sound plays a key part in the story, this reminded me of Loom, so I'd better revisit that first</description>
		<content:encoded><![CDATA[<p>And thus more games are added to my, ah, monotonically increasing to-play list. Although since you mentioned sound plays a key part in the story, this reminded me of Loom, so I'd better revisit that first</p>
]]></content:encoded>
	</item>
</channel>
</rss>
