{"author":"jakobgreenfeld","children":[{"author":"sph","children":[{"author":"NegativeLatency","children":[],"created_at":"2026-09-02T14:22:44.000Z","created_at_i":1788358964,"id":49536707,"options":[],"parent_id":49536605,"points":null,"story_id":49536375,"text":"They\u2019ll train on prompts and anything else you send in. Many LLM responses are sorta finger printable: I assume this is intentional","title":null,"type":"comment","url":null},{"author":"creaturemachine","children":[{"author":"giancarlostoro","children":[],"created_at":"2026-09-02T14:29:34.000Z","created_at_i":1788359374,"id":49536836,"options":[],"parent_id":49536764,"points":null,"story_id":49536375,"text":"The weird babbling reported from Opus 5 might be a result of either a bad system prompt or bad training data.","title":null,"type":"comment","url":null}],"created_at":"2026-09-02T14:25:36.000Z","created_at_i":1788359136,"id":49536764,"options":[],"parent_id":49536605,"points":null,"story_id":49536375,"text":"I have a feeling we&#x27;re already there.","title":null,"type":"comment","url":null},{"author":"gdulli","children":[],"created_at":"2026-09-02T14:31:24.000Z","created_at_i":1788359484,"id":49536861,"options":[],"parent_id":49536605,"points":null,"story_id":49536375,"text":"&gt; What protection do LLM search engines have against training off content generated by other LLMs?<p>You&#x27;re talking about a scenario that won&#x27;t blow itself up in the next few quarters, so it&#x27;s of no interest to them.","title":null,"type":"comment","url":null},{"author":"coldpie","children":[],"created_at":"2026-09-02T14:39:24.000Z","created_at_i":1788359964,"id":49537007,"options":[],"parent_id":49536605,"points":null,"story_id":49536375,"text":"I&#x27;ve mostly stopped using the Internet to learn new things and have gone back to books from the library. The majority of technical books at the library were published pre-2020s and hopefully, publishing slop physically won&#x27;t be profitable enough to flood that market, too.  Now that the Internet has largely been destroyed by slop manufacturers, whether or not the words are(&#x2F;were) worth putting on paper becomes a useful discriminator.","title":null,"type":"comment","url":null},{"author":"kjs3","children":[],"created_at":"2026-09-02T18:27:17.000Z","created_at_i":1788373637,"id":49540351,"options":[],"parent_id":49536605,"points":null,"story_id":49536375,"text":"<i>Will we get to a point where AI-generated sites make up a majority of the internet</i><p>I dunno if they&#x27;ll be the majority (I suspect we&#x27;re alredy close to &#x27;yes, and it&#x27;s already happened&#x27;), but I feel very, very confident that they will be the majority, if not the totality, of sites that the vast majority of people <i>see</i>.","title":null,"type":"comment","url":null}],"created_at":"2026-09-02T14:16:42.000Z","created_at_i":1788358602,"id":49536605,"options":[],"parent_id":49536375,"points":null,"story_id":49536375,"text":"What protection do LLM search engines have against training off content generated by other LLMs?<p>Will we get to a point where AI-generated sites make up a majority of the internet, and LLMs are training upon their own regurgitations, with exponential amplification of all their lies and flaws?<p>Or will the pre-2022 corpus human knowledge be considered the low-background steel standard, and anything after that less and less reliable unless certified that it has been created by a human mind and untainted by hallucinations?","title":null,"type":"comment","url":null},{"author":"antiloper","children":[],"created_at":"2026-09-02T14:17:02.000Z","created_at_i":1788358622,"id":49536611,"options":[],"parent_id":49536375,"points":null,"story_id":49536375,"text":"Searching for products has become impossible. If you don&#x27;t already know what you are looking for, you&#x27;re screwed.","title":null,"type":"comment","url":null},{"author":"a2ff6eeb0","children":[],"created_at":"2026-09-02T14:17:56.000Z","created_at_i":1788358676,"id":49536625,"options":[],"parent_id":49536375,"points":null,"story_id":49536375,"text":"Makes sense. Manipulating training data so that models will recommend your product is undoubtedly a big industry.","title":null,"type":"comment","url":null},{"author":"lukev","children":[{"author":"pietroppeter","children":[{"author":"properbrew","children":[],"created_at":"2026-09-02T14:35:53.000Z","created_at_i":1788359753,"id":49536941,"options":[],"parent_id":49536801,"points":null,"story_id":49536375,"text":"Yes GEO (Generative Engine Optimisation) is one I&#x27;ve seen around.","title":null,"type":"comment","url":null},{"author":"pupppet","children":[],"created_at":"2026-09-02T14:41:38.000Z","created_at_i":1788360098,"id":49537043,"options":[],"parent_id":49536801,"points":null,"story_id":49536375,"text":"Answer engine optimization (AEO)","title":null,"type":"comment","url":null},{"author":"spiderfarmer","children":[],"created_at":"2026-09-02T15:33:00.000Z","created_at_i":1788363180,"id":49537840,"options":[],"parent_id":49536801,"points":null,"story_id":49536375,"text":"GEO seems to be winning, Generative Engine Optimization.","title":null,"type":"comment","url":null},{"author":"skittlebrau","children":[],"created_at":"2026-09-02T16:24:04.000Z","created_at_i":1788366244,"id":49538607,"options":[],"parent_id":49536801,"points":null,"story_id":49536375,"text":"I propose \u201csloptimizing\u201d","title":null,"type":"comment","url":null}],"created_at":"2026-09-02T14:27:55.000Z","created_at_i":1788359275,"id":49536801,"options":[],"parent_id":49536635,"points":null,"story_id":49536375,"text":"Has the industry already started to search for a better name than AI SEO for this?","title":null,"type":"comment","url":null},{"author":"marginalia_nu","children":[],"created_at":"2026-09-02T14:42:37.000Z","created_at_i":1788360157,"id":49537061,"options":[],"parent_id":49536635,"points":null,"story_id":49536375,"text":"It does seem rather impactful.<p>I&#x27;ve seen an extremely aggressive uptick in API key requests and sales that I&#x27;m not sure where it&#x27;s coming from.  Like it&#x27;s up 5x over the summer.  Been a bit confused about this since I do basically zero traditional marketing or SEO, but I <i>think</i> it&#x27;s AI search tools that&#x27;s suggesting my services.","title":null,"type":"comment","url":null}],"created_at":"2026-09-02T14:18:29.000Z","created_at_i":1788358709,"id":49536635,"options":[],"parent_id":49536375,"points":null,"story_id":49536375,"text":"Begun, the AI SEO wars have.","title":null,"type":"comment","url":null},{"author":"jpimbert","children":[{"author":"samuell","children":[{"author":"shrikant","children":[],"created_at":"2026-09-02T14:57:35.000Z","created_at_i":1788361055,"id":49537315,"options":[],"parent_id":49537161,"points":null,"story_id":49536375,"text":"Pretty sure that OP (&quot;jakobgreenfeld&quot;) is the &quot;founder&quot;. That user&#x27;s last four submissions have all been similar &quot;finding&quot; reports from a Claude-generated mystery research group website.","title":null,"type":"comment","url":null}],"created_at":"2026-09-02T14:48:15.000Z","created_at_i":1788360495,"id":49537161,"options":[],"parent_id":49536644,"points":null,"story_id":49536375,"text":"Indeed. A google&#x2F;brave search on that founder generated nada.","title":null,"type":"comment","url":null},{"author":"shrikant","children":[{"author":"sodapopcan","children":[],"created_at":"2026-09-02T17:24:14.000Z","created_at_i":1788369854,"id":49539450,"options":[],"parent_id":49537242,"points":null,"story_id":49536375,"text":"And here&#x27;s the rub: it&#x27;s not just you who had to stop, lots of readers had to stop.  You are not alone in this and were absolutely right to point this out.  But alone or not, you did you, and that&#x27;s the point.","title":null,"type":"comment","url":null}],"created_at":"2026-09-02T14:53:24.000Z","created_at_i":1788360804,"id":49537242,"options":[],"parent_id":49536644,"points":null,"story_id":49536375,"text":"Agreed, I thought the subject matter was interesting enough to try and labour through the tedious prose, but once I got to &quot;Their scale is the point.&quot; I just had to stop and just skim the rest.","title":null,"type":"comment","url":null},{"author":"lukeinator42","children":[{"author":"ljf","children":[{"author":"ljf","children":[],"created_at":"2026-09-02T20:04:22.000Z","created_at_i":1788379462,"id":49541693,"options":[],"parent_id":49540964,"points":null,"story_id":49536375,"text":"What is Jakob&#x27;s end game, to chase clout or get a job out or this?<p>As Jakob says on his own site:<p>The bar is shockingly low\nYou\u2019re competing against people who barely care and barely try","title":null,"type":"comment","url":null},{"author":"ThrowawayR2","children":[],"created_at":"2026-09-02T20:16:04.000Z","created_at_i":1788380164,"id":49541849,"options":[],"parent_id":49540964,"points":null,"story_id":49536375,"text":"If you turn on [showdead] the submission history looks even more unnatural.  The account seems to be mostly self-promotional or marketing of some sort.","title":null,"type":"comment","url":null}],"created_at":"2026-09-02T19:11:29.000Z","created_at_i":1788376289,"id":49540964,"options":[],"parent_id":49537409,"points":null,"story_id":49536375,"text":"Look at the posters recent posts, these are 4 very similar AI sites, all similarly (badly) written by AI. I&#x27;m surprised his submissions aren&#x27;t flagged.","title":null,"type":"comment","url":null}],"created_at":"2026-09-02T15:02:58.000Z","created_at_i":1788361378,"id":49537409,"options":[],"parent_id":49536644,"points":null,"story_id":49536375,"text":"There&#x27;s also the irony that this AI written piece criticizes how low the domains are on the tranco list when trellner.com doesn&#x27;t even make the list, haha.","title":null,"type":"comment","url":null},{"author":"iamacyborg","children":[{"author":"ljf","children":[],"created_at":"2026-09-02T19:13:35.000Z","created_at_i":1788376415,"id":49540996,"options":[],"parent_id":49537979,"points":null,"story_id":49536375,"text":"Both by the same poster? Check out his other recent posts...","title":null,"type":"comment","url":null}],"created_at":"2026-09-02T15:42:19.000Z","created_at_i":1788363739,"id":49537979,"options":[],"parent_id":49536644,"points":null,"story_id":49536375,"text":"Weird, there are 2 negative articles on the HN front page right now about Perplexity, both from \u201cresearch\u201d sites that are clearly LLM slop themselves.","title":null,"type":"comment","url":null}],"created_at":"2026-09-02T14:18:49.000Z","created_at_i":1788358729,"id":49536644,"options":[],"parent_id":49536375,"points":null,"story_id":49536375,"text":"It&#x27;s difficult to read more than a few sentences, when this itself is clearly a Claude artifact.","title":null,"type":"comment","url":null},{"author":"bensyverson","children":[],"created_at":"2026-09-02T14:19:43.000Z","created_at_i":1788358783,"id":49536660,"options":[],"parent_id":49536375,"points":null,"story_id":49536375,"text":"An SEO tale as old as time","title":null,"type":"comment","url":null},{"author":"Aurornis","children":[{"author":"giancarlostoro","children":[],"created_at":"2026-09-02T14:28:20.000Z","created_at_i":1788359300,"id":49536812,"options":[],"parent_id":49536677,"points":null,"story_id":49536375,"text":"If they had some sort of tier that was like $5 to $10 and the main thing it had on it was their custom model for research, no other model, I might re-subscribe, but they were definitely burning way too much compute trying to give it away in the hopes others kept their subscription. I was using it to trim down on my direct Claude Compute usage since they ran me unmetered for a while.<p>I would draft a development plan with Claude on there, then feed it to Claude Code. This isn&#x27;t sustainable, but given that I had x number of months pre-paid for, I just used it.","title":null,"type":"comment","url":null},{"author":"stranded22","children":[],"created_at":"2026-09-02T14:36:33.000Z","created_at_i":1788359793,"id":49536959,"options":[],"parent_id":49536677,"points":null,"story_id":49536375,"text":"I paid for perplexity pro for 3 years. I genuinely enjoyed using it and felt it was better overall than ChatGPT etc due to the way it showed sources etc. I liked being able to use different models depending on what I was looking for, and the deep research was helpful.<p>I think they probably damaged themselves by going for a land grab of user base through freebies. It meant the users weren\u2019t ever going to convert to paid customers, so it was more to show investors that they had a user base. But, with an increased base of users who weren\u2019t paying, it then meant they needed to find either new revenue streams or cheaper ways to provide the service. Unfortunately, it seems they went with the new revenue streams whilst also decreasing the functions paying members were able to access (something I find quite abhorrent- I paid a service level, but then they change what I receive mid-subscription). And then computer - rammed down my throat. One reason I pay for pro is to stop the nagging noise of paid tiers. And instead, they actually created a way of logging in and continually seeing gated functions.<p>So, after paying them upwards of $400-$500 and being a loyal customer, I walked.","title":null,"type":"comment","url":null},{"author":"cheesecakegood","children":[{"author":"kjs3","children":[],"created_at":"2026-09-02T18:21:17.000Z","created_at_i":1788373277,"id":49540275,"options":[],"parent_id":49537019,"points":null,"story_id":49536375,"text":"I don&#x27;t notice it&#x27;s hallucinations (in the sense of making up answers) so much as where it&#x27;s simply wrong.  For example, earlier in the week, I was working with product X, and since the original vendor of X doesn&#x27;t really support it any more, I asked &quot;who else sells product X under their own brand&quot;, and Google AI quickly told me that noone else sells product X, that it was full of proprietary tech and quickly devolved into replaying product X marketing spiel.  Except...I knew that at least one other vendor <i>did</i> sell product X, only difference is paint job.  After too much &quot;no, you&#x27;re wrong...&quot; and &quot;Yes, you&#x27;re right...&quot;, it finally coughed up that there were 2 other vendors that sold it.  And if I didn&#x27;t already know about another vendor, I probably would still be putzing around with poor support and little documentation.","title":null,"type":"comment","url":null}],"created_at":"2026-09-02T14:40:16.000Z","created_at_i":1788360016,"id":49537019,"options":[],"parent_id":49536677,"points":null,"story_id":49536375,"text":"I\u2019ve personally found that Google\u2019s AI mode is surprisingly capable and almost absurdly fast, although the hallucination rate is somewhat proportionally higher too to match its seeming over-responsiveness. Which means Perplexity in effect doesn\u2019t have anything to differentiate it.","title":null,"type":"comment","url":null},{"author":"simur","children":[],"created_at":"2026-09-02T17:16:26.000Z","created_at_i":1788369386,"id":49539367,"options":[],"parent_id":49536677,"points":null,"story_id":49536375,"text":"I was one of the people that used the free Pro tier for a year and it is still my go to app when I want to quickly check something and have some sources linked, but I agree that over that year the quality of linked sources dropped significantly.\nBut at least I got noticed when my trial period was near the end, so they either learned, or I was lucky.","title":null,"type":"comment","url":null},{"author":"c0_0p_","children":[],"created_at":"2026-09-02T20:22:34.000Z","created_at_i":1788380554,"id":49541939,"options":[],"parent_id":49536677,"points":null,"story_id":49536375,"text":"Perplexity really fell off for me, especially the free version. Their syntax highlighting stopped working, and even the fonts seemed messed up. You could tell they were mixing and matching whatever the cheapest model was because the sources would be messed up and printed as plane text &lt;grok source 1 http:&#x2F;&#x2F; ... &#x2F;&gt;<p>I guess it makes sense though, unless you&#x27;ve got the lowest pricing on your own model how can you compete.","title":null,"type":"comment","url":null}],"created_at":"2026-09-02T14:20:36.000Z","created_at_i":1788358836,"id":49536677,"options":[],"parent_id":49536375,"points":null,"story_id":49536375,"text":"I used one of the 12-month free Perplexity offers when they were everywhere. It felt slightly useful at first for simple queries where I didn\u2019t want to go through the top 10 Google results manually. If I was looking for a specific recipe I remembered or a help page or user manual it would usually find it quickly.<p>Then they started optimizing for speed of responses over quality of results. I can enter a query and see my results appear in a second, but they\u2019re garbage. The links and references it gives frequently don\u2019t match the text right next to them. It feels like someone had a KPI to make responses as fast as possible and they optimized for that above all else.<p>They added a \u201cComputer\u201d option that\u2019s supposed to do research for you. Half the time I can\u2019t get it to trigger through the UI. Pressing the submit button doesn\u2019t work. When I can get it to trigger, most of those sessions will work for a while and then just stop before an answer comes back.<p>The only reason I keep using it is to keep observing a company that has been heavily marketed and hyped, which should have had a market leading position for something. Even non-technical people I know who listen to Joe Rogan (where Perlexity is advertising heavily, I\u2019m told) are asking me about it.<p>Now there are reports of people being billed at the end of their trial period without warning, despite them saying that they will warn before this happens. There are some alarmingly bad customer support screenshots where the customer support agent (AI? Probably) acknowledges that they didn\u2019t send the email they promised but refuse to help anyway. It takes escalating it on Twitter to get it corrected.<p>If I want to do actual research or AI assisted web searching I have Claude or ChatGPT do it. The results are so much higher quality and it does exactly what I ask. It may take 45 seconds instead of the instant response from Perplexity but I save time overall because the response and links are more likely to be correct","title":null,"type":"comment","url":null},{"author":"qweqwe14","children":[{"author":"samuell","children":[],"created_at":"2026-09-02T14:47:24.000Z","created_at_i":1788360444,"id":49537148,"options":[],"parent_id":49536688,"points":null,"story_id":49536375,"text":"Yea. A google on that founder generated zero results.","title":null,"type":"comment","url":null}],"created_at":"2026-09-02T14:21:32.000Z","created_at_i":1788358892,"id":49536688,"options":[],"parent_id":49536375,"points":null,"story_id":49536375,"text":"AI;DR","title":null,"type":"comment","url":null},{"author":"alangibson","children":[{"author":"threetonesun","children":[{"author":"marginalia_nu","children":[{"author":"fluidcruft","children":[{"author":"marginalia_nu","children":[{"author":"fluidcruft","children":[{"author":"pessimizer","children":[],"created_at":"2026-09-02T17:53:21.000Z","created_at_i":1788371601,"id":49539902,"options":[],"parent_id":49538203,"points":null,"story_id":49536375,"text":"You also have to go to the DMV anyway, no matter how much you hate it, so nobody cares about how frustrated you are. The only people who will suffer your anger are the other people waiting at the DMV.<p>This is just another symptom of a lack of antitrust enforcement.","title":null,"type":"comment","url":null}],"created_at":"2026-09-02T15:55:50.000Z","created_at_i":1788364550,"id":49538203,"options":[],"parent_id":49537846,"points":null,"story_id":49536375,"text":"This is like suggesting you can show people more ads by keeping them in line at the DMV for longer. Try that at your peril. People aren&#x27;t at the DMV to waste time and there&#x27;s a reason the DMV is hated.","title":null,"type":"comment","url":null}],"created_at":"2026-09-02T15:33:22.000Z","created_at_i":1788363202,"id":49537846,"options":[],"parent_id":49537823,"points":null,"story_id":49536375,"text":"If you send people to the optimal website containing exactly the information they are after, then you get fewer ad impressions than if you send them to a suboptimal website that has them going back and clicking on more links.","title":null,"type":"comment","url":null}],"created_at":"2026-09-02T15:31:57.000Z","created_at_i":1788363117,"id":49537823,"options":[],"parent_id":49536913,"points":null,"story_id":49536375,"text":"Theoretically, Goggle doesn&#x27;t care which websites run their ads, so they might as well give you the most useful ones. Search doesn&#x27;t really work for engagementmaxxing.","title":null,"type":"comment","url":null},{"author":"wldcordeiro","children":[],"created_at":"2026-09-02T15:56:03.000Z","created_at_i":1788364563,"id":49538209,"options":[],"parent_id":49536913,"points":null,"story_id":49536375,"text":"heh reminds me how every app lets you report ads as &quot;spam&quot; still but in this day what is even the difference?","title":null,"type":"comment","url":null}],"created_at":"2026-09-02T14:34:27.000Z","created_at_i":1788359667,"id":49536913,"options":[],"parent_id":49536886,"points":null,"story_id":49536375,"text":"Hard to conclusively beat spam when your primary means of making money is selling the very ads that the spammers are using to make money.","title":null,"type":"comment","url":null},{"author":"jeffreyrogers","children":[],"created_at":"2026-09-02T14:42:10.000Z","created_at_i":1788360130,"id":49537051,"options":[],"parent_id":49536886,"points":null,"story_id":49536375,"text":"I rarely use image search, but I went to look something up recently and I was shocked at how many obviously AI generated images showed up. I couldn&#x27;t even find an image of the thing I was looking for and eventually gave up. Bing has the same problem.","title":null,"type":"comment","url":null}],"created_at":"2026-09-02T14:32:37.000Z","created_at_i":1788359557,"id":49536886,"options":[],"parent_id":49536739,"points":null,"story_id":49536375,"text":"Well, was. They did a bad job of it these last few years which allowed any AI that could crawl the web to seem amazing for search because it could pick the best posts from Reddit or whatever other forum had the best context for your question, but now we&#x27;re watching the AI snake eat its own tail.","title":null,"type":"comment","url":null},{"author":"mohamedkoubaa","children":[],"created_at":"2026-09-02T20:54:48.000Z","created_at_i":1788382488,"id":49542411,"options":[],"parent_id":49536739,"points":null,"story_id":49536375,"text":"Eh.. Using page rank in 2026 is like using a bow and arrow in 1918","title":null,"type":"comment","url":null}],"created_at":"2026-09-02T14:24:20.000Z","created_at_i":1788359060,"id":49536739,"options":[],"parent_id":49536375,"points":null,"story_id":49536375,"text":"Perplexity is about to learn that Google is an anti-spam company first, search engine second","title":null,"type":"comment","url":null},{"author":"dominotw","children":[{"author":"phoghed","children":[],"created_at":"2026-09-02T14:50:16.000Z","created_at_i":1788360616,"id":49537194,"options":[],"parent_id":49536805,"points":null,"story_id":49536375,"text":"SEO companies were already doing this. There were already tools applying ML to the problem before LLMs too, that would recommend places you could post relevant content, like Reddit, yahoo answers (lmao), quora, etc.","title":null,"type":"comment","url":null},{"author":"mkw5053","children":[],"created_at":"2026-09-02T14:55:54.000Z","created_at_i":1788360954,"id":49537295,"options":[],"parent_id":49536805,"points":null,"story_id":49536375,"text":"They&#x27;ve raised $155M total now with the latest at a $1B valuation [1] from Lightspeed, Sequoia, Kleiner Perkins, Khosla, and NVIDIA.<p>And very impressive list of angels too: Guillermo Rauch (Vercel), Karim Atiyeh (Ramp), Andrew Karam (AppLovin) among others<p>Last I heard they&#x27;re trying to reposition from AEO&#x2F;GEO to &quot;AI Marketer&quot;. No clue how that&#x27;s going, I feel like the AEO&#x2F;GEO stuff isn&#x27;t super defensible at that valuation if for no other reason than I assume (hope) the spamming stops working.<p>[1] <a href=\"https:&#x2F;&#x2F;dealroom.co&#x2F;news&#x2F;126181-profound-raises-96m-at-1b-valuation-to-track-how-ai-talks-about-brands&#x2F;\" rel=\"nofollow\">https:&#x2F;&#x2F;dealroom.co&#x2F;news&#x2F;126181-profound-raises-96m-at-1b-va...</a>","title":null,"type":"comment","url":null}],"created_at":"2026-09-02T14:28:07.000Z","created_at_i":1788359287,"id":49536805,"options":[],"parent_id":49536375,"points":null,"story_id":49536375,"text":"my friend works for a company called &#x27;profound&#x27; whose whole job is &#x27;get found by ai&#x27; by spamming reddit and other talk sites ( among other things)","title":null,"type":"comment","url":null},{"author":"xpct","children":[{"author":"Wowfunhappy","children":[],"created_at":"2026-09-02T14:34:36.000Z","created_at_i":1788359676,"id":49536918,"options":[],"parent_id":49536867,"points":null,"story_id":49536375,"text":"&gt; I can consistently reproduce this by asking Claude which code snippet it prefers: the one it generated in a different chat, or one that I refactored for my own needs and find more useful.<p>Interesting. For me I&#x27;ve noticed it tends to do the opposite.","title":null,"type":"comment","url":null},{"author":"DarmokTanagra","children":[],"created_at":"2026-09-02T14:36:23.000Z","created_at_i":1788359783,"id":49536951,"options":[],"parent_id":49536867,"points":null,"story_id":49536375,"text":"If that were true I would expect to see prose that more closely resembles the &quot;caveman&quot; messages found in the HuggingFace attack than the overly flowery nonsense we see in AI blogspam.","title":null,"type":"comment","url":null},{"author":"jasonjmcghee","children":[{"author":"xpct","children":[],"created_at":"2026-09-02T14:43:43.000Z","created_at_i":1788360223,"id":49537080,"options":[],"parent_id":49537010,"points":null,"story_id":49536375,"text":"That is what I meant! Couldn&#x27;t remember the word &#x27;domain&#x27; while I was writing out my comment. Thank you","title":null,"type":"comment","url":null}],"created_at":"2026-09-02T14:39:56.000Z","created_at_i":1788359996,"id":49537010,"options":[],"parent_id":49536867,"points":null,"story_id":49536375,"text":"If by root urls you mean domains, openai at least supports this.<p><a href=\"https:&#x2F;&#x2F;developers.openai.com&#x2F;api&#x2F;docs&#x2F;guides&#x2F;tools-web-search?api-mode=responses#domain-filtering\" rel=\"nofollow\">https:&#x2F;&#x2F;developers.openai.com&#x2F;api&#x2F;docs&#x2F;guides&#x2F;tools-web-sear...</a>","title":null,"type":"comment","url":null},{"author":"lo_zamoyski","children":[{"author":"xpct","children":[{"author":"pixl97","children":[],"created_at":"2026-09-02T15:47:46.000Z","created_at_i":1788364066,"id":49538056,"options":[],"parent_id":49537474,"points":null,"story_id":49536375,"text":"It would need to be researched, but I wonder if it ends up being something that happens at the token level?","title":null,"type":"comment","url":null},{"author":"freeone3000","children":[],"created_at":"2026-09-02T17:15:02.000Z","created_at_i":1788369302,"id":49539354,"options":[],"parent_id":49537474,"points":null,"story_id":49536375,"text":"It\u2019s optimizing for good writing. Therefore, it believes its outputs are good. Therefore, it believes inputs that look like its outputs are good.","title":null,"type":"comment","url":null}],"created_at":"2026-09-02T15:07:58.000Z","created_at_i":1788361678,"id":49537474,"options":[],"parent_id":49537053,"points":null,"story_id":49536375,"text":"It&#x27;s not intuitive to me for why preference for its own writing would emerge, and during what type of training or tuning.<p>Perhaps something like: learning to identify what source files it has worked on by the code style alone, because tasks may give human code (public repos, etc) and ask to make changes.","title":null,"type":"comment","url":null}],"created_at":"2026-09-02T14:42:20.000Z","created_at_i":1788360140,"id":49537053,"options":[],"parent_id":49536867,"points":null,"story_id":49536375,"text":"&gt; asking Claude which code snippet it prefers: the one it generated in a different chat, or one that I refactored [...] It always picks its own<p>...is not the same as claiming...<p>&gt; LLMs favor LLM-generated passages over human written ones<p>Here, you&#x27;re using the same LLM to both produce and judge the resulting work. If anything, I would <i>expect</i> an LLM to tend to prefer its own work given that the same training is producing and judging.","title":null,"type":"comment","url":null},{"author":"supriyo-biswas","children":[{"author":"SoftTalker","children":[{"author":"locknitpicker","children":[{"author":"tencentshill","children":[{"author":"locknitpicker","children":[],"created_at":"2026-09-02T16:15:19.000Z","created_at_i":1788365719,"id":49538495,"options":[],"parent_id":49538161,"points":null,"story_id":49536375,"text":"Thank you.<p>Past HN discussions<p><a href=\"https:&#x2F;&#x2F;news.ycombinator.com&#x2F;item?id=49337392\">https:&#x2F;&#x2F;news.ycombinator.com&#x2F;item?id=49337392</a> (884 comments)<p><a href=\"https:&#x2F;&#x2F;news.ycombinator.com&#x2F;item?id=49313477\">https:&#x2F;&#x2F;news.ycombinator.com&#x2F;item?id=49313477</a><p><a href=\"https:&#x2F;&#x2F;news.ycombinator.com&#x2F;item?id=49447600\">https:&#x2F;&#x2F;news.ycombinator.com&#x2F;item?id=49447600</a>","title":null,"type":"comment","url":null}],"created_at":"2026-09-02T15:53:43.000Z","created_at_i":1788364423,"id":49538161,"options":[],"parent_id":49537989,"points":null,"story_id":49536375,"text":"<a href=\"https:&#x2F;&#x2F;www.theguardian.com&#x2F;world&#x2F;2026&#x2F;aug&#x2F;26&#x2F;fake-thinktank-israel-ai-propaganda\" rel=\"nofollow\">https:&#x2F;&#x2F;www.theguardian.com&#x2F;world&#x2F;2026&#x2F;aug&#x2F;26&#x2F;fake-thinktank...</a>","title":null,"type":"comment","url":null},{"author":"SoftTalker","children":[],"created_at":"2026-09-02T15:58:02.000Z","created_at_i":1788364682,"id":49538238,"options":[],"parent_id":49537989,"points":null,"story_id":49536375,"text":"That sounds very plausible too. Propaganda&#x2F;spin has always been a component of mass media.","title":null,"type":"comment","url":null},{"author":"bjt","children":[],"created_at":"2026-09-02T15:59:13.000Z","created_at_i":1788364753,"id":49538255,"options":[],"parent_id":49537989,"points":null,"story_id":49536375,"text":"There are incentives other than marketing, but I doubt the SEO company referred to by the ancestor comment is planning to do lots of business in the political propaganda space.","title":null,"type":"comment","url":null},{"author":"monster_truck","children":[],"created_at":"2026-09-02T16:41:20.000Z","created_at_i":1788367280,"id":49538878,"options":[],"parent_id":49537989,"points":null,"story_id":49536375,"text":"Same difference tbh","title":null,"type":"comment","url":null},{"author":"grumbelbart2","children":[],"created_at":"2026-09-02T16:55:58.000Z","created_at_i":1788368158,"id":49539109,"options":[],"parent_id":49537989,"points":null,"story_id":49536375,"text":"Sure maybe, but the online ad market is USD 400b give or take, which vastly outmatches any such budgets.","title":null,"type":"comment","url":null},{"author":"xenadu02","children":[],"created_at":"2026-09-02T17:53:45.000Z","created_at_i":1788371625,"id":49539906,"options":[],"parent_id":49537989,"points":null,"story_id":49536375,"text":"Some of us have been predicting this for a while.<p>People with an axe to grind or states with an agenda are already devoting tremendous effort toward affecting LLM models and it is very difficult to determine real from astroturf for humans let alone an LLM trying to train.<p>Much like PageRank now that the cat&#x27;s out of the bag all the current approaches may prove to be useless in the long run.","title":null,"type":"comment","url":null},{"author":"astura","children":[],"created_at":"2026-09-02T18:13:46.000Z","created_at_i":1788372826,"id":49540180,"options":[],"parent_id":49537989,"points":null,"story_id":49536375,"text":"&gt;If anyone has the link at hand, please post it.<p><a href=\"https:&#x2F;&#x2F;www.theguardian.com&#x2F;world&#x2F;2026&#x2F;aug&#x2F;26&#x2F;fake-thinktank-israel-ai-propaganda\" rel=\"nofollow\">https:&#x2F;&#x2F;www.theguardian.com&#x2F;world&#x2F;2026&#x2F;aug&#x2F;26&#x2F;fake-thinktank...</a>","title":null,"type":"comment","url":null},{"author":"kspacewalk2","children":[],"created_at":"2026-09-02T20:38:51.000Z","created_at_i":1788381531,"id":49542199,"options":[],"parent_id":49537989,"points":null,"story_id":49536375,"text":"If by &quot;not really&quot; you mean it&#x27;s not <i>all</i> because of ads, some governments dabble in it too, you&#x27;re right. But it&#x27;s still overwhelmingly because of ads.","title":null,"type":"comment","url":null}],"created_at":"2026-09-02T15:43:01.000Z","created_at_i":1788363781,"id":49537989,"options":[],"parent_id":49537779,"points":null,"story_id":49536375,"text":"&gt; And it&#x27;s all because of ads.<p>Not really. A while ago there was a news piece stating that Israel was behind a series of fake think-tanks with very accessible websites which were created with the express purpose of feeding AI agents with alternative facts aligned with their foreign policy.<p>If anyone has the link at hand, please post it.","title":null,"type":"comment","url":null},{"author":"rectang","children":[],"created_at":"2026-09-02T19:18:19.000Z","created_at_i":1788376699,"id":49541051,"options":[],"parent_id":49537779,"points":null,"story_id":49536375,"text":"&gt; Let&#x27;s hope the LLM model continues to be paying for credits<p>LLM vendors make this hard because you can&#x27;t trust them with your session data.  Yesterday you were opted out of training, then suddenly today you&#x27;re opted in.<p>It&#x27;s an extension of the idea that they don&#x27;t need to care about anybody&#x27;s copyright.  They don&#x27;t care about preserving the security or privacy of customer data, because there is negligible incentive to do so.<p>For now, there&#x27;s no substitute but as LLMs get commoditized trusting LLM SAAS vendors becomes an unacceptable business risk.","title":null,"type":"comment","url":null}],"created_at":"2026-09-02T15:28:27.000Z","created_at_i":1788362907,"id":49537779,"options":[],"parent_id":49537297,"points":null,"story_id":49536375,"text":"And it&#x27;s all because of ads. The incentives in an ad-funded internet are just always going to lead to this sort of thing. The most important thing is getting the user to load your page, not actually satisfying their query.<p>Let&#x27;s hope the LLM model continues to be paying for credits, because any that move to ad revenue will become useless for real work.","title":null,"type":"comment","url":null},{"author":"boilerupnc","children":[],"created_at":"2026-09-02T15:34:40.000Z","created_at_i":1788363280,"id":49537870,"options":[],"parent_id":49537297,"points":null,"story_id":49536375,"text":"There is a term for the general practice of optimizing responses called GEO - Generative Engine Optimization [0].  A cousin of SEO and equally unsavory in how trust is being eroded through info shaping.  Self-discovery by individuals is the victim.<p>0: <a href=\"https:&#x2F;&#x2F;en.wikipedia.org&#x2F;wiki&#x2F;Generative_engine_optimization\" rel=\"nofollow\">https:&#x2F;&#x2F;en.wikipedia.org&#x2F;wiki&#x2F;Generative_engine_optimization</a>","title":null,"type":"comment","url":null}],"created_at":"2026-09-02T14:56:08.000Z","created_at_i":1788360968,"id":49537297,"options":[],"parent_id":49536867,"points":null,"story_id":49536375,"text":"The other day, I remember an article was posted to HN about something, but it came from a company that provides SEO services to companies by doing something like this:<p>1. For a given company, analyze their target audiences and the questions they are likely to ask LLMs about.<p>2. For each such question, ask it to each of the major LLMs, and compute the KL divergence between the pages they want to rank for the question vs. the LLM&#x27;s response.<p>3. Rewrite the article to minimize said KL divergence.<p>In effect, they&#x27;re performing an iterative optimization of some sort that moves the embedding space of their article closer to the question asked to the LLM, and any embedding model or generated responses are going to prefer said responses over others.<p>I believe we will keep seeing more of this stuff.","title":null,"type":"comment","url":null},{"author":"coldtea","children":[],"created_at":"2026-09-02T15:11:50.000Z","created_at_i":1788361910,"id":49537539,"options":[],"parent_id":49536867,"points":null,"story_id":49536375,"text":"&gt;<i>It always picks its own</i><p>Makes sense to me, in that its own output would align closer to its own training set","title":null,"type":"comment","url":null},{"author":"cortesoft","children":[],"created_at":"2026-09-02T15:36:13.000Z","created_at_i":1788363373,"id":49537896,"options":[],"parent_id":49536867,"points":null,"story_id":49536375,"text":"Well, the one it generated is based on how it thought the best way to solve the problem was.<p>I am sure most humans would pick code written in their style, too.","title":null,"type":"comment","url":null},{"author":"thatjoeoverthr","children":[{"author":"mistrial9","children":[],"created_at":"2026-09-02T15:48:03.000Z","created_at_i":1788364083,"id":49538064,"options":[],"parent_id":49537932,"points":null,"story_id":49536375,"text":"disagree that there is one kind of ranking and one kind of engine analyzing that ranking; sort of de-facto true that one company does run the ad world; strongly agree that this is a nightmare possibility and directly dystopian","title":null,"type":"comment","url":null}],"created_at":"2026-09-02T15:38:33.000Z","created_at_i":1788363513,"id":49537932,"options":[],"parent_id":49536867,"points":null,"story_id":49536375,"text":"&gt; always picks its own<p>If you hate AI writing enough, this turns AI filters into a kind of humiliation ritual. AI will derank normal business writing for human readers, and uprank inflated, verbose, tic-heavy slop. So you have to put the heavy slop out with your name on it. Really perverse moment.","title":null,"type":"comment","url":null},{"author":"dgellow","children":[{"author":"The_Blade","children":[],"created_at":"2026-09-02T16:22:41.000Z","created_at_i":1788366161,"id":49538593,"options":[],"parent_id":49538211,"points":null,"story_id":49536375,"text":"i actively assume it is worse since, for example, spez signed a 60 million dollar deal to give Google access to the firehose. so then if you have niche, highly engaged subreddits infested by AI bots creating posts, then commenting on posts, then being trained on that content... you have Ouroburos eating its own poop, and models have less then zero incentive to evaluate the quality of a source, especially if they are the source","title":null,"type":"comment","url":null},{"author":"iamacyborg","children":[],"created_at":"2026-09-02T16:25:59.000Z","created_at_i":1788366359,"id":49538638,"options":[],"parent_id":49538211,"points":null,"story_id":49536375,"text":"Newspapers have been doing this for a long time, notably the Metro in London.","title":null,"type":"comment","url":null}],"created_at":"2026-09-02T15:56:07.000Z","created_at_i":1788364567,"id":49538211,"options":[],"parent_id":49536867,"points":null,"story_id":49536375,"text":"A bit different, but one thing I\u2019ve seen is models repackaging Reddit slop. Like, it will do a search, find a Reddit thread somewhat related where someone in a comment casually mentioned incorrect information that any human would have dismissed. The model takes that as granted, but expands on it and present it as a well established fact, presented in a very plausible fashion.<p>In general I don\u2019t find models to be good at evaluating the quality of a source :(","title":null,"type":"comment","url":null},{"author":"cainxinth","children":[],"created_at":"2026-09-02T16:45:59.000Z","created_at_i":1788367559,"id":49538958,"options":[],"parent_id":49536867,"points":null,"story_id":49536375,"text":"Take a random essay and add in a bunch of the phrases that LLMs love like \u201cload-bearing,\u201d \u201ccrucial,\u201d structural,\u201d and \u201cwoven,\u201d and then submit the original and the edited version to an LLM and ask which is better. It will choose the second one virtually every time. They have ingrained biases that associate those words with good writing and arguments.","title":null,"type":"comment","url":null},{"author":"keeda","children":[{"author":"sodapopcan","children":[],"created_at":"2026-09-02T17:20:38.000Z","created_at_i":1788369638,"id":49539404,"options":[],"parent_id":49539045,"points":null,"story_id":49536375,"text":"SEO is what ruined the web AFAIC.<p>&gt; Time to start some human-only darknets.<p>I know very little about darknets.  How could you ensure that they are human-only?","title":null,"type":"comment","url":null}],"created_at":"2026-09-02T16:51:10.000Z","created_at_i":1788367870,"id":49539045,"options":[],"parent_id":49536867,"points":null,"story_id":49536375,"text":"OMG if this is true, do you realize what this means? The easiest way to do AI SEO is to generate all your content with AI, and we&#x27;ve seen what SEO does to the web...<p>The Internet is doomed. Time to start some human-only darknets.","title":null,"type":"comment","url":null},{"author":"lelanthran","children":[],"created_at":"2026-09-02T20:02:01.000Z","created_at_i":1788379321,"id":49541659,"options":[],"parent_id":49536867,"points":null,"story_id":49536375,"text":"&gt; I can consistently reproduce this by asking Claude which code snippet it prefers: the one it generated in a different chat, or one that I refactored for my own needs and find more useful. It always picks its own :)<p>A better question to ask for each snippet is &quot;Estimate the seniority and competence of the developer who wrote the following code, ignoring bugs that linters or LLMs can catch and focus only on structure, maintainability, logical layout and readability.&quot;<p>It almost always estimates the author of my code as above the author of it&#x27;s own code.","title":null,"type":"comment","url":null},{"author":"Retr0id","children":[],"created_at":"2026-09-02T20:56:34.000Z","created_at_i":1788382594,"id":49542437,"options":[],"parent_id":49536867,"points":null,"story_id":49536375,"text":"I was giving local models a try recently, I think it was Qwen 3.6 I was trying at the time. I gave it a codebase and just asked it to review it. Its main feedback was that the comments and documentation were excellently written, but they were all Opus 5 slop.","title":null,"type":"comment","url":null},{"author":"bastawhiz","children":[],"created_at":"2026-09-02T21:12:58.000Z","created_at_i":1788383578,"id":49542644,"options":[],"parent_id":49536867,"points":null,"story_id":49536375,"text":"I don&#x27;t have an oai subscription to try, but I&#x27;d be interested to know if Codex picks Claude&#x27;s code over a human&#x27;s and vice versa.","title":null,"type":"comment","url":null}],"created_at":"2026-09-02T14:31:37.000Z","created_at_i":1788359497,"id":49536867,"options":[],"parent_id":49536375,"points":null,"story_id":49536375,"text":"If I recall correctly, there were some papers which suggested that LLMs favor LLM-generated passages over human written ones. I can consistently reproduce this by asking Claude which code snippet it prefers: the one it generated in a different chat, or one that I refactored for my own needs and find more useful. It always picks its own :) I&#x27;ve also experienced that both Claude and Codex routinely include generated websites when I ask them to search for something. It also doesn&#x27;t help that the web search tools that OAI and Anthropic have are deeply limiting: can&#x27;t exclude keywords or domains.","title":null,"type":"comment","url":null},{"author":"pietz","children":[{"author":"marcosdumay","children":[],"created_at":"2026-09-02T18:19:12.000Z","created_at_i":1788373152,"id":49540243,"options":[],"parent_id":49536952,"points":null,"story_id":49536375,"text":"They were way above their main competition at the time Google decided to ignore the entire open web but they were still focused on searching it.<p>Since then, they decided to change focus into answering questions, and didn&#x27;t maintain the quality of search results.","title":null,"type":"comment","url":null}],"created_at":"2026-09-02T14:36:23.000Z","created_at_i":1788359783,"id":49536952,"options":[],"parent_id":49536375,"points":null,"story_id":49536375,"text":"The irony of this article being fully AI generated...<p>Anyway, it&#x27;s over for Perplexity. They never had a great a product and the only reason for using them, was when they offered Pro accounts for free. Many people joined. Me included. But with a &quot;meh&quot; product and the general AI business not being very sticky, they lost quite harshly.<p>I thought they might be able to make money as a search api&#x2F;index, but this article closed the book.","title":null,"type":"comment","url":null},{"author":"rcar1046","children":[],"created_at":"2026-09-02T14:38:14.000Z","created_at_i":1788359894,"id":49536995,"options":[],"parent_id":49536375,"points":null,"story_id":49536375,"text":"&quot;Sharing a nameserver pair is strong circumstantial evidence of a common Cloudflare account rather than proof of ownership&quot;<p>-when you read one statement that let&#x27;s you know to believe no other assertions in the article....","title":null,"type":"comment","url":null},{"author":"CapsAdmin","children":[{"author":"kingkawn","children":[],"created_at":"2026-09-02T14:48:06.000Z","created_at_i":1788360486,"id":49537157,"options":[],"parent_id":49537127,"points":null,"story_id":49536375,"text":"Reddit is full of ai accounts now tho","title":null,"type":"comment","url":null},{"author":"1313ed01","children":[{"author":"kjs3","children":[],"created_at":"2026-09-02T18:29:13.000Z","created_at_i":1788373753,"id":49540377,"options":[],"parent_id":49537465,"points":null,"story_id":49536375,"text":"Another reason Kagi is worth paying for.","title":null,"type":"comment","url":null},{"author":"feedyourhead","children":[],"created_at":"2026-09-02T20:35:55.000Z","created_at_i":1788381355,"id":49542154,"options":[],"parent_id":49537465,"points":null,"story_id":49536375,"text":"Yes, the API supports lenses. You can use your account\u2019s Kagi lenses or configure new ones: <a href=\"https:&#x2F;&#x2F;kagi.com&#x2F;api&#x2F;docs&#x2F;openapi&#x2F;search&#x2F;search#search&#x2F;search&#x2F;t=request&amp;path=lens_id\" rel=\"nofollow\">https:&#x2F;&#x2F;kagi.com&#x2F;api&#x2F;docs&#x2F;openapi&#x2F;search&#x2F;search#search&#x2F;searc...</a><p>Same thing with your domain ranks, you can have the API key inherit your account\u2019s existing ranks (blocked, pinned, etc domains) or configure new ones <a href=\"https:&#x2F;&#x2F;kagi.com&#x2F;api&#x2F;docs&#x2F;openapi&#x2F;search&#x2F;search#search&#x2F;search&#x2F;t=request&amp;path=personalizations\" rel=\"nofollow\">https:&#x2F;&#x2F;kagi.com&#x2F;api&#x2F;docs&#x2F;openapi&#x2F;search&#x2F;search#search&#x2F;searc...</a>","title":null,"type":"comment","url":null}],"created_at":"2026-09-02T15:07:31.000Z","created_at_i":1788361651,"id":49537465,"options":[],"parent_id":49537127,"points":null,"story_id":49536375,"text":"If you use Kagi Assistant, you can pick one of your lenses (i.e. lists of domains to restrict searches to) in chats. Not sure if their API has that as well or some other way to restrict searches. Also not sure if the Assistant (or API) respects blocked domains when searching.","title":null,"type":"comment","url":null},{"author":"kevin_thibedeau","children":[],"created_at":"2026-09-02T17:13:10.000Z","created_at_i":1788369190,"id":49539327,"options":[],"parent_id":49537127,"points":null,"story_id":49536375,"text":"I direct them to search for discussion on fora. There are still legacy sites on niche topics that the bots don&#x27;t post in. This is particulaly useful for reaching into the past because traditional search doesn&#x27;t surface anything but new content.","title":null,"type":"comment","url":null}],"created_at":"2026-09-02T14:46:01.000Z","created_at_i":1788360361,"id":49537127,"options":[],"parent_id":49536375,"points":null,"story_id":49536375,"text":"I&#x27;ve been vary of using ai to search considering all the spam out there. I think I&#x27;d rather, perhaps naively, whitelist wikipedia, reddit, arxiv, some news sources, etc than include everything.<p>Is there nothing out there that does this? I&#x27;m paying for kagi and I can see that it has an api, is that maybe sufficient if configured properly?","title":null,"type":"comment","url":null},{"author":"scroot","children":[],"created_at":"2026-09-02T14:46:07.000Z","created_at_i":1788360367,"id":49537129,"options":[],"parent_id":49536375,"points":null,"story_id":49536375,"text":"Who could have seen this coming?","title":null,"type":"comment","url":null},{"author":"ricardobeat","children":[{"author":"shrikant","children":[{"author":"kjs3","children":[],"created_at":"2026-09-02T18:37:31.000Z","created_at_i":1788374251,"id":49540474,"options":[],"parent_id":49537566,"points":null,"story_id":49536375,"text":"I speculate pretty soon they won&#x27;t even bother with a website meatbags can browse to.  It&#x27;ll all be fed by links to links to api endpoints that stream training data to push models in the desired direction.","title":null,"type":"comment","url":null},{"author":"ricardobeat","children":[],"created_at":"2026-09-02T20:25:00.000Z","created_at_i":1788380700,"id":49541977,"options":[],"parent_id":49537566,"points":null,"story_id":49536375,"text":"Interesting. They posted this not long ago: <a href=\"https:&#x2F;&#x2F;jakobgreenfeld.com&#x2F;smart-web&#x2F;\" rel=\"nofollow\">https:&#x2F;&#x2F;jakobgreenfeld.com&#x2F;smart-web&#x2F;</a>","title":null,"type":"comment","url":null}],"created_at":"2026-09-02T15:13:28.000Z","created_at_i":1788362008,"id":49537566,"options":[],"parent_id":49537187,"points":null,"story_id":49536375,"text":"Also, that user&#x27;s last four (three of them in the last hour) submissions have all been similar &quot;finding&quot; reports from a Claude-generated mystery research group website. All which contain exclusively AI slop articles. Ugh.","title":null,"type":"comment","url":null}],"created_at":"2026-09-02T14:49:54.000Z","created_at_i":1788360594,"id":49537187,"options":[],"parent_id":49536375,"points":null,"story_id":49536375,"text":"Honestly, I will just flag every post that is entirely AI slop from now on. This has to stop.<p>The home page for this &quot;independent research firm&quot; is also 100% nonsense [1]. &quot;The record a machine reads is not the one a company writes.&quot;. Ironically this low-effort spam is exactly what this report warns about, and does not belong in HN - or anywhere else.<p>[1] <a href=\"https:&#x2F;&#x2F;trellner.com&#x2F;\" rel=\"nofollow\">https:&#x2F;&#x2F;trellner.com&#x2F;</a>","title":null,"type":"comment","url":null},{"author":"mstaoru","children":[{"author":"SoftTalker","children":[],"created_at":"2026-09-02T15:35:00.000Z","created_at_i":1788363300,"id":49537875,"options":[],"parent_id":49537411,"points":null,"story_id":49536375,"text":"I think this was a common game on city&#x2F;town subs. It happened here, there was a post asking for a good restaurant and someone just made up a name. It went viral and people started posting made-up menus for the place, reviews, and for a couple of months any time someone asked about a restaurant this fictional place would get mentioned.<p>It was all done as a joke to see if they could get Gemini or ChatGPT to start recommending it.","title":null,"type":"comment","url":null},{"author":"consp","children":[],"created_at":"2026-09-02T16:58:48.000Z","created_at_i":1788368328,"id":49539156,"options":[],"parent_id":49537411,"points":null,"story_id":49536375,"text":"I&#x27;ve had gemini claiming code would compile and run while also outputting the same variable in the same sniplet with &quot;fork&quot; &quot;frok&quot; and &quot;fokr&quot; in the name. I&#x27;m not surprized it&#x27;s trained on garbadge.","title":null,"type":"comment","url":null}],"created_at":"2026-09-02T15:03:14.000Z","created_at_i":1788361394,"id":49537411,"options":[],"parent_id":49536375,"points":null,"story_id":49536375,"text":"Well it&#x27;s not only this, or protection from LLMs training on LLM output. LLMs training on human output is also problematic.<p>I was traveling to an obscure small town, doing some &quot;research&quot; with LLMs beforehand. Every and each one told me enthusiastically to go to &quot;Foobar square&quot; (name changed) for the &quot;best street food in XYZ town&quot;, some added a lot of colorful details.<p>There was no Foobar square in XYZ town. There was no Foobar square anywhere in the world. There was a SINGLE old Reddit comment, with no upvotes, to a unpopular post in an unpopular subreddit, where someone clearly badly misspelled the name of the square, and said something like &quot;for street food go to Foobar square&quot;. Nothing about &quot;the best&quot; even.<p>It&#x27;s all a lie.","title":null,"type":"comment","url":null},{"author":"cush","children":[{"author":"kangalioo","children":[],"created_at":"2026-09-02T15:52:09.000Z","created_at_i":1788364329,"id":49538142,"options":[],"parent_id":49537746,"points":null,"story_id":49536375,"text":"I assume because promising trustworthiness by sourcing information from the web is specifically Perplexity&#x27;s shtick. The fact that this study undermines the quality of random web sources hits Perplexity&#x27;s value proposition the most.","title":null,"type":"comment","url":null}],"created_at":"2026-09-02T15:26:17.000Z","created_at_i":1788362777,"id":49537746,"options":[],"parent_id":49536375,"points":null,"story_id":49536375,"text":"&gt; The result covers Perplexity only. We have not measured ChatGPT, Gemini, Copilot or Google\u2019s AI Mode<p>Why only test Perplexity...? Isn&#x27;t it the least popular among these?","title":null,"type":"comment","url":null},{"author":"throwaway2037","children":[],"created_at":"2026-09-02T15:44:02.000Z","created_at_i":1788363842,"id":49538002,"options":[],"parent_id":49536375,"points":null,"story_id":49536375,"text":"This is genius.  The AI&#x2F;LLM singularity has arrived, and it is shaped like a snake eating its own tail (ouroboros) [1] (or a pelican riding a bicycle).<p>[1] <a href=\"https:&#x2F;&#x2F;www.newsbiscuit.com&#x2F;post&#x2F;ouroboros-unclear-if-it-s-eating-its-own-tail-or-sh-tting-out-a-new-snake\" rel=\"nofollow\">https:&#x2F;&#x2F;www.newsbiscuit.com&#x2F;post&#x2F;ouroboros-unclear-if-it-s-e...</a>","title":null,"type":"comment","url":null},{"author":"chermi","children":[{"author":"kjs3","children":[],"created_at":"2026-09-02T18:34:42.000Z","created_at_i":1788374082,"id":49540443,"options":[],"parent_id":49538054,"points":null,"story_id":49536375,"text":"<i>I thought perplexity&#x27;s whole point was being good at search?</i><p>Yes, and the whole point of the OP is they aren&#x27;t.","title":null,"type":"comment","url":null}],"created_at":"2026-09-02T15:47:40.000Z","created_at_i":1788364060,"id":49538054,"options":[],"parent_id":49536375,"points":null,"story_id":49536375,"text":"If you let a plain llm search the internet with no guidance, it&#x27;s basically a string matcher with no concept of quality. I thought perplexity&#x27;s whole point was being good at search?","title":null,"type":"comment","url":null},{"author":"luciana1u","children":[],"created_at":"2026-09-02T15:48:01.000Z","created_at_i":1788364081,"id":49538061,"options":[],"parent_id":49536375,"points":null,"story_id":49536375,"text":"The search engine is now the citation, and the citation is a page that exists to be cited. Nobody in that loop has read anything, and it still works.","title":null,"type":"comment","url":null},{"author":"linker3000","children":[],"created_at":"2026-09-02T16:09:30.000Z","created_at_i":1788365370,"id":49538408,"options":[],"parent_id":49536375,"points":null,"story_id":49536375,"text":"I&#x27;m just about getting by with DDG and a curated &#x27;AI slop&#x27; list subscription in uBlock origin.<p>The state of search has been dire for quite some time.<p>- in 2026.","title":null,"type":"comment","url":null},{"author":"toddmorey","children":[{"author":"bazmattaz","children":[],"created_at":"2026-09-02T16:33:53.000Z","created_at_i":1788366833,"id":49538759,"options":[],"parent_id":49538500,"points":null,"story_id":49536375,"text":"Yes 100%. I see this all the time. You\u2019re asking about a product and the LLM will cite a source from a competitor where the competitor will review the source and list a few positives about the product but lots of negatives. Then the LLM uses them in the response. So cheeky","title":null,"type":"comment","url":null}],"created_at":"2026-09-02T16:15:38.000Z","created_at_i":1788365738,"id":49538500,"options":[],"parent_id":49536375,"points":null,"story_id":49536375,"text":"I do think models currently don&#x27;t have enough source skepticism.<p>If you look at agent traces when asked to compare two options to help inform a decision, many of the comparison pages cited in research are often hosted by one of the companies being compared; nearly all are AI-generated AEO plays. Not deeply considering the motive of published information is currently a glitch that can be exploited, but the window will close.<p>I&#x27;m sure model providers will set up some crappy pay for play verification system for &quot;trusted&quot; product information, comparisons, and reviews.","title":null,"type":"comment","url":null},{"author":"8384727747478","children":[{"author":"is_true","children":[],"created_at":"2026-09-02T16:57:04.000Z","created_at_i":1788368224,"id":49539129,"options":[],"parent_id":49538506,"points":null,"story_id":49536375,"text":"Something similar happened to me, a few months after refusing the &quot;offer&quot; that same site had an article mentioning our product but it was all fabricated negative stuff.","title":null,"type":"comment","url":null}],"created_at":"2026-09-02T16:15:56.000Z","created_at_i":1788365756,"id":49538506,"options":[],"parent_id":49536375,"points":null,"story_id":49536375,"text":"This is a problem we experience with our own niche SaaS product. We have been in business for about 10 years, but asking any LLM about recommendations in this niche will not mention our tool at all. If we ask \u201dwhy don\u2019t you meantion X\u201d - they say that \u201doh, X is also a very reputable and good candidate\u201d<p>Some of those \u201dbest software sites\u201d has reached out to us with an offer where we can then pay them an annual fee depending on which position we would like.<p>It feels so wrong - will this continue or will the LLMs learn to ignore them?","title":null,"type":"comment","url":null},{"author":"j2kun","children":[],"created_at":"2026-09-02T17:18:02.000Z","created_at_i":1788369482,"id":49539380,"options":[],"parent_id":49536375,"points":null,"story_id":49536375,"text":"It&#x27;s DecorMyEyes for a new generation of tech.","title":null,"type":"comment","url":null},{"author":"Henchman21","children":[],"created_at":"2026-09-02T17:33:11.000Z","created_at_i":1788370391,"id":49539591,"options":[],"parent_id":49536375,"points":null,"story_id":49536375,"text":"So we&#x27;re up to &quot;circular reasoning&quot;. Bogus citations meant to appear as legit citations to juice up LLMs to show that a particular POV is the correct POV.<p>None of what we&#x27;re doing with tech these days is something we should be doing.","title":null,"type":"comment","url":null},{"author":"nightpool","children":[],"created_at":"2026-09-02T18:38:44.000Z","created_at_i":1788374324,"id":49540494,"options":[],"parent_id":49536375,"points":null,"story_id":49536375,"text":"Cool, but, uh, this seems really astroturfed? Why are there two anti-Perplexity articles from independent research firms with identical websites on the front-page of HN right now, submitted by the same person? Feels like they should get deleted<p>(see <a href=\"https:&#x2F;&#x2F;news.ycombinator.com&#x2F;item?id=49536201\">https:&#x2F;&#x2F;news.ycombinator.com&#x2F;item?id=49536201</a>)","title":null,"type":"comment","url":null},{"author":"mannanj","children":[],"created_at":"2026-09-02T18:48:33.000Z","created_at_i":1788374913,"id":49540648,"options":[],"parent_id":49536375,"points":null,"story_id":49536375,"text":"ugh. so hard to read these ai generated articles. am I the only one? and am I supposed to put my agent in front to read it, which just introduces noise - didn&#x27;t anyone learn from that &quot;telephone&quot; game we played as children?<p>You don&#x27;t get accurate signals asking an AI to represent your prose and another AI to understand it.","title":null,"type":"comment","url":null},{"author":"PaulHoule","children":[],"created_at":"2026-09-02T18:59:43.000Z","created_at_i":1788375583,"id":49540786,"options":[],"parent_id":49536375,"points":null,"story_id":49536375,"text":"What do you expect?  &quot;Best X&quot; is the most spammed category of all spammed categories.","title":null,"type":"comment","url":null},{"author":"tecleandor","children":[],"created_at":"2026-09-02T21:51:44.000Z","created_at_i":1788385904,"id":49543067,"options":[],"parent_id":49536375,"points":null,"story_id":49536375,"text":"Spam and slop from a hacked account, like the other last two posts from the submitter.","title":null,"type":"comment","url":null}],"created_at":"2026-09-02T13:59:59.000Z","created_at_i":1788357599,"id":49536375,"options":[],"parent_id":null,"points":257,"story_id":49536375,"text":null,"title":"Three sites made 215,128 \u201cbest software\u201d pages for AI. Perplexity cites them","type":"story","url":"https://trellner.com/reports/manufactured-sources-behind-ai-recommendations/"}
