TIME Is Serving AI Bots a Different Website, with Ads Built In
Posted by vincent_s 1 day ago
Comments
Comment by skeledrew 1 day ago
Comment by pohl 1 day ago
Comment by skeledrew 1 day ago
Yep, only a matter of time.
Comment by nilamo 1 day ago
Why do we choose to live in the worst timeline?
Comment by butlike 1 day ago
This could take up to 30 minutes for subagents to complete. So you go to the store, but the milk is already sold out. Walmart will ship you the milk if you pay the S&H. You drive back home and check the terminal. "Completed in 6m37s."
is the worst timeline. That's close though
Comment by scubbo 1 day ago
Comment by polotics 1 day ago
Comment by marcosdumay 1 day ago
Comment by mschuster91 1 day ago
Unfortunately you'd have to wire your entire kitchen in cameras, your scale needs to be smart and all of it needs to be sent / processed in real time to track consumption which means it will take a lot of compute power and a level of data mining that could be abused by anybody from enterprising break-and-entry crews to the police, and on top of that if it's done by a cloud provider it's probably gonna end up in a data lake for targeted advertising.
Comment by MattDaEskimo 15 hours ago
What, last I checked I had 15%!
> Apologies, I was searching for the closest <Large Franchise Restaurant> since you are typically hungry around this time
Comment by keiferski 1 day ago
It ended up recommending and linking me to an entirely different one, that I hadn't linked to at all.
Don't know if it was just sloppy LLM "thinking," or a genuine ad.
Comment by cyanydeez 1 day ago
Comment by ccgreg 1 day ago
Comment by jodacola 1 day ago
Comment by OptionOfT 1 day ago
Comment by DANmode 1 day ago
Comment by cyanydeez 1 day ago
Comment by rglover 1 day ago
Comment by HPsquared 1 day ago
Comment by zeratax 21 hours ago
Comment by skeledrew 19 hours ago
Comment by fhdkweig 1 day ago
https://www.theguardian.com/technology/2016/mar/24/microsoft...
https://www.cbsnews.com/news/microsoft-shuts-down-ai-chatbot...
Comment by skeledrew 1 day ago
Comment by esafak 1 day ago
Comment by reaperducer 1 day ago
Sounds like politics.
"Tell a lie enough times, and it becomes the truth."
Comment by ticulatedspline 1 day ago
Works on humans too. Garbage in garbage out.
Comment by notjes 1 day ago
Comment by embedding-shape 1 day ago
Maybe now we'll get something similar, just happens to be for LLMs, but for us who like less bloat, it can be a better viewing/reading alternative. Hope it spreads :)
Comment by amiga386 1 day ago
It's effectively dead today, and while it was dominant, Google abused it to get people to view and link to their cached copies of AMP pages, rather than the original site.
It also led to widespread abuse where the AMP version of the page differed in content significantly from the regular page. The same issue existed for WAP.
You may also remember browsing the web using Opera Mini, which wasn't a direct user-agent but used Opera's backend systems as a proxy that stripped, minified and compressed HTML/CSS, and resized and recompressed images. It let you browse the web using massively less mobile data, but raised a lot of privacy and security issues.
Comment by masklinn 1 day ago
That was the entire point of the tech. If Google had wanted they could have upranked simple and lightweight websites but they didn’t do that (at least not until whoever had used amp for their promo package left and the project was killed).
Comment by Groxx 1 day ago
AMP was just one of many smash-and-grab attempts to steal the internet, wrapped in a fake-gold-encrusted PR-infused box. Google has quite a lot of them.
Comment by customguy 1 day ago
I remember how amazing that mode was on desktop, I'd sometimes enabled it on site and shrink it horizontally just to look at it handle anything perfectly. Don't remember if that was independent of the proxy though, I just cared about the removed padding and the responsiveness. Make all sites look like that and add some gutter to the side I can toggle and resize, and I'll use your browser kthx.
Comment by jraph 1 day ago
Comment by masklinn 1 day ago
Comment by Telemakhos 1 day ago
Comment by bjackman 1 day ago
Comment by radley 1 day ago
Comment by dieselgate 1 day ago
Comment by samtheDamned 1 day ago
Comment by bjackman 22 hours ago
JS annoyances it won't fix are like idiotic scrolling behaviour and stuff. And for that, yeah reader mode is ideal.
Comment by ToucanLoucan 1 day ago
It's frankly wild how many of my favorite tools for Internet browsing have nothing to do with connectivity, solving bugs, or any of that and it's just stripping out all the fucking BULLSHIT that comes on a modern website.
Comment by inigyou 1 day ago
Comment by ihuman 1 day ago
Comment by mananaysiempre 1 day ago
Comment by madebysnacks 1 day ago
Comment by WJW 1 day ago
Comment by jerf 1 day ago
This sort of reminds me of that. LLMs are by their nature credulous. They can be trained to not give in easily to some things, like the capital of the US, but in general they constitutionally have a tendency to believe what they read. What they read is basically their universe. There's only so much room and so much training data to really strongly pin raw facts in their weights. The only way they can not believe some marginal fact presented to them in their input is to possibly have read something that contradicts it in the same session... and the vast, vast majority of the world is those marginal facts, not really objective things like capital names.
So if you can work a confident statement in to an LLM's input about some semi-relevant topic, it's truth to the LLM. And, being truth, the LLM will then happily and confidently elaborate on it quite a bit.
Of course, if it's irrelevant to the current query, it probably won't have much effect. Ads have always been a game of numbers, anyhow. Even a query about a science topic has some probability of eventually turning to a question about banking in the same session. It's probably a good idea to rather strictly partition your conversations to stick to a single topic, not to defend against this but just to maximize the effectiveness of what is in the context window by keeping it focused, but I have to imagine there's plenty of people out there who reuse conversations all the time and end up with single conversations covering a huge array of topics.
The good news, and the bad news, all at once, is that Google isn't going to take this one sitting down. If they're going to replace the search engine box with an LLM, well, they're using the same LLMs we're all using, if not in fact a bit cheaper one for the work they do, and by golly, that bot should be serving up Google's ads, not Time's ads! Who do these uppity content creators think they are, anyhow?! So there is definitely going to be work done in the field of ad-blocking content served to LLMs.
Comment by philistine 1 day ago
You assume that UI is sacrosanct and the same everywhere. Those LLM providers are already offering to mingle all your conversations together. Gemini from Google for example defaults to memory from every conversation, and it's safe to assume the option to disable it will be eventually removed.
Comment by jerf 1 day ago
But to your implied point about getting an advertisement into a memory file... that makes it even more amusing to hack Google's own AI to put Time's ads into it. I think that's probably an easier problem for Google to solve, too, though. The small size and the way that a memory is going be a stereotypical summary makes it easier to filter out the ads Google doesn't want...
... but it'll cost them. That's an AI-complete problem and they're going to have to run LLMs over the memories to filter them, at their expense.
The most obvious fix to me is to have an LLM try to pre-filter out the ads from Time's content before feeding that as pristine content to the "core" AI so it won't be corrupted by the advertisement, but the LLM doing the filtering has to be at least as smart as the one using the content and/or the one inserting the ads. (A dumber one can filter the obvious stuff, but then the obvious next step in the arms race is for Time to tell their AI to be more clever about it, and a smarter AI will dominate the dumb cheap AIs here.) This is going to be an expensive setup. And there will be semantic loss in any such filter, too.
Comment by tclancy 1 day ago
Comment by StableAlkyne 1 day ago
There was a brief moment in the early Internet before it was all hyper-optimized... Until the parasites in the advertising industry started attaching themselves to every page.
The year before ChatGPT, the first page of Google was SEO-optimized blogspam designed to say as little in as many words as possible, to splice ads between every paragraph. This was the net result of 20 years of SEO. I suspect this is why Google's AI search didn't get as much pushback as other tools, since its summarization of pages functions is a form of adblock.
Given the ecological impact of AI, I wonder how much damage the ad industry will be causing in 10 years once they figure out how to trick LLMs into manipulating their own users.
Comment by KeplerBoy 1 day ago
Comment by InsideOutSanta 1 day ago
Probably something like this?
Comment by gmerc 1 day ago
Comment by yccs27 1 day ago
Comment by gypsy_boots 1 day ago
Comment by MetaWhirledPeas 1 day ago
Comment by ASalazarMX 1 day ago
Comment by everdrive 1 day ago
Comment by xmcp123 1 day ago
Comment by zb3 1 day ago
Comment by pitchlatte 1 day ago
Comment by everdrive 1 day ago
- reject all 3rd party cookies
- selectively refuse to load 3rd party domains
- block javascript
- block webgl
- block webrtc
- block canvas
- use a generic user agent because "resist fingerprinting" is checked
- etc.
To a lot of sites, this makes you look like a bot. Most people don't go around deliberately spoofing their user agent these days. (and of course bots themselves can present whatever user agent they want.)
Comment by guywithahat 1 day ago
Comment by red_admiral 1 day ago
Comment by keane 1 day ago
Comment by inigyou 22 hours ago
> Your client does not have permission to get URL /search/docs/essentials/spam-policies from this server. That’s all we know.
What a utopia.
Comment by ButlerianJihad 21 hours ago
I do not believe that this is an accident. I suggest that you consider this is a "you problem", inigyou, and not the webmaster's fault.
https://news.ycombinator.com/item?id=49194606
https://news.ycombinator.com/item?id=49165757
https://news.ycombinator.com/item?id=49145010
https://news.ycombinator.com/item?id=49144842
https://news.ycombinator.com/item?id=49136992
Comment by inigyou 21 hours ago
Comment by raggi 1 day ago
Comment by apocalyptic0n3 1 day ago
Be careful not to do this with Googlebot, though. Google would consider this "cloaking" and could ban your entire domain for it.
Comment by throwawayffffas 1 day ago
I find that odd, If I were doing that, that's what I would want to target the bots used for training in order to get the ad content in the training corpus.
Comment by nairboon 1 day ago
Comment by int0x29 1 day ago
Comment by Magicrafter13 1 day ago
I'm also completely unbothered by the precedent of drip feeding product ads to someone's LLM chat history, and having that influence future conversations, because if you're going to outsource your buying decisions to an LLM, frankly I don't really care if you buy stupid products at that point - you brought that on yourself.
Comment by dust42 1 day ago
Comment by inigyou 22 hours ago
Comment by netsharc 1 day ago
Comment by furst-blumier 1 day ago
Comment by ASalazarMX 1 day ago
Comment by inigyou 1 day ago
Comment by Havoc 1 day ago
LLMs being indirectly tainted that way seems like a serious problem
Comment by bigfishrunning 1 day ago
Comment by dspillett 1 day ago
Could it be that other scrapers are pulling the pages, without making the extra requests to get ad related resources, to present the content to people ad-free, and embedding the ads in the main response body is a way to get around that so the human sees an advert at least, even if it isn't the one they might see if the stalky-adtech-algo could deliver something more targetted.
Comment by Joel_Mckay 1 day ago
For media/news businesses that make their money from readership, it is unsustainable to subsidize Alphabets content farm. Thus, understandable people would change their corporate posture with a search turned scraper company. =3
Comment by inigyou 1 day ago
Comment by lostmsu 1 day ago
Comment by Joel_Mckay 1 day ago
Let us not forget they are a business =3
Comment by ccgreg 1 day ago
Comment by AkbarHabeebB 1 day ago
If there is any article mentioning something around this, can someone share it please? Im curious to know about it
Comment by BehanPrW 1 day ago
Comment by spiderfarmer 1 day ago
Comment by DANmode 1 day ago
Also: what’s the name of the ag project?!
Comment by gostsamo 1 day ago
Comment by Apreche 1 day ago
Comment by altmanaltman 1 day ago
Maybe they are banking on LLMs including it in their training data while scrapping or something like that. It does track impressions so they are definitely up to something.
Comment by nairboon 1 day ago
Comment by tclancy 1 day ago
Comment by nojs 1 day ago
Comment by shevy-java 1 day ago
Most people will say "yay, it is great you waste the time of AI bots via ads!". Well, I already think ads should not exist in the first place, nor bots, but both exist - but the real issue is when websites now steal my time. That was different in the 1990s. I think mankind made several missteps here.
Comment by bombela 1 day ago
Combined with how slow modern website are, I feel a sense of dread opening any website. Asking an LLM feels lower friction... but at what cost...
Comment by ForHackernews 1 day ago
Comment by forinti 1 day ago
Comment by zb3 1 day ago
Comment by bigfishrunning 1 day ago
Comment by krapp 1 day ago
I want every LLM that touches my site without my permission (which is any of them) to become obsessed with the perfect recipe for key lime pie. Or something.
It's just nature. LLMs are an unwanted invasive species on the internet, so the ecosystem must adapt.
Comment by jademcd 1 day ago
Comment by EtienneDeLyon 1 day ago