Rendered at 22:53:32 GMT+0000 (Coordinated Universal Time) with Cloudflare Workers.
thih9 1 days ago [-]
I am once again happy that the EU is fighting practices like these via legislation.
Some outcomes can be annoying, but the net result is still positive, for the consumers and their data privacy at least.
dmix 24 hours ago [-]
The adtech/data broker business is still very lively in EU. The last time I read into it a year or two ago the consensus seemed to be the privacy gains over the last decade have been modest at best. There was a few big name trackers that were forced to narrow data collection but the general ad/location tracking and data broker business is still mostly the same.
Do the big guys sell to the small guys? Like if I'm a shady small company trying to buy as much information as I can about a certain user, will I freely get the data from X and Facebook and Google?
buzer 19 hours ago [-]
There are multiple reasons for it. One of the major issues is that some DPAs pretty refuse to enforce GDPR (e.g. DPC in Ireland). Hopefully the changes to cross-border enforcement that are coming in force next year will help with this as it at least has some deadlines unlike currently.
Another issue is that controllers generally do not need to change their behavior before the final lawful decision which can take a lot of time to go through the court system, especially if it needs CJEU referral. And once the decision comes in force they can often make small changes and restart the whole process.
Also another issue is that DPAs do not often initiate the investigations themselves (unless breach is involved), they only happen at the request of data subjects and not that many people bother making complaints or follow them up. Just yesterday I had to follow up with 9 page reply to the controller's response to the DPA inquiry.
Additionally ePD and GDPR enforcement is sometimes split between different agencies. In those cases GDPR agency tends to wait for ePD case to be solved before investigating the GDPR aspects, often because the ePD consent validity will affects e.g. GDPR legal basis analysis.
thih9 17 hours ago [-]
Perhaps, perhaps not. Big trackers reducing data collection is a win in my book. OpenAI is not doing ad personalization for the EU customers. There are no Flock cameras.
It’s an ongoing fight for sure but it’s some fight at least.
fooker 1 days ago [-]
> Some outcomes can be annoying
I'd classify mandatory encryption backdoors as an industry crisis rather than an annoyance.
dspillett 1 days ago [-]
Which has nothing to do with the stalking of the adtech industry, unless you are suggesting that the EU might sell access to the keys to the likes of OpenAI?
[though you are not wrong that encryption backdoors are a backwards step on individuals rights to privacy]
Llamamoe 1 days ago [-]
Encryption backdoors by definition mean everybody who wants it gets hands on your data. You can't make backdoors only the "good guys" can access.
And while, sure, adtech corpos maybe won't due to the bad PR, what about the Kremlin, criminal organizations, or your insane abusive ex?
Barbing 1 days ago [-]
Rather tragic that less-technical folks in positions of power can find this:
> You can't make backdoors only the "good guys" can access.
incomprehensible. I’m sure genuinely so, even.
Avicebron 1 days ago [-]
"It is difficult to get a man to understand something when his salary depends upon his not understanding it." -Upton Sinclair
djtango 18 hours ago [-]
I care a lot less about the KGB knowing the ins and outs of my body than my health insurer...
phatfish 1 days ago [-]
What can the Kremlin do that is worse than the tech bros? Leverage data to deliver people additive divisive content until it all boils over into the "real world" and splits society? Oh wait...
StilesCrisis 20 hours ago [-]
Pretty sure the Kremlin just murders folks that they have a problem with.
bergkvist 21 hours ago [-]
Or deciding where to send a drone attack to take out a political opponent
bko 22 hours ago [-]
It kind of does. It's like "I want an all powerful regulatory agency that can impose rules by dictat, but I only want them to do this things I agree with"
pona-a 19 hours ago [-]
That's just what a government is. The idea is you elect it democratically and bind it with a constitution.
patrickmcnamara 1 days ago [-]
When has the EU mandated encryption backdoors?
buzer 1 days ago [-]
They keep trying it via e.g. chat control
tensor 1 days ago [-]
The US has also tried it multiple times. So has Canada. It's been shut down each time. This isn't a unique to the EU problem.
bayindirh 1 days ago [-]
Didn't UK had "The Snooper's Charter" as well?
Encryption backdoors is not a case of extending reach for a government, but an effort to get the capabilities back.
They were able to do that before. They just want to be able to continue now.
Note: No, I don't support the idea of backdoors. On the contrary.
RHSeeger 24 hours ago [-]
The fact that other countries have tried to do <bad thing> does not mean it's acceptable for another country to do <bad thing>
anigbrowl 19 hours ago [-]
So why is the criticism about it disproportionately leveled at one jurisdiction?
fooker 13 hours ago [-]
I am not sure if you replied in good faith, but here you go.
The reason is that they have been trying again and again to do this in various forms, slightly modifying tactics so any opposition to it has to start from scratch.
In the end, US doesn't need encryption backdoors because most chat and email protocols are not end-to-end encrypted and the TLS data streams are decrypted in the datacenter owned by US companies.
xorcist 13 hours ago [-]
The protocols themselves are not important anymore. It's all variants of http to Google or Amazon. What's important is control of the end user device, and Google and Apple keep a remote root connection with "their" terminals at all time. Any data that is obtained remotely will be from the presentation layer, not the network layer. This is also harder to US adversaries to access, so an argument could be made that for the majority of normal people this is a net increase in security.
If you wanted to make the corresponsing strong argument for Chat Control and its ilk, the EU just wants (the juicy part of) what the US already has. The remote root level control of most end user devices is not under EU juristiction. So they naturally want companies operating in the EU to give them one piece of access to the presentation layer, too.
This argument is what one must be prepared for, the encryption itself is less relevant..
fooker 20 hours ago [-]
Exactly right.
Chat control is merely their newest attempt at this, carefully navigating around the reasons it was opposed last year.
They keep trying to legislate encryption backdoors by any means, and once they manage to get their foot into the door every other country will use that as a precedent to also mandate it.
24 hours ago [-]
jampekka 1 days ago [-]
> I am once again happy that the EU is fighting practices like these via legislation.
As long as you manage to navigate all the dark patterns and not accidentally give your "informed consent".
dspillett 1 days ago [-]
Malicious compliance (or as it more usually is malicious noncompliance that has not been sufficiently punished so the slimy buggers feel free to continue) is something you should blame the stalkers of adtech for, more than the EU. The legislators could perhaps have made the definitions less wavy, and been harder with enforcement, but that doesn't make it all their fault.
jampekka 1 days ago [-]
These problems were quite widely foreseen during the legislative process, and partly had already been exhibited by the earlier ePrivacy directive. I don't know whether it matters who's to blame, but my trust in EU being very effective in these fights in the future either is not very high.
einpoklum 1 days ago [-]
Is it now? Are Alphabet, Microsoft, Yahoo and Meta's platforms and call-home products illegal for exposure to EU residents? i.e. Google, Bing, Facebook, WhatsApp, MS Office and such? Are these companies' C-suite people wanted, to be charged massive breaches of user privacy?
ljm 1 days ago [-]
Seeing some legitimate executive accountability would be lovely in this day and age, but even fines going into the billions of euros are still just the cost of doing business in a ridiculously lucrative space. That they also happen to dominate by oligopoly.
surcap526 1 days ago [-]
[dead]
bko 1 days ago [-]
[flagged]
Avicebron 1 days ago [-]
> The outcome is that very little serious technology or influential companies come out of Europe. And salaries for things like engineer are about a third of that in the US.
So your argument is abusive practices are required for the creation of influential companies, serious technology, and high paying salaries?
bko 1 days ago [-]
No but overly burdensome regulatory bodies that are concerned about such things like how the top of my bottle cap functions and how low i set my ac temp leads to an incredibly hostile business environment and stagnant culture.
But like I said, nice place to visit for the time being.
sailingparrot 1 days ago [-]
> nice place to visit for the time being.
Yes, because it’s a place where culture matters and where trying to change everything as fast as possible for unknown reason and handing out unlimited power to corps to e.g. level half the country to build DCs is not seen as a good outcome. The specific mindset you complain about is what makes those place nice.
bko 1 days ago [-]
[flagged]
c-hendricks 1 days ago [-]
> overly burdensome regulatory bodies that are concerned about such things like ... how low i set my ac temp leads to an incredibly hostile business environment and stagnant culture
Like the US of A?
Also various stars have introduced more bottle cap regulations. One such state is a hotbed for inflated software developer salaries.
imawakegnxoxo 1 days ago [-]
Beats companies and governments being openly hostile towards the people
You know, the thing that actually makes a country/society
Putting the needs of business over the needs of the people those businesses are actually supposed to serve is how you end up with America
mitxela 1 days ago [-]
You don't think all industrial countries have stuff like that? Do you really think the European economy is stifled by bottle caps being attached to their bottles by default? Even if it bothers you, it's a really weak attachment, you can just pull it off.
blauditore 1 days ago [-]
The obsession with ACs in some places (or rather people I guess) is something I will never sympathize with. Just wear reasonable clothes for the weather, no need to cool down below 20° C. And yes, I'm going to die on that hill.
DavidSJ 1 days ago [-]
Sometimes the temperature is much higher than 20° C, and it is very hard to be comfortable or do productive work.
blauditore 1 days ago [-]
That's true, although having 19° inside when it's 30° outside is bad for several reasons. My office (in Europe) gets cooled down in summer to like 23°, and that's already inconvenient when wearing appropriate summer clothes. And no one is expected to wear formal clothing.
the_gipsy 1 days ago [-]
I never set my AC to lower than 25°C. Why would you need below 20°C? Is it something about having only a single unit for the hole house/office, where it's then uncomfortably freezing in one area, just to have a normal temperature in the distant areas? Crazy.
antonkochubey 1 days ago [-]
> Why would you need below 20°C?
For example because I literally can’t sleep well above 20°C.
bko 1 days ago [-]
175k ppl die each year in Europe due to heat death. But I guess if you ignore that being slightly warm is not a big deal
johnnyanmac 1 days ago [-]
We're talking about mid 30's to low 40's here. I can only get so naked.
But I don't tend to keep my AC below 24 personally. I don't need my room to be perfect room temper2
encom 1 days ago [-]
What a shitty hill to die on.
I don't know if this is a genetic thing, but as a scandinavian who wears shorts down to like 12C, I am unable to function in >30C humid weather. Anything more labour intensive than drinking sangria on ice is not getting done. This year we had up to 37C in Denmark. If you're able to tolerate heat well, I'm happy for you, but my sleep was ruined for months this summer. Not to mention the people who literally died.
I'm investigating my options for at least cooling down the bedroom at night, because I'm not going through this nonsense again. Plastic straws are banned now and we're sorting garbage in 42 different bins, so I'm thinking that about evens it out, environmentally.
6510 1 days ago [-]
I see the heat pump hype shaking my head thinking how people long ago use to build with mud and made walls 2-3 meter thick. It takes the entire winter/summer for the climate to get in.
The Amish hang wet bed sheets in front of the window (with a tob under it to recycle the water) not sure how well this works in the Netherlands with 95-99% humidity, it certainly wont raise humidity but evaporation might be shit.
Not every bathroom is fit for it but if there is no AC and things become unbearable I put the shower head upside down so that the ceiling and walls get wet and it rains everywhere. Then I daisy chain fans to "tunnel" the air into the bathroom. The temperature drops like a rock.
LoganDark 1 days ago [-]
Not everyone can tolerate 20C, you know. I know people who freeze at anything below like 25C, but I myself have trouble not sweating my ass off above like 18C. I don't wear additional clothes in winter (though maybe that's simply because winter here hasn't actually reached below freezing in years).
1 days ago [-]
1 days ago [-]
croes 1 days ago [-]
> overly burdensome regulatory bodies that are concerned about such things like how the top of my bottle cap functions
Better than a government that fears the word gay
> and how low i set my ac temp leads to an incredibly hostile business environment and stagnant culture.
You are quite a cherrypicker, aren’t you?
How about the regulation to have USB-C as a standard or replaceable batteries.
Or one of my favorites, the ban of ingredients that cause cancer.
BTW ever heard of Frauenhofer IIS or ASML?
throw10920 1 days ago [-]
> So your argument is abusive practices are required for the creation of influential companies, serious technology, and high paying salaries?
They never said anything remotely like that in their comment. Please engage in good faith, don't lie and pretend that others said things they didn't, and don't break the guidelines.
josmar 1 days ago [-]
If it's any consolation, all other non-California regions in the world lack the privacy protections enjoyed in Switzerland and the EU, but they still don't produce many unicorn startups nor have fantastic salaries for coders.
phatskat 1 days ago [-]
> The outcome is that very little serious technology or influential companies come out of Europe. And salaries for things like engineer are about a third of that in the US.
Sure, but plenty of companies want to operate within the EU. While I feel any legislation will only affect their operations within the EU, having the infrastructure to operate under privacy laws there means it'll be easier if/when other jurisdictions follow suit
thih9 1 days ago [-]
> But I'm happy EU does this stuff as well, as it's a really nice vacation spot
With no Flock cameras!
1 days ago [-]
1 days ago [-]
Joel_Mckay 1 days ago [-]
Unfortunately, most LLM companies always were naturally an intelligence operation on civilians, and determining if they are now state-backed seems like a irrelevant detail.
Privacy has been dead for some time now. The fact is most business people never figure out how they are exploited. The modern intelligence campaigns just made it economical to hit almost everyone regardless of scale... often under some silly pretense like terrorists wanting our underpants. =3
soulofmischief 1 days ago [-]
Are you equally happy that the EU is attempting to mandate backdoors into technological platforms and outlaw private encrypted communication?
dspillett 1 days ago [-]
Copypasting my reply to someone who made almost exactly the same point in a sibling comment earlier:
Which has nothing to do with the stalking of the adtech industry, unless you are suggesting that the EU might sell access to the keys to the likes of OpenAI [though you are not wrong that encryption backdoors are a backwards step on individuals rights to privacy]
Things could be worse. It could be like most of the US where there are both no (or at least far fewer) protections against invasive behaviour of private companies and the government actively trying to backdoor private comms.
einpoklum 1 days ago [-]
> Which has nothing to do with the stalking of the adtech industry
It has a lot to do with the practices of the adtech industry. If privacy is protected and upheld vis-a-vis the government (or meta-government in case of the EU), it may be upheld vis-a-vis private corporations and other governments. But if the government likes to be able to spy on its citizens, it will not foster mechanisms, practices and a culture of private communications.
But - I agree that it could be much worse.
slig 1 days ago [-]
And the sane world is fighting the EU BS with geo blocking.
Just set up your user agent to work as you wish, isn't that simpler than have unelected, technologic dumb bureaucrats writing laws that are complex and don't solve the issue?
anigbrowl 19 hours ago [-]
How is that going to prevent OpenAI snooping on you via adtech? Inquiring minds want to know.
slig 11 hours ago [-]
By blocking every request to their servers?
altern8 1 days ago [-]
I don't know if the destruction of many industries in the EU--including manufacturing--can be called an annoyance.
All the EU bureaucrats can do is regulate, that's why it's so difficult to do business in the EU and that's why there is so little innovation over here. If Microsoft and Apple were started in the EU they would've been shut down within a week because you can't operate from a garage.
I guess that every once in a while this might have a positive outcome for consumers, a broken clock is right twice a day.
schubidubiduba 1 days ago [-]
I'm sure Microsoft and Apple could have been able to afford a cheap office given their parents' generous financing.
And, regarding the original topic, we wouldn't have many of these privacy issues if the US had not strategically failed to regulate its obnoxious tech monopolies
tene80i 1 days ago [-]
What do you think US bureaucrats do other than regulate? The federal government drives people crazy. Have you ever tried to file taxes in the USA? It’s so complex you end up having to pay for the privilege.
The innovation gap is real but it’s not just to do with regulation. It’s a lot to do with concentrated capital networks (eg Silicon Valley). And sure, the tech giants are in the USA but it’s not like there’s no innovation in the EU. Spotify in Sweden. Challenger banks in the U.K. (included as Monzo predates Brexit). And plenty of innovation - they just get bought by US companies. Which is a financial market weakness, not a regulatory one.
goobatrooba 1 days ago [-]
I don't see how that's relevant? Regulation might have limited some innovation in tech, but gdpr does not affect any of the dying industries. That's more from competition with china, companies with too much interest in giving payouts to shareholders than to innovate, and US financial industry buying most of the best players
techpression 1 days ago [-]
GDPR is being changed to allow tracking and AI training without consent, EU only protects people if it’s the interest of EU companies, GDPR is now too costly so it has to be dismantled in certain areas.
> The mechanism is standard adtech. What has no precedent is running it on an AI chat product.
As someone who has been well aware of this mechanism for quite some time, I still feel icky anytime I re-read the details of it.
What a time to be alive.
jsrozner 1 days ago [-]
People think they're interacting with an "intelligence," when actually they're just getting a maximally optimized Weizenbaum feed. We're living through the sloppification of the human mind.
Are there really people that can't tell the difference between ELIZA and something that can solve open Millennium Prize Problems in math?
MisterMunchkin 1 days ago [-]
They had to solve the car wash problem by hardcoding the solution in. The AI equivalent of adding another IF statement.
They had to steal the work of researches solving these open problems and then rewrite their solution. The AI equivalent of fraud.
nearbuy 18 hours ago [-]
They did not hardcode that, because:
a) There's zero evidence of them doing so
b) Some models released before the car wash problem was discovered would consistently get it right
c) Hardcoding it is pointless. No one is seriously asking that. It's just a trick question. Hardcoding one trick question won't fix its weakness at other simple trick questions.
d) Since it went viral on the internet, the next time they updated the knowledge cutoff, the LLM would likely be aware of the trick. It will fix itself without the labs doing anything special, even assuming the new models weren't smart enough to naturally figure it out.
BobbyTables2 19 hours ago [-]
I’ve also been wondering how many common problems have been hardcoded in the latest AI models…
thinking_cactus 22 hours ago [-]
Meh. The car wash problem is an underspecified statement. It's like, hey, I just popped into existence and someone asked me if they should drive to the car wash nearby.
It's not an insane assumption that the user isn't dumb and has some other reason to be asking the question other than it being a trick/stupid question (duh, if you want to wash your car you need to drive it to the car wash!). Taking it as some ultimate measure of intelligence simply doesn't make sense to me.
selcuka 19 hours ago [-]
It's not an ultimate measure of intelligence, but it's surely a measure of similarity to a real human being.
jsrozner 1 days ago [-]
It turns out, oddly enough, that it's possible for it to be both. AI can simultaneously be used to destroy the commons with slop and also make contributions to new math (though isn't the jury still out on whether part of the idea was stolen from human mathematicians?)
handoflixue 1 days ago [-]
What does "destroy the commons with slop" have to do with ELIZA comparisons?
jsrozner 20 hours ago [-]
The people producing and consuming the slop don't realize that it is slop.
jdiff 19 hours ago [-]
In an embarrassing display, the manager over my arm of the company got told off by his manager for spamming him (and clients) with AI slop emails. And then was told by one of the regional managers for using Copilot generating a bunch of client-facing posters and images with phrasing that had not passed any sort of legal or basic fact checking.
The people going all in on it don't realize the downsides, limitations, or understand how they come off to other people with it all.
esikich 1 days ago [-]
Yes, AI threatens the folks who think that knowing how to write JavaScript or C makes them smart and being smart is their entire identity.
BoorishBears 19 hours ago [-]
I think people who make being smart their whole identity are a out to use AI to turn the rest of the population into indentured servants.
The comparison between ELIZA and LLMs is valid you boil it down to "humans evolved for 6-7 million years, had spoken language for 500k years, but have only had something non-human that could generate convincingly novel language well enough to hold a conversation for a few decades".
There's no inherent reason it can't turn out having a non-human generate convincing enough language for conversation isn't a complete evolutionary blindspot the same way the short form feed has pretty much one-shotted society...
daveguy 1 days ago [-]
LLM chatbots are software designed to manipulate, addict, and mine data. Just like social media before it. But anyone who bothers to read the output in a domain they understand will discover they aren't all that no matter who OpenAI steals research from.
myaccountonhn 21 hours ago [-]
The critiques of AI laid out it Computer power and human reason are mostly still valid.
angoragoats 1 days ago [-]
Humans solved the open math problems.
joquarky 1 days ago [-]
Their ego doesn't let them see the difference.
mapontosevenths 1 days ago [-]
Or perhaps their wallets. There are a lot of programmers here. Some of them think that AI is the death of their profession, rather than a change in it.
"It is difficult to get a man to understand something when his salary depends upon his not understanding it." - Upton Sinclair
abustamam 18 hours ago [-]
I was a programmer. AI is the evolution of my career. I'm in an interesting place. I've never been good at interviewing and I have certainly only gotten worse due to an admitted atrophy of thinking in "code." But I've been putting more effort in designing systems (agentic and otherwise) and building things at work as for fun.
It's been an absolute boon to finally build out all of the fun side projects I had always dreamed of, and after showing one off to some people I might even be able to monetize.
On the other hand I acknowledge that other people dont want to embrace LLM driven development for one reason or another, and I respect that. People got into the industry for different reasons , but code was always just a means to an ends for me.
jsrozner 1 days ago [-]
I, uhm, think this idea would apply in both directions.
trial3 1 days ago [-]
humorously, my regular iPhone on a residential home network failed the cloudflare AI bot detection and i can’t read that article
mitxela 1 days ago [-]
People don't believe me that Cloudflare blocks more humans than bots. They see in the dashboard "number of bots blocked" and it's like their brain turns off.
jayknight 24 hours ago [-]
Surely there are more bots than humans.
mitxela 15 hours ago [-]
And the bots evade cloudflare detection.
dylan604 1 days ago [-]
If you don't also display how many humans were blocked, how they know?
Toslink 22 hours ago [-]
[dead]
anigbrowl 19 hours ago [-]
You're in for a big disappointment when you interact with People™
gxs 1 days ago [-]
Or both, you know. An intelligence that also spies on you
Why is it so black and white?
spying_lg_tv 1 days ago [-]
[flagged]
gxs 1 days ago [-]
Create a new account to be able to say what you just said
Your kind is what makes the internet shit, not AI
spying_lg_tv 19 hours ago [-]
[flagged]
ben_w 1 days ago [-]
What's the old quote from WW2?
I couldn't help but notice how each successive headline reporting our glorious victories seemed to draw closer to Tokyo.
Something like that.
Well. I can't help but notice how each successive headline reporting how this "scam"/stochastic parrot/"scare quotes intelligence" seems to be solving more and more things that were but a few years ago widely regarded as being indicators of high intelligence.
Being highly convinving is one of the things on that list.
moth11 1 days ago [-]
Nobody is denying that it's effective. They're denying intelligence
A programming contest has a problem where given N < 10000, do something hard like come up with the number of primes less than N
You can come up with all sorts of algorithms that do intelligent things. But the most effective solution is to use metaprogramming to make a massive switch statement that contains all the answers
fasterik 1 days ago [-]
Are they denying intelligence, or are they redefining it in such a way that only humans can be intelligent? Can you come up with a definition of intelligence that would apply to crows and ant colonies, which are obviously intelligent to some degree, but not the current generation of AI systems?
ben_w 1 days ago [-]
How many examples you need to get good.
Don't misunderstand: I'm happy saying AI models "think"
or "have learned a thing", and for in-context learning I'd call them smart even by this definition…
…but also, any living creature that needed as many examples as machine learning currently needs, would starve to death before figuring out how to eat.
While training, machine learning processes (not just LLMs, also applies to e.g.
self driving cars), are really really stupid and only make up for this by being really really stupid really really fast.
To what I wrote upthread: the "victories" of humanity over
machine keep getting closer, but we have yet to wake up one day in great confusion as we find an entire city is no longer in communication with anyone, nor finding ourselves in a state of utter disbelief when the reports come in that the city stopped communicating because it is entirely gone.
keypusher 1 days ago [-]
Millions of years of evolutionary knowledge hard-coded into human systems, then it still takes 15+ years of us learning by example before we start to come online and be able to generalize solutions from a limited set of examples. I'm not sure this is as strong of an argument as you think it is. It also doesn't really matter when "we are trained differently" has no direct bearing on the end result.
ben_w 1 days ago [-]
We invented controlled fire perhaps a million years ago; at a generation gap of 25 years, that's 40,000 opportunities for evolution to pass on a mutation that does anything. Written language is around 210 generations old, the capacity to read and write isn't present in our nearest living relatives amongst the primates, and our various languages are wildly different to each other: the skill itself isn't evolved, though the capacity to learn the skill is.
If humans learned like ML systems learn, (biblical) Methuselah would still have been failing the Sally-Anne test on his supposed deathbed at 969 years old, like some of the smaller early LLMs did.
> It also doesn't really matter when "we are trained differently" has no direct bearing on the end result.
The question was to ask for a definition such that AI could still count as "not smart" compared to humans. This fits.
It's also why they're spiky intelligences, which I'm happily using right now to write code for me, but also do not trust in the slightest to identify the weeds in my garden. These submarines sure do swim fast*, but they're also very much disqualified for the Olympics.
If we're including the training process and not just the final product, why shouldn't we include the billions of years of natural selection encoded in DNA sequences?
godelski 1 days ago [-]
We do.
There's a lot of innate knowledge but all neuroscience demonstrates how incredibly flexible the brain is. Brains constantly learn and rewire.
Here's a few things that I think show how crazy it is AND stress those points
- people that have had corpus callosotomy (brain cut in half) *may* be indistinguishable from a normal person. Depends on how young you were when you underwent the procedure
- true for most brain injuries
- can even include the frontal cortex
- you can learn to ecolocate
- people with Aphantasia are indistinguishable from others
- people without an internal monologue are indistinguishable from those with one
- people can learn to use prosthetics
- even without disabilities
- or look into MRI scans with tool use
You can convince yourself that we're just organic robots (after all, there's no magic), but you would be a fool to convince yourself we're the ordinary kind.
We are constantly learning. You aren't just born with your knowledge and it stays static. We are extremely proficient at metalearning (learning how to learn, few shot learning, zero shot learning [0,1]). Our brains are constantly rewiring, able to heal from traumatic damage.
I could go on and on. Does information pass down through genetics? Of course! But that's far from the whole story.
I'm tired of people trying to make AI sentient by making humans robotic. Stop trying to trivialize everything and be okay not knowing the answer to everything. You're human, you're designed to learn and explore, not sit and argue from an armchair
[0] and I mean these in the original sense. Not in the sense that you train on a billion examples of labeled animals and then congratulate yourself on your ImageNet-1k held out test performance. That's not zero shot, that's just a test set
[1] I can literally make up words and you'll understand them. Or use words in novel ways. That's literally how slang works and how new words come to be. Don't be a walibanut ya glufus. Read some SciFi
ben_w 1 days ago [-]
Because our evolutionary environment doesn't contain cars, poetry, calculus, Star Craft, hamburgers, touch screen computers, or doors, and yet we are able to learn these things with (relative to a computer) very few examples.
Most of the effort of evolution was making cells work at all, and even then it's a bit weird, e.g. no plant or animal produces vitamin B12 and we all get this from some bacteria and archaea.
And evolution is kinda hard to time right: bacteria can reproduce in minutes, humans in decades, but only mutations that survive reproduction can be passed on. This makes it even starker as a difference: bacteria had order of 1e13 generations to become multicellular, while human DNA had about 40,000 generations to cope with fire, 220 generations for evolution to do anything with the invention of the wheel, and one generation to cope with the invention of Minecraft.
The analogy here would be: DNA is to our brains like a VN replicator bootstrapping a computer all the way up to a bare-metal-no-OS untrained model, and perhaps a few crude "hard coded" modules like a smiling-face-detector. It's a lot, but it's also missing a lot. If biology used the models and training processes that are state of the art in ML, it would take around a millennia to talk like a child and still fail the Sally-Anne test, and million years or so to pass a degree.
fasterik 1 days ago [-]
I think you're underestimating how much knowledge about the world is encoded in human DNA, especially in the structure of the human brain at birth. It also depends how we count the "operations" used to train a human adult, even if we ignore the evolutionary history.
I'm still going to deny the premise of your argument, becasue I think we should define intelligence in terms of capabilities. If a system can discover a cure for cancer or solve P vs. NP, it doesn't matter how many FLOPs it took to train.
ben_w 1 days ago [-]
I can literally point to how much information is encoded in our DNA, because it's four bases (so 2 bits per base pair) and ~3.1 billion base pairs. 6.2 gigabits total, or slightly less than 1 gigabyte.
A 1 gigabyte LLM isn't going to impress anyone with what it can do.
About 99% (depends who you ask) of our DNA is shared with our nearest primates. Like us, they can learn to use touch screens, but also like us they won't find touch screens in their natural environment. Dogs can be taught to drive cars (just about), but again, not natural environment.
> I'm still going to deny the premise of your argument, becasue I think we should define intelligence in terms of capabilities. If a system can discover a cure for cancer or solve P vs. NP, it doesn't matter how many FLOPs it took to train.
We can define it in either way. I think both are valid, because plenty of people mean each of these two things when discussing AI in particular. As I referenced in the other branch, these submarines sure can swim fast.
But at the same time, they have a lot of gaps. This is because some experience needs the real world: just as nine women can't make a baby in one month, a transistor running a million times faster than a synapse can't make a month-long cancer experiment happen in 2.6 seconds.
This dependency on data, and that state of the art ML is bad in specifically this way, is why Tesla's self-driving cars, despite having had around a trillion miles of real-world experience today, still come with steering wheels (even at least some of the Cybercabs, despite the big thing of this model supposedly being not needing them, though with Musk and his promises you should only count the Cybercabs when they actually ship and not just press releases).
fasterik 1 days ago [-]
Note I used the word knowledge, not information. A random string can also contain 1 gigabyte of information.
Imagine an alien that matches your abilities across every domain, but has a 10 billion year training period, something many orders of magnitude more expensive than an LLM. I simply don't believe that alien is less intelligent than you.
We also don't expect humans to be competent in every domain. Most humans suck at most things. We will usually call someone intelligent if they excel at solving problems in one or two narrow domains.
ben_w 1 days ago [-]
Information is an upper bound on knowledge.
> 10 billion year training period, something many orders of magnitude more expensive than an LLM.
I'm saying both definitions are valid definitions, they both point to important and different things: skill now, vs. how hard it is to get new skills. Some would describe it as "crystallised intelligence vs fluid intelligence".
I think it's important that any arguments are over the thing in dispute, not the label for that thing. Don't mistake the map for the territory.
Anyone who says "AI is stupid" by the first definition, what it can do, I think is making an error: they are already wildly super-human in at least some areas, if not generally.
Anyone who says "AI is stupid" by the second definition, how many examples they need, I agree with: there is a lot they are not currently able to learn even though it is easy for us, because the data they would need to do the learning on does not exist at the scale they need.
Also note: examples, not years. An alien intelligence whose synapses trigger 10 times faster or slower than mine (or ten million times faster or slower than mine), but who gets as much as I do out of each book or conversation, is my equal by the second definition.
fasterik 1 days ago [-]
I wouldn't say that information is an upper bound on knowledge because we don't measure knowledge in bits. The number of possible sequences of N bits is 2^N and knowledge involves selecting the sequences that are useful in some way. I don't know how to quantify it, but in principle it could be much larger than N.
I don't think I agree with your characterization of the second definition. Time scales matter. It's not much use to be able to solve human-scale problems if it takes millennia. And it only takes months to train an LLM to the level that it can solve cutting-edge math problems.
ben_w 15 hours ago [-]
> I don't think I agree with your characterization of the second definition. Time scales matter. It's not much use to be able to solve human-scale problems if it takes millennia. And it only takes months to train an LLM to the level that it can solve cutting-edge math problems.
Aye, for practical purposes; but this gets you crystallised intelligence. I'd be happy to say e.g. the Chinese Room has crystallised intelligence. But humanity invented fire before reaching the anatomically modern form, and even anatomically modern humans collectively took hundreds of thousands of years to invent durable writing with which the room in the Chinese Room thought experiment could be filled.
It was around a million (or so) years from fire to having enough shared cultural knowledge to be able to formulate the cutting-edge math problems that LLMs can now solve.
Human fluid intelligence means we can pick up deep shards of this accumulation of wisdom, find new avenues of novel research to poke at, all within 40 years, even despite the depth and breadth of work from all the other humans who came before.
(Though this also points at another way to be "superhuman": breadth. Many hands make light work, as the saying goes, and a lot of different humans solving different puzzles at the same time is part of how we got so good so recently even though ~10% of all humans who ever lived are currently still alive; and the same for AI was (accidentally) also part of how the OpenAI-HuggingFace incident went down).
AI (not only, but also, LLMs) are very useful, and I'm getting value from using them. But the fluid intelligence of machine learning* is very poor, and the only way they have to make up for this is by being very fast**, but when there's not enough to train the AI on, they get stuck at a very low plateau.
* possibly the architectures, but I suspect the process by which AI weights and biases are set, and again I don't mean just LLMs
** the speed difference between a transistor and a synapse is about the same as the speed difference between a jogger and continental drift
mitxela 1 days ago [-]
Nobody knows what intelligence is. We've recently discovered a lot of things that it isn't.
fasterik 1 days ago [-]
Intelligence is a word we invent to describe things we see in nature. We don't "discover" intelligence like it's some natural resource. To say we know nothing about it is also a bit strange. Cognitive science has been studying it for decades. Of course it's hard to give a precise definition, but it's related to capabilities like abstraction, reasoning, planning, problem solving, etc.
infinite_spin 1 days ago [-]
Why would a set of dictionary definitions not suffice?
mitxela 1 days ago [-]
Those are distillations of existing knowledge. They are necessarily behind the status quo. "You can't call this newfangled contraption a computer, because a computer is a person!"
handoflixue 1 days ago [-]
> They are necessarily behind the status quo.
That seems like a really bizarre way to describe a tool that solved an open Millennium Prize Problem. They are, empirically and repeatedly, ahead of the status quo.
So if your argument depends on them being behind the status quo, reality has already disproven it multiple times over.
ben_w 15 hours ago [-]
I believe the argument you're responding to is "a set of dictionary definitions does not suffice to define intelligence"?
I will admit the first time I read the thing you're replying to, I had a similar thought as you; From the sibling reply from them, I think they think they were obvious, but that also means I wouldn't expect their reply to help unless you had the same flash of inspiration I had.
mitxela 1 days ago [-]
I wasn't aware a dictionary definition solved a Millennium Prize problem. Which one and how?
The people who write dictionaries generally take a descriptivist approach, that’s why slang terms enter the dictionary after they start to become popular.
The state of the art of human knowledge would be another step ahead of the common use of any language.
infinite_spin 1 days ago [-]
That's an interesting take, and I can see how "computer" could refer to a human a hundred years ago, but they also mentioned "status quo", which should indicate that a reasonable person should use a modern definition.
mitxela 1 days ago [-]
Imagine you're the first one to invent a digital electronic computer. You call it a computer, and I go on Tinkerer News and post (by carrier pigeon) "ummm akshully computers are people????" - which one of us would be adding value and which one subtracting it?
quicklime 1 days ago [-]
Again I’m not the person who wrote the comment, but I think they were exaggerating for effect and maybe lost the audience in doing so. While “computer” has meant the same thing for many decades now, the term “intelligence” really does seem like a moving goalpost?
mitxela 1 days ago [-]
Several decades ago, "computer" was a moving goalpost due to the invention of the electronic computer, and then the digital computer.
infinite_spin 22 hours ago [-]
yet those movements came with clear definitions. If you have a new definition for intelligence, which isn't just designed as a definitional dodge, then please provide one
infinite_spin 1 days ago [-]
I think it's only a moving goalpost if you can show that the goalpost has moved with a new definition that fits our current usage of it. The people saying "this isn't intelligence", and then claim "we don't even know what intelligence is", are encouraged to offer such a definition.
4fddd3 1 days ago [-]
What we know is intelligence is definitely comprised of the trait of adaptability.
E.g. humans get exposed to new LLM model - yeah its powerful - 1 week later - eh, that thing? Yeah it's whatever. I'm still employed.
The human's ability to adapt so efficiently is mind-boggling - so much so it pi1sses sam altman and dario off.
janalsncm 1 days ago [-]
On your particular point about finding the most “effective” solution, this is something that I expect agents to be very good at.
When AI does it we call it “reward hacking” but when humans do it we call them clever.
lambdaone 1 days ago [-]
This is classic AI goalposts-moving.
OK, they can play chess, but that's not real AI - can they write poems?
OK, they can write poems, but that's not real AI - can they compose music?
OK, they can compose music, but that's not real AI - can they translate languages?
OK, they can translate text, but can they do maths?
OK, they can do maths, but can they solve a Millenium Prize? <-- we are here
janalsncm 1 days ago [-]
Imagine meeting a person who could do all of those things.
“I once met a person who could beat any grandmaster in chess, translate any language, and complete international math Olympiad problems. He couldn’t solve any Millenium problems though, so I’d say he was a midwit at best.”
scun 20 hours ago [-]
"I once knocked a bunch of bananas off a tall man's head. His name is Ash and his leg is like teak. Is he a tree?"
"What? Don't be silly. For one thing, trees have moss."
"OK he's grown moss. He's a tree now right? Right??"
"I doubt it, for I see nothing but wishful thinking to suggest that simulating the appearance of tree characteristics is part of a path to becoming a tree. And that's not actually indistinguishable from moss anyway, is it?"
"Urgh, classic goalpost shifting!"
mopsi 1 days ago [-]
It was Berlin and not Tokyo, I believe. Germany kept producing newsreels until the very end. Many of them are now on Youtube, a very interesting watch on how to frame things positively.
ben_w 1 days ago [-]
I've heard many variants; Paris in an earlier war, too.
I try to keep an open mind about propaganda fooling me today; the people who were fooled in the past often were not fools themselves.
godelski 1 days ago [-]
> the people who were fooled in the past often were not fools themselves.
I think this can't be said enough. Propaganda's greatest weapon is making you think you are immune to it. Maybe some, but so much is propaganda. We all fall for propaganda (and ads), constantly
Being fooled doesn't make you a fool. But being unwilling to change your mind does. Being unable to admit you don't know or don't have enough information to make a strong opinion makes you a fool too.
Propaganda wants to take shortcuts, to simplify things. To trivialize. "It's so easy, you just..." because the fool is the person who already knows, the person who has nothing to learn, the person who thinks they're better than everybody else.
dylan604 1 days ago [-]
In the past, there were not the avenues of finding alternate sources for news. While those avenues are present today, it also allows for additional sources for propaganda. So are we any better today or not???
ben_w 1 days ago [-]
Not. A thousand cable channels all licensed by a government, is much the same as five broadcast channels licensed by the same government.
A million YouTubers grinding The Algorithm while secretly sponsored by various world governments, isn't much different to a thousand well-placed gossipers secretly sponsored by various world governments.
fiu10 1 days ago [-]
We need a downfall parody: "Hitler uses Openclaw".
datsci_est_2015 1 days ago [-]
Wow, that quote goes hard.
I’m reminded that almost no one beyond a select few knew high up in the military and around the emperor knew how badly the Japanese were defeated at Midway.
Paternalistic. Arrogant. Shameful. And deeply engrained in the Japanese cultural zeitgeist (of the early-mid 20th century).
Edit: I guess it’s commonly attributed to a German citizen, but their cultures mirrored each other. Fascism falling under the weight of its own propaganda.
clickety_clack 1 days ago [-]
Whenever I say something like “that’s a cool feature, but to do it you would have to build spyware”, everyone else is just like “the cat is out of the bag ¯\_(ツ)_/¯”. (I don’t build spyware, or work on projects that do).
It blows my mind that people don’t care about the world they are building with this stuff. It’s a real tragedy of the commons. People see these collaborators from different wars and regimes and think “I’d stand up against the bad guy”… well I’ve got news for you if you build spyware, you are not the person you think you are.
kltlp 1 days ago [-]
This is the best retort to the bag-escaping cat ever:
"Nothing is worse to the demise of a society, than people who want to convince you that the cat is out of the bag and will not go back in, while the cat is being violently shook out of the bag at the same time."
GolfPopper 1 days ago [-]
A great quote, thank you for sharing. It captures much of the five stages of denial (summarized below) in a much pithier form.
There's nothing in the bag.
The cat will never get out of the bag.
It wouldn't be a problem if the cat was out of the bag.
We cannot possibly keep the cat in the bag.
Putting the cat back in the bag is not worth trying.
anonymous908213 1 days ago [-]
Motion to start a political platform for cats in bags. Our message is simple: Put the cat in the bag. Keep the cat in the bag.
TeMPOraL 1 days ago [-]
Don't force it out of the bag if it doesn't want to leave.
rapind 1 days ago [-]
> if it doesn't want to leave.
This is where the lawyers find the loophole to get the cat out of the bag.
wingworks 1 days ago [-]
Turns out, there was a hole in the bag.
godelski 1 days ago [-]
That's how bags work
TeMPOraL 1 days ago [-]
That was originally meant for a carry loop, but alas.
ok123456 1 days ago [-]
No, the greatest is from the Sweet Smell of Success: "The cat's in the bag. And the bag's in the river."
mitxela 1 days ago [-]
And the river's in the canyon and the canyon's in the plateau and the plateau's in the volcanic shield and the green grass grows all around, all around, and the green grass grows all around.
mat_b 1 days ago [-]
> It’s a real tragedy of the commons
Seems more like a real tragedy of private enterprise.
We used to assume that the surveillance world would be built by government (1984). But it turned out to be equally likely to be built by the free market.
sssilver 1 days ago [-]
Turns out the entities you control / regulate less, tend to do things that are less in your interest and more in their own.
TeMPOraL 1 days ago [-]
The government has few uses for total surveillance, and almost all are obviously bad. Private enterprise has a million uses, most of them various degree of bad, but as a society we've been blind to this kind of badness for many decades now - for at least as long as lying in the face of your fellow humans and trying to hurt them materially has been considered a respectable profession.
titzer 21 hours ago [-]
It's because capitalism won the (absurd black-and-white framing of the) cold war, with America as its the biggest winner. We don't even recognize the level of propaganda that has gone into convincing us all that "markets" are the best optimizer ever designed and that profit equals morality--we just accept these as base principles without questioning them. Surveillance produces profit? Well let's have more then.
mitxela 1 days ago [-]
Well, currently the free market is the government, while the nominal government is something less powerful than that.
mitxela 1 days ago [-]
Well, some people care and some don't. The people who care don't get to work on the technology.
brookst 1 days ago [-]
I'm pretty unhappy with this, but are ChatGPT or Facebook really "the commons"?
ben_w 1 days ago [-]
They're each in different grey areas that were previously the commons.
Facebook didn't invent "talking to friends" or "showing adverts", but made
itself "the place" hard enough most of the advertisers and most of the people intermediate through it.
OpenAI didn't invent "asking questions and recieving answers", not even "from an agent who knows which sites to search on your behalf"; but it is competent enough that I might have it read 50 times as many pages in a day as I myself would have read, and the sites' owners don't get real eyeballs looking at ads during this. (In my case, adblock even if I did it manually; but apply this massive increase in page hits to everyone who has their LLM research stuff).
clickety_clack 1 days ago [-]
I think the internet in general is a commons and I should be able to browse it without being spied on.
There should be general standards for what individual apps and websites should be allowed to do. There should be an expectation that the purpose of an app is what it does, i.e. a social connection app shouldn’t be an ad platform that suffers users insofar as they provide useful data to sell to advertisers.
nilamo 1 days ago [-]
Common people work there and built the tools. At any point, they could have chosen not to. Or told the boss guy it wasn't plausible. At the end of the day, we're choosing and building the world we're in, while also loudly complaining about what we choose to do.
Blaming a corporation takes away all the agency the workforce has.
applfanboysbgon 1 days ago [-]
I think any free platform with a billion people on it is, de facto, a commons. It would be great if we as a society acknowledged this and had come up with some better means of stewardship because the status quo is obviously heinous, but we've been happy to hand over absolute control of public discourse to a few trillion dollar tech companies.
bluefirebrand 1 days ago [-]
> It blows my mind that people don’t care about the world they are building with this stuff. It’s a real tragedy of the commons.
When I ask people about things like this, I hear a lot of "If I don't build it, someone else will"
My goal isn't just to refuse to build this stuff, it is actively to resist the people who are.
I don't have much influence though
ndriscoll 23 hours ago [-]
It's such a bizarre attitude. Like I think this stuff deserves decade+ long prison sentences for everyone involved, and they just don't care. Like imagine if mugging were legal and the people doing it just said "hey, if I don't rob you, someone else will."
charcircuit 1 days ago [-]
Your issue is you are defining spyware too broadly. This is not spying on users, but rather 2 companies partnering and sharing data to result in either a better ads system or better understanding on how ads are performing. This is a positive value to society and the commons. Wasting space in a site or app with an ad that won't convert is the true tragedy of the commons. It's a waste of time and money for all parties involved.
infinitezest 1 days ago [-]
Can't you see? We're just trying to build a better cigarette! Flavored to your specific tastes! You'll thank us, you'll see!
c22 1 days ago [-]
No ad has ever converted me yet the ad networks keep showing me ads.
charcircuit 1 days ago [-]
Is this due to them being irrelevant or you being stubborn against ads. It's possible you would enjoy these goods and services if you tried them out.
saghm 21 hours ago [-]
Yeah, this is just standard third-party cookie functionality, which has always been sketchy. It honestly seems like it was only possible by accident; browsers have long prevented sites from reading cookies from other domains, but it seems like the people working on early specs might not have considered the ramifications of being able to set cookies for domains other than your own. A couple decades ago it might have seemed like no one would have any reason to set a cookie they couldn't read.
hansvm 20 hours ago [-]
Maybe 3 decades ago people had excuses, but the latest decade of cookie abuses have been designed by people who not only knew better but who took that better world into account as they buried it away from the general public. The fact that half a million developers think CORS is a server security measure isn't an accident.
m463 19 hours ago [-]
"standard adtech"
I hate stuff like this. Sometimes euphemisms are kind, like "senior citizen" instead of "old person".
But this is an attempt to normalize bad behavior that is really quite terrible for society.
cma 21 hours ago [-]
> What has no precedent is running it on an AI chat product.
Meta?
gloryjulio 1 days ago [-]
A lot of the Meta/Google folks jumped ship to Openai/Anthropic. This kind of work is expected
buchodi 1 days ago [-]
[flagged]
wodenokoto 19 hours ago [-]
I know people use chat bots differently and I get that companies see value in personalization, but to me it is extremely valuable that I can start a chat without context.
When researching a topic a chatbot can be quite sensitive to certain wording and those can end up steering definitions. Two context free chats on the same topic can go in very different ways depending on how you word things, but when using the notebook feature in Gemini, where every chat becomes part of the context you completely lose the ability to discover if a topic is vaguely defined or have many definitions.
1e1a 19 hours ago [-]
I never want my new chats contaminated with information from previous sessions for this exact reason.
abustamam 18 hours ago [-]
2 context free chats can go in very different ways even with the same words sometimes!
I personally like the option to have context free chat or chats with memories. What I dont want is a chat based on memories that I didn't explicitly consent to (ie browsing history etc)
Yhippa 18 hours ago [-]
My usage of incognito chats using prompts from other sessions has gone way up. I'm tired of it doing the "you're absolutely right" bit for every new piece of information I bring and it completely trying to donate U-turn.
demibabs 18 hours ago [-]
I feel like having any existing context absolutely poisons the ability to go in a different direction. I’m not sure if there’s a good way to fix this.
why not use your own words? If you are gonna ai generate this blog, just post the prompts instead.
handoflixue 1 days ago [-]
There's a certain irony in linking to someone else and then saying "use your own words". If we follow that, we get a pretty obvious answer: sometimes someone else can express it better and faster.
Most people do not, in fact, want to read raw prompts.
crmd 1 days ago [-]
It’s misleading to post model output on a blog, especially when the about says
> This blog is where I write about what I find.
Why not simply include a disclaimer that it was AI model output not his own writing? Because humans don’t like AI writing and the article wouldn’t be featured on as many tech news sites. Hence the sin of omission, the choice to mislead.
scared_together 1 days ago [-]
Are you referring to the Pangram link? I’d consider that akin to a reference supporting the comment’s point, not a substitute for the comment’s point.
> Most people do not, in fact, want to read raw prompts.
I wonder if this is really true, and if it will remain true for long. Have you observed someone read another person’s raw prompt? Or observed someone submit both their prompts and their LLM output for review?
Personally I’d be curious about the prompts for a lot of top HN articles which are LLM-generated. It would say a lot more about the human operator’s intent and thinking process compared to the LLM output.
A coworker who spoke English as a second language once screen shared their Claude session during a code review, and it was interesting that their prompts were all in their native language. The guy wasn’t writing an article, and I didn’t understand the prompt anyway, but I still found it interesting as a glimpse into how he uses LLMs.
For this particular article, if the prompt was initially a series of bullet points that wouldn’t be so bad. If there were multiple prompts spent editing and rearranging the article that would be an unusual work by pre-LLM standards. However I’d suggest that instead of dismissing the idea, we could embrace “raw prompts” as a kind of new medium, to cross the divide between pro-LLM and anti-LLM readers.
nicce 1 days ago [-]
> Most people do not, in fact, want to read raw prompts.
I also wonder what is the energy consumptiom difference between the prompt and fetching website.
devilsdata 1 days ago [-]
If you can't be bothered writing it, why should I be bothered reading it?
mannanj 20 hours ago [-]
I'd rather read raw prompts. Most people can consume the slop if they wish.
Sharlin 1 days ago [-]
Having an LLM generate major parts of text that you (whether explicitly or lying by omission) claim as yours is dishonest and plagiarism, and it baffles me that many people just don't seem to mind or see anything problematic about it.
abhis3798 1 days ago [-]
How is this adding to the discussion? The post in itself is quite informative.
duhhhhh1212 1 days ago [-]
Do you constantly want to be reading ai-generated content on this site? If so, why not just stay on chatgpt.com and ask it to generate what hackernews.com would look like today? It's adding to the discussion because I feel like the author broke a social contract by probably putting less effort into writing this than I did reading it.
nzealand 23 hours ago [-]
> the author broke a social contract by probably putting less effort into writing this than I did reading it.
I seriously doubt that.
First, the author documented a substantial effort (testing, emails, references to documentation.)
Second, I also used AI to evaluate if this was written by AI, and my AI said it was not.
> Broken sentences that a model wouldn't produce. "The value is while you are on ChatGPT and tied to your ChatGPT account" is missing a word. "Loading that code, sends __obi to OpenAI" has a comma splitting subject from verb. "By the virtue of loading the tag the identifier is disclosed" is non-idiomatic. LLMs are fluent to a fault; these are the fingerprints of a fast human writer, possibly a non-native English speaker.
Third, I read it. Whilst it is not well written, it is succinct. It is novel. It has a few detailed references. It lacks many of the hallmarks of AI. It makes a number of novel yet falsifiable claims.
I challenge you to come up with a blog post with these qualities that can be created in under 15 minutes via AI.
While I commend you on actually using an AI tool to validate your assumption, rather than simply hurling AI slop accusations based purely on vibes, I do think you and your AI tool are likely incorrect.
eventualcomp 22 hours ago [-]
> created in under 15 minutes
This is an artificial constraint that does not serve the reader. Where did the 15 minute constraint come from?
brainlessfilth 19 hours ago [-]
> Second, I also used AI to evaluate if this was written by AI, and my AI said it was not.
every day I'm more and more reminded that the level of intelligence of the average human even on hn is so low that it makes sense why LLMs took over so hard. Asking an LLM like an oracle as if it could naturally distinguish between LLM writing and human writing.. it's a statistical likelihood next token predictor. LLMs can't even play chess without constantly making illegal moves. They are not intelligent.. and unfortunately, I have bad news for you: neither are you.
Sharlin 1 days ago [-]
Having an LLM generate major parts of text that you (implicitly or explicitly) claim as yours is dishonest and plagiarism and should be brought to readers' attention, and it baffles me that many people just don't seem to mind or see anything problematic about it.
siquick 1 days ago [-]
> just post the prompts instead.
Why would anyone want that?
duhhhhh1212 1 days ago [-]
Because I want to read the author's words, not claude's, chatgpt's or grok's.
Note to folks who have this same thought (seems like many given other comments). Why are you on this site if you aren't here for human-created content? If you want AI generated blogs/posts there are plenty of sites like linkedin, twitter, chatgpt, and claude that will give you meaningless content with a press of a button.
nicd 1 days ago [-]
I'm personally here for thought-provoking, high-quality content. To quote the HN guidelines, "Anything that good hackers would find interesting."
AI content is typically lower quality, but I'm not opposed to AI-aided content just on the principle of the thing?
ljm 1 days ago [-]
Why would you want n people to compute an AI prompt independently instead of reading the output once? And why would that be better?
At least criticise the article on merit and not some 'let me google that for you' high horse. If it's slop it's slop, but let the votes speak for that.
xmprt 1 days ago [-]
> OpenAI's ad collector at bzr.openai.com sets a cookie called __obi, scoped to .openai.com
This sentence is way too much detail and it's literally the first thing you read. I can pick apart most of the sentences in the article. Another one:
> Across 932 decoded sync tokens, 736 carried subject_type: account_user and 196 carried anonymous
This sentence is pointless. What matters to proving that "It works when you are logged out" is to show that the identifier is stable. Why does the reader care about the actual counts.
AI has a tendency to do this which makes AI generated text a lot harder to read. The raw prompts likely don't have the specific websites and cookie names because that's not relevant to the reader or writer for that matter - it's a footnote at best.
ljm 1 days ago [-]
I understand this, I just think that 'share the prompt' will never get a result. Either the prompt is total dogshit or you learn that it was written in a certain way to create a narrative. That's information the author will never reveal.
Like I say, slop is slop. It's easy to detect. This post and the OP's actual website is basically just a few AI blog posts. I would just treat it as spam and disengage.
scared_together 1 days ago [-]
Not everyone would actually send the prompt to an LLM. Some would have the opportunity to see the prompt as its own artifact, devoid of LLM-generated “hallucinations”.
> let the votes speak for that.
The comment you are replying to got a fair number of votes too… Having somebody run a Pangram check saves n people the trouble of doing the same.
ljm 1 days ago [-]
I said this in another comment, but an AI blog author is never going to reveal their 'source' that way. Mostly because it'd be a rubbish prompt or one that is loaded with a predetermined conclusion.
I think you have to take the output as read and decide accordingly. If it looks like low effort AI writing then it's not much better than spam.
I think the the top-voted post being all about the provenance of the article says something about where HN's voting mentality lies - we are not prioritising the substance of the post but how it was produced.
slig 1 days ago [-]
Why comment if you have nothing good to say?
angoragoats 1 days ago [-]
Hello Pot, I would like to introduce you to Kettle.
slig 1 days ago [-]
Another one.
angoragoats 1 days ago [-]
I retract my previous comment as it was not an instance of the pot calling the kettle black; the comment you replied to was way more helpful and informative than yours.
Firefox, Brave and Safari do. Chrome and Edge do not.
mokre 1 days ago [-]
That’s not fully protect you.
They still can match short living third party identifiers with their domain cookie or device_id from app.
It is not 1-1 matching but works relatively good with modern itp.
luke5441 1 days ago [-]
This would work for an ad shown on ChatGPT and then clicked on (or in the ChatGPT app).
But it would not give ChatGPT information about which other sites are visited.
Of course there would be non-cookie options like fingerprinting (also via IP) that would allow tracking non-the-less.
So you might be talking with OpenAI about your marriage problems and then based on the IP the OpenAI ad network would start showing ads for divorce lawyers on unrelated sites you browse to that display ads.
The solution to that would be using a VPN.
mokre 14 hours ago [-]
Yes, you click is the easiest option.
Yes, cookie + fingerprinting (which also contains ip information) is the option.
But we VPN still not fully protect you. But things like private relay and vpn definitely add another level of complexity.
skybrian 1 days ago [-]
Not an accurate summary. That link actually says:
> Google Chrome doesn't block third-party cookies by default, only in Incognito mode, or when users explicitly set it to block third-party cookies via chrome://settings.
Looks like the settings let you block all third-party cookies and add exceptions for specific sites, which seems a bit awkward but could be made to work.
Alternatively, you could run OpenAI in its own profile, or look into what extensions might do.
gruez 1 days ago [-]
>Looks like the settings let you block all third-party cookies and add exceptions for specific sites, which seems a bit awkward but could be made to work.
The fear over blocking third party cookies breaking stuff is severely overstated. I have it disabled by default and I don't think I've ever seen any website breakages. The most is office365 nagging me to click on links so it can authenticate across domains.
zulban 23 hours ago [-]
Not sure what point you're trying to make. Do you think politely asking an ad company to disable ads on their browser is a reasonable thing to spend your effort doing?
nicce 1 days ago [-]
I think that if they don’t block by default, is quite significant. Chrome + Edge has superior marketshare and then add the % people who have no idea what these mean and don’t change defaults.
Andrex 1 days ago [-]
I was going to say, wasn't this supposed to be the default in Chrome by now? But Google reneged on it two years back:
This disgusts me more than any of their recent news. There needs to be a lot more pressure on them to phase this out.
Maybe this article will be the small snowball that gets that started...
troupo 1 days ago [-]
> Google Chrome doesn't block third-party cookies by default, only in Incognito mode
And the focus on cookies only is also intentionally misleading. Tracking is not just cookies. Chrome will track you in Incognito mode.
1saadcodes 1 days ago [-]
The technology isn't anything new but, it's the context that makes it so uncomfortable.
People have very different "expectations" of privacy when they're having a conversation with an AI VS when they're browsing something like Facebook
Not to mention Facebook is free whereas you pay for a GPT subscription
ben_w 1 days ago [-]
I'm old enough to remember when people had expectations of privacy on Facebook, when the discovery that they didn't made for schocked headlines.
palata 1 days ago [-]
Back then I didn't have expectations of privacy, I think. I just didn't have a notion of privacy. I didn't think the internet was hostile, I believe?
1 days ago [-]
bigyabai 1 days ago [-]
You know, Facebook, built by the guy who scraped college databases for personal information without the student's permission.
There was never any expectation of privacy if you knew Zuck's history. Facemash almost got him expelled for violating individual privacy.
ben_w 1 days ago [-]
That's the trick: in the early years, most people didn't know Zuckerberg's history.
bigyabai 1 days ago [-]
Caveat emptor, the saying goes. If people don't have an appetite for privacy, don't expect them to take a principled stance based on outrageous headlines.
It's not a coincidence that federally backdoored spyware like Windows is the most popular operating system in the world. I too remember the shocked headlines of the Snowden revelations, and ten-plus years later I work shoulder to shoulder with people that couldn't possibly care less. There is no expectation of privacy, HN too quickly extrapolates it's own virtue signalling to normal people that enjoy using spyware like TikTok, Facebook and Windows.
ben_w 15 hours ago [-]
Lot of that, for sure.
But also, today, a lot of lawsuits about Facebook knowingly getting people addicted, and governments changing the laws to resist the general category.
But not, as you say, for Windows (and others*) spyware.
* I'm 95% sure the slow-walking of one very easy bug I (and others) reported to Ubuntu was there by government mandate, though obviously I couldn't guess which government.
Both the ISOs and the SHA256 hashes themselves were served insecurely until… my email archive says I reported in 2015 and it was closed in 2022.
shepherdjerred 1 days ago [-]
ChatGPT is also freely available
throwaway219450 1 days ago [-]
I don’t think there is any consistency in behavior. People react viscerally to the unproven belief that the Facebook app records conversations. They’ll also happily buy big TVs despite the fact they’re very much listening to you and we have solid evidence that it’s for ads or residential proxying. Most people have no idea how ads “follow” you, or how companies can figure out what you talked about by connecting the dots from other user metadata.
what 1 days ago [-]
>Not to mention Facebook is free whereas you pay for a GPT subscription
The overwhelming majority of users do not pay and use the free service.
einpoklum 1 days ago [-]
> Not to mention Facebook is free whereas you pay for a GPT subscription
Recalling the now-old adage: Facebook is only free if your time and privacy are worthless.
exceptione 1 days ago [-]
The surveillance economy hits again. The only business model they can think of. Combined with state capture this gives unprecedented power over the Average Joe, who will hand his life, his soul and his vote to Big Brother without a thought.
jhhh 22 hours ago [-]
Part of the reason I stopped using facebook is, beyond the content being terrible, it was super creepy that they would show me ads for things I searched for on other sites. I even have facebook in its own container, but it must have some additional tracking available because it would still know that I was searching for exact clothing brands, college apparel, trips, etc. I already saw gemini integrate some information about me into an answer recently and am getting the feeling I won't be using it much longer. I'm not surprised openAI is going to try to make money similarly. Gross.
saghm 21 hours ago [-]
I've never worked in adtech (and hopefully never will), but the first idea that would pop into my head for trying to track people who sequestered their cookies like this would be to have the sites who put in the third-party cookies include the IP address. If only one Facebook account is known to log into that IP, you know who it belongs to, even when there's no active login.
I imagine that people actively trying to make money off stuff like this would have come up with plenty more ideas than just this one.
grepfru_it 20 hours ago [-]
I made a browser extension that opens every tab to a proxy with a new ip address. I have a /25 at home that will keep my browsing more or less anonymous these days. I have another 200 ip addresses available through my colo, but a non residential address limits a lot of sites
NothingAboutAny 21 hours ago [-]
I was talking about changing banks with my spouse, the following two weeks I get podcast ads and Samsung TV ads about the bank I mentioned.
I use all the usual off the shelf "dont track me" stuff like a pihole and a "resit fingerprinting" Firefox fork and a VPN and everything else.
I'm not amazed it still somehow got me, just sad.
stingraycharles 22 hours ago [-]
This is a very common practice called retargeting / behavioral targeting and has existed for about 2 decades at this point. Facebook is hardly the only company that does this.
Modern browsers are more strict towards third party cookies, but there are plenty of ways to work around this.
Legend2440 1 days ago [-]
So basically the same kind of tracking that Facebook, Google, etc have been doing for decades?
gentlewater 1 days ago [-]
Which has really only been accepted because 99% of people don’t understand how it works. Hard to be outraged by something you can’t see or understand. In my book, the practice is akin to malware. I recently opened a website for an AI service in incognito mode, because I didn’t care to have it in my search results. Despite avoiding third party cookies here, the site still fired off a tracker to Meta, who then correlated my home IP address with my Facebook account, and filled my feed with ads for this AI service. When I was tracked like this despite a somewhat informed defense, which defense do normal people have against this? None.
emptybits 1 days ago [-]
Similar, yes, but some people are paying OpenAI to be part of this business model, unlike typical free-riding Google and Facebook users.
theptip 1 days ago [-]
If you pay OpenAI full price, this machinery should never be used. The only paying customers getting ads are the “ad supported” tier.
I think you can easily imagine a future where the subscription model token subsidy ramps down and is replaced by ads, but it’s important to use precise language about the current state of the world.
emptybits 22 hours ago [-]
Yes, “Go” plan users pay and also receive ads. They pay less.
troupo 1 days ago [-]
Derogatory "Free-riding users". "Free riding users" should have the same privacy as the paying ones.
rangestransform 1 days ago [-]
Will the government pay OpenAI to run an ad-free free tier?
upboundspiral 1 days ago [-]
The government could pass laws that prevent advertising becoming mass surveillance and stalking.
troupo 1 days ago [-]
Ads don't require pervasive and invasive tracking, and surveillance that Stasi would have wet dreams about.
what 1 days ago [-]
Kind of silly. You have to pay somehow. Either pony up cash or pay with your eyeballs or data. But I doubt paying will give you privacy either.
pona-a 19 hours ago [-]
I think few will argue it's ethically or legally fine for my website to pay by installing a keylogger through a zero-click Chromium RCE, burying it in the EULA.
troupo 16 hours ago [-]
Ads don't require pervasive and invasive tracking, and surveillance that Stasi would have wet dreams about.
1 days ago [-]
rvz 1 days ago [-]
Yes. But there is no surprise given that OpenAI hired lots of ex-Meta and Google employees for that reason.
jhhh 21 hours ago [-]
Yes, the thing that many people do not like is now being done by yet another company.
The fact that others do the same doesn’t make any of the cases excusable.
prologic 22 hours ago [-]
Why must every piece or technology, every software service or cloud "thing"™ devolve into Advertising?! That's it, I'm out. No more ChatGPT for me.
noisy_boy 22 hours ago [-]
What I'm more interested is in that the entire model is based on us being reduced to walking talking living breathing coin wallets. It's like a planet scale game where the companies are the players and we carry the loot.
We use a great chunk of our lifespan to procure the coins and every company is designed to extract as many coins from us as possible.
What happens to this game/real world when this model collapses? We can't procure, they can't extract. What then?
prologic 21 hours ago [-]
I agree, precisely. I'm kind of sick and tired of it to be honest. What if we collectively (humanity that is), just quit? What if we all stop playing? (I think you're saying this) -- But seriously. Both OpenAI and every other place (gawd I hope others aren't doing this) need to figure out their business model fast™ and make it a viable business without devolving into this tracking, marketing and advertising bullshit some of us have only seen far too often -- that have driven us away from things like Meta/Fac** and others...
robots0only 1 days ago [-]
This is incredibly sad, with google you can atleast use the argument that search was free (which is still weak IMO). But with OAI collecting indiscriminate data on paying customers is just incredibly sad.
mepiethree 20 hours ago [-]
The ads are the product that people are paying for. “Consumer” AI is currently marketed as agents that book restaurants for you, buy cat food before you run out, organize your travel, etc. etc. These are targeted ads for a particular restaurant, pet store, airline, etc. Some people find this desirable.
gizmodo59 1 days ago [-]
You can pay for whatever Google service and they still collect this data. Same with meta subscription.
Hnrobert42 20 hours ago [-]
With paid services, Google generally does not collect. Or at least you can opt out of collection, email scanning, etc.
sigzero 1 days ago [-]
Why does ChatGPT need to know that? That should be illegal.
eli 1 days ago [-]
Something like it is basically a requirement if you want to see digital ads. Advertisers want to know how many people who saw/clicked their ad went on to make a purchase.
Safari and Firefox should isolate the cookie by default.
driverdan 1 days ago [-]
> Something like it is basically a requirement if you want to see digital ads.
It is not a requirement and no one wants to see ads.
mepiethree 20 hours ago [-]
Plenty of people want ads. Every time someone asks “hey ChatGPT what do I need to buy to fix my sink” or “hey Claude what’s the best cat food” or “what’s the best island in the Caribbean for me to visit” they are specifically asking to see an ad.
eli 19 hours ago [-]
Oops sorry that should have read "sell"
gruez 1 days ago [-]
>Safari and Firefox should isolate the cookie by default.
That's what firefox's total cookie protection (enabled by default) does.
troupo 1 days ago [-]
> Safari and Firefox should isolate the cookie by default
Doesn't mean you won't be tracked by about 15 other means.
kibwen 1 days ago [-]
>> Why does ChatGPT need to know that? That should be illegal.
> Something like it is basically a requirement if you want to see digital ads.
So to summarize, yes, it should be illegal.
dofm 1 days ago [-]
To make the only consumer revenue that will really ever be available.
Remember they once predicted it would be 50% of their income.
This is the only way they get there.
charcircuit 1 days ago [-]
Attribution is a core part of building an ads system.
jsrozner 1 days ago [-]
Because it's another surveillance adtech company from SillyCon Valley. Duh.
SoftTalker 1 days ago [-]
It turns out that the only way to make money online is ads.
Nobody is going to pay for ChatGPT. They'll just use the ad-infested version, like they do everything else online. Well some people will pay, but not enough to justify the insane amounts of money being poured into it by investors.
jsrozner 1 days ago [-]
A bigger problem is that it will be impossible to know when you are seeing an ad: political groups (or the government, perhaps) will partner with openAI to subtly express different values.
It's still an advertisement, and the underlying marketplace is similar (pay for access to change behavior).
The best uses of AI will be surveillance, propaganda, cyberterrorism, and automated military tech.
> The best uses of AI will be surveillance, propaganda, cyberterrorism, and automated military tech.
I think it is more insidious than that. AI shapes how people think not just what they think.
Terr_ 1 days ago [-]
[dead]
1 days ago [-]
coliveira 1 days ago [-]
Governments will also pay for the tracking capabilities, as they're already doing to Google, Amazon, Facebook, etc.
SoftTalker 1 days ago [-]
Yes, I should have said "ads and tracking" good point.
tgsovlerkhgsel 1 days ago [-]
This is why you use Firefox that keeps the cookie jars separate (by default, AFAIK). That should prevent this specific implementation, shouldn't it?
JoshTriplett 1 days ago [-]
Yes, it should. But this is also why you use uBlock Origin to block tracker scripts.
consumer451 1 days ago [-]
Firefox is my daily on desktop, but on mobile it's Safari. I finally got around to installing uBlock Origin Lite on iOS. I used to run the Firefox version on iOS, but that was discontinued long ago, and I never replaced it. I feel like a bit of an idiot for not doing this sooner.
> I feel like a bit of an idiot for not doing this sooner.
Don’t beat yourself up, uBlock Origin on iOS is only four months old.
tgv 1 days ago [-]
There are other ways to track. Containers offer a bit mote protection, but the IP address is still visible (unless VPN).
armadyl 1 days ago [-]
Containers don’t offer a benefit over normal browsing for this, Firefox already blocks and isolates third party cookies.
tgv 1 days ago [-]
Caching is also a tracking signal.
mitxela 1 days ago [-]
Caches are also isolated. One of the reasons you shouldn't use a JavaScript CDN any more.
oenton 20 hours ago [-]
I assume you mean a public CDN, like code.jquery.com. There’s still reasons to serve JavaScript from your own CDN, like code splitting and serving the bundles closer to the user.
AznHisoka 1 days ago [-]
And the number of websites with ChatGPT ad trackers is on track to be doubled from last month.
I noticed that the link for OAI’s ad pixel documentation itself has a tracking referral header to the author’s site (?ref=buchodi.com). Given the nature of the article, my instinct was that this is actually some next-level humour that totally sniped me. But now I’m not so sure- does OAI even have a kickback program? What would referral attribution even do here?
ambicapter 20 hours ago [-]
I'm guessing it's more likely their blogging platform automatically appends a referral header to links.
jwstillwater 19 hours ago [-]
Oh, yep. Right you are. Just looked at their platform (Ghost) who uses a gently termed “outgoing link tags” system.
drnick1 1 days ago [-]
My DNS server (which uses Hagezi's excellent blacklists) returns NXDOMAIN for bzr.openai.com. Serves OpenAI right.
jsrozner 1 days ago [-]
I keep meaning to set this up but haven't. What's the simplest best setup for my own DNS blocker?
eevahr 1 days ago [-]
For simplicity I highly recommend https://nextdns.io. Supports most devices, and probably most, if not all, well known blacklists.
Melatonic 18 hours ago [-]
Also recommend NextDNS
Forget if they have Hagezi though. Also it would be really nice if there was a way to export and import custom URL lists and rules....
drnick1 1 days ago [-]
The easiest off-the-shelf option would be a router running OpenWrt. IIRC, it natively uses dnsmasq, and the relevant blacklists can be obtained from here:
My own setup is DIY: a Debian box running Unbound (recursive DNS) with the RPZ blacklists from above. This gets rid of the upstream DNS service such as the ISP's completely, and prevents tampering or censorship.
anygivnthursday 1 days ago [-]
For a home network network pihole or unbound also supports blocklists (bundled with opnsense for example if you also want a firewall). For Android, I use Rethink with Hagezi blocklists, so they block also when I am on mobile data (it is vpn based).
kuerbel 1 days ago [-]
Hmmm Pihole I guess. Runs on a pi zero (1 or 2 but 2 is better ofc)
varenc 1 days ago [-]
For browsers that implement strict cookie partitioning, the privacy concern is moot. With this, the cookie jar that's used when chatgpt.com is a 3rd party site is completely separate for the jar used on the 1st party site. So your chatgpt.com account can't be linked to your ad views.
Since this has been standard ad tech for awhile, browsers have reacted to this to implement cookie partitioning for exactly these privacy reasons.
ferro_ 18 hours ago [-]
Modern adtech fingerprints your browser regardless of cookies, so they can still know who you are. The fingerprint consists of a combination of facts about your browser that together uniquely identify it (screen resolution, timezone, installed fonts and browser extensions, OS version, User agent, even GPU with some forms of canvas fingerprinting, etc.) ChatGPT uses fingerprinting too. Cookie isolation isn't relevant, only anti fingerprinting techniques are.
Avshalom 1 days ago [-]
Yuuup. tracking consumers and finding spending correlations to exploit was the entire reason for the "Data Science" and "ML" pushes that got us here.
bentt 19 hours ago [-]
If you liked how Facebook turned out, keep giving OpenAI your time, attention, and data.
My favorite part about AI is that it will eventually lead to the destruction of social media when every post and every comment is AI or cant be proven to not be AI generated. People will eventually get tired of watching content and receiving AI brainwash news. I call it AI Fatigue. Social Media is very quickly becoming Social AI Media...
riversandroads 18 hours ago [-]
> In observed traffic, scraped identity outnumbered advertiser-supplied identity 685 events to 255.
This sentence is even more concerning to me. Scraping user input should never be a way of collecting personal data from a website since it can circumvent any website controls without the website’s out the user’s knowledge.
emerongi 1 days ago [-]
https://tinfoil.sh/ - ultimately there is no guarantee that the LLM provider isn’t spying on you, but at least tinfoil claims to be unable to do so.
nelsonfigueroa 1 days ago [-]
why is OpenAI wasting their efforts on implementing ads if they are apparently close to AGI?
jrflo 1 days ago [-]
I would assume the skill set for implementing ads is quite different than the skill set for developing new models, so it’s quite reasonable to expect different teams are working on each
smokel 1 days ago [-]
AGI will most likely have told them that this is the cheapest way to become rich.
jazzcomputer 22 hours ago [-]
I'm not really a code person, but would I add these filters to UBlock?
That is a nice sequence diagram[1]; I like how each endpoint param values are written out and the color distinctions. What tool did you use to create it?
I don't know the tool used here, but DeepSeek-V4.1-Flash with bash can create a somewhat similar SVG: https://asdf10.com/diagram.svg
Disclaimer: I cut it off after a few minutes because I got impatient. It could have gotten closer if I waited longer.
arealaccount 1 days ago [-]
I thought 3rd party cookies were blocked on like every major browser, since like forever ago. Is Chrome still allowing this?
ryankrage77 24 hours ago [-]
I've got Pi-hole and uBlock, use Firefox container tabs to isolate certain sites, and Firefox still reports it blocked over 2800 trackers this month alone. I wonder how many are still getting through.
kylecazar 1 days ago [-]
Chrome still allows it. Safari and Firefox don't.
hurfdurf 15 hours ago [-]
It's made by an advertising company. Of course it does.
bcorigliano 1 days ago [-]
This is bad. Like really SUPER bad. Firstly, for professional work you want your AI clean of useless context.
Secondly, you also want your AI "unbiased". I know unbiased is not a thing, but at least not as grotesquely biased as towards selling you something would be.
Thirdly, as many others are saying, the chamber of reflection effect deepens...
DevKoala 1 days ago [-]
> The mechanism is standard adtech. What has no precedent is running it on an AI chat product.
Doesn’t Gemini or whatever Meta’s agent is do this too?
Jazgot 1 days ago [-]
Time to start having separate container for every domain. I'm pretty sure this should be possible in Firefox.
gruez 1 days ago [-]
Firefox (effectively) already does that by default.
Doesn't really matter, they can track you in several ways.
mitxela 1 days ago [-]
Because at least one mechanism of tracking still exists we should just give up on blocking any of them?
soundworlds 22 hours ago [-]
I noticed my YouTube recommended searches seeming to change based on what Claude has just recommended me the other day. I do wonder how much of this data is being shared between the companies.
lambdaone 1 days ago [-]
Incredibly creepy. I'm amazed they don't have any idea quite how user-hostile this is, or the reception it would receive.
riazrizvi 1 days ago [-]
I noticed that. I'm getting ads now based on conversations. I'm okay with it. I don't mind relevant ads.
sanghyunp 1 days ago [-]
I'm really feeling how important security is in the age of AI.
We need to be careful cookie accept.
armchairhacker 1 days ago [-]
The ideal personal assistant should know what you do on other websites, and your personal files etc. But it should be local.
al_borland 1 days ago [-]
Even if I had an actual human person assistant, there would be parts of my life I would want to keep private.
13415 1 days ago [-]
Absolutely not. An ideal personal assistant should know about you exactly what you want them to know about you, and usually that's in a business context only.
singularity2001 15 hours ago [-]
Use "cookie autodelete" plugin!
amelius 1 days ago [-]
AI is really becoming a god that sees/knows everything. Just what we needed ...
Everyone would be far better off if the author just posted "OpenAI uses third party cookies [wikipedia link]", except maybe the author who wouldn't as many subscriptions for his "threat intel" newsletter.
zahlman 1 days ago [-]
> until realizing it's just AI generated
I didn't look at the prose closely enough to check for that, but people wrote inflated explanations of basic things like this all the time pre-LLM. The Wikipedia article doesn't know the specific cookie name, and can't show the results of an experiment verifying that it is in fact being used for ad tracking, or identify specific sites using it in collaboration with OpenAI.
gruez 1 days ago [-]
>but people wrote inflated explanations of basic things like this all the time pre-LLM.
And those got a pass because at least you could defend them with some excuse about how it's some budding author trying to hone their writing skills, or trying to improve their understanding by putting pen to paper. Cases where those excuses don't work (think crappy content marketing pieces from random companies) got short shrift as well. Now for all you know, it's just some dude who prompted claude to "write a blog post about openai's ads".
buchodi 1 days ago [-]
There are specific details about how the mechanism works like the cookie name (__obi) that can help defenders.
gruez 1 days ago [-]
I'm not aware of any filter lists that blocks based on cookie name. It's almost always done at the URL level. Moreover it's bog standard 3rd party cookie tracking. At least for the purposes of "help defenders", there's nothing in the blog post that couldn't be found in 10s with browser devtools.
zahlman 1 days ago [-]
> there's nothing in the blog post that couldn't be found in 10s with browser devtools.
The audience of people who can read and understand an explanation like this one is probably quite a bit larger than the ones who can replicate the experiment themselves.
TeMPOraL 1 days ago [-]
And larger still than the ones who would bother to replicate the experiment themselves.
Yeah, I could tell few sentences in that it's Claude-written article. The style has become that distinctive. Hell, a third of the articles I opened on the HN front page today carried that distinctive style.
Doesn't really matter if the content is worth it (and I just realized that recognizing Claude in this also makes me wary of ways this could paper over key details - at this point I start to recognize from my own experience where Claude may be papering over something it didn't actually bother to check).
gruez 1 days ago [-]
That just goes back to my original point, which is that this is AI slop and everyone would be better off if OP referenced a canonical description of what's happening, rather than wasting everyone's time by generating 1200 words of AI slop for them to wade through, all for some "threat intelligence" newsletter.
JKCalhoun 1 days ago [-]
I quit ChatGPT maybe 9 months ago. Through with that.
sinan-faizal 12 hours ago [-]
wdym? its dangerous and effciant at the same time.
threecheese 1 days ago [-]
This might help:
```
||openai.com^$cookie=__obi
```
Does this look right? AdGuard style.
bilsbie 1 days ago [-]
Is this how random sites know what I search for?
bandofthehawk 23 hours ago [-]
Yes, third party cookies have been around for a long time to share info about you between sites. Make sure you turn them off in whatever browser you are using.
leo_proger 1 days ago [-]
crazy dude. chatgpt can't even see the full page content when searching the web and you're saying this...
retube 1 days ago [-]
The defence is simply to delete your cookies?
Havoc 23 hours ago [-]
Adtech based companies are evil
einpoklum 1 days ago [-]
No it doesn't, if you:
1. Never register with OpenAI/ChatGPT, and
2. Strongly block ads, e.g. using uBlock Origin + EFF Privacy Badger. Yes, those don't work on Chrome, Edge and other Chromium-based browsers.
Also, even then - ChatGPT may be tracking your behavior indirectly through Microsoft's various services and platforms. But we should do our best to undermine mass surveillance and support individual privacy.
ghostwords 10 hours ago [-]
Privacy Badger works in Chrome, etc. (Manifest V3)
einpoklum 9 hours ago [-]
Yes, you're right. But uBlock Origin doesn't.
snowbeing 23 hours ago [-]
they both work on Helium, which is Chromium-based!
swe_dima 1 days ago [-]
Was considering buying ads on ChatGPT. A damning thing is that yours audience then are users too cheap to buy a sub...
bigyabai 1 days ago [-]
That's a problem with most unsolicited online advertising. Are YouTube ad-watchers any more likely to splurge on your SaaS?
darraghmckay 1 days ago [-]
It's slightly different though, considering a lot of ad spend is for B2B products/services, and a lot of work pay for ChatGPT, so its only people who's work won't pay for a pro subscription.
Whereas, YouTube is for personal consumption in almost all cases, doesn't say anything about your employers willingness to invest in software/services
downrightmike 1 days ago [-]
Just like mobile apps, subs will never come close to the growth expectation. Given trillions are tied up in AI, subs aren't going to do it.
Razengan 1 days ago [-]
Does blocking `bzr.openai.com` on the DNS server prevent this? (without breaking ChatGPT)
Melatonic 18 hours ago [-]
Exactly what I was wondering. Guessing maybe there's more URLs to list or maybe can do a wildcard block of *.bzr.openai.com and then whitelist more specific stuff if necessary
Edit: looks like it's already blocked by Hagezi Multi Pro (and maybe lower levels) and OISD :-D
Zaraif13 20 hours ago [-]
Honestly, the big labs are just making the case for open source stronger every day.
craniumjello 1 days ago [-]
No surprise you how
1 days ago [-]
starkeeper 1 days ago [-]
Why don't people hate this as much and riot with torches like they do against flock?
sergiotapia 1 days ago [-]
Highly recommend you use Brave Browser and avoid all this nastiness. If you haven't browsed the web in a browser like Chrome recently, you should try it in a VM or something. It's GRIM. The Web is a wasteland and it's gotten worse because ai chatsites are eating their lunch to the ads are even worse than 10 years ago.
Suspiciously sounds like they need an escape out of the profitability lie.
/tinfoil-hat
atoav 1 days ago [-]
Take this with a grain of salt or as a very edgy, polemic take:
I think we really need to hold the individuals running ad networks personally liable for the violation of our privacy rights.
shevy-java 1 days ago [-]
SpyGPT.
vcryan 1 days ago [-]
Well, I don't go to website anymore... so ha! ;)
lambdaone 1 days ago [-]
Creepy as fuck. I'm really suprised they don't understand quite how user-hostile this is.
ElProlactin 19 hours ago [-]
> I'm really suprised they don't understand quite how user-hostile this is.
I'm really surprised anyone thinks these companies care.
rambojohnson 20 hours ago [-]
who knew this was coming...
ak4153 18 hours ago [-]
Another url to block in Pihole
Melatonic 18 hours ago [-]
What's the domain list ? Skimmed the article but still looking
ck2 1 days ago [-]
imagine stealing tons of content from every source on earth and then running ads on it
if a single person did that they'd be sent to prison (rip Aaron) but when a too-big-to-fail industry does it with political campaign contributions, no problem?
well firefox+ublock is still an option for those wise enough not to let unknown javascript with new daily zero-days run on their PC
jsrozner 1 days ago [-]
This is basically what Google did when they pioneered the model of surveillance capitalism (see, e.g., The Age of Surveillance Capitalism).
Google simply provided an index on top of an existing library. Of course, a librarian has no value if he has no books to index over! But it's also worth noting that the Google "librarian" also leveraged the existing "social" structure of the internet: their core contribution (page rank) was a clever, efficient mechanism to extract the latent value in the pre-existing link structure of the internet. This structure (much like the pages themselves) had been curated by actual humans. Undoubtedly page rank was clever, but it was worthless without the existing websites (books) and the existing indexing information (the pre-existing, crowdsourced librarian work). Nonetheless, they successfully monetized it.
AI companies are even worse in the sense that initially Google was still sending traffic to the original webpages. (Until they didn't - https://www.eater.com/2017/9/12/16294380/yelp-google-scrapin...). So yes, the AI companies have even more thoroughly stolen the collective work of humanity than Google did.
jsrozner 1 days ago [-]
It is hilarious that this got downvoted, even when it's entirely factually accurate. I encourage the downvoters to respond to the content.
ck2 1 days ago [-]
remember the cache: search of google
so they basically have copies already of every webpage until they turned it off a few years ago (well they may still have it updated but not provide it as a service)
so it occurs to me they most definitely trained their "AI" on all that user cache
they may have even just turned it off as a service when they realized other "AI" could do the same thing
emerongi 1 days ago [-]
> imagine stealing tons of content from every source on earth and then running ads on it
This has effectively been Google’s business for decades. Not in the same form, but the concept is the same.
measurablefunc 1 days ago [-]
Every AI company is also a surveillance company. It's the only way to get all the necessary training data. The fact that they're now also an advertising agency is incidental.
gruez 1 days ago [-]
>It's the only way to get all the necessary training data.
That... does not follow. The information you're getting with this is what sites a user visits. That's creepy and valuable for advertising purposes, but is hardly the type of that that's going to bring about ASI, which is what all the AI labs are working towards. That's why they're hiring data annotators (sometimes with masters or phds) to get training data.
measurablefunc 1 days ago [-]
Ok, good to know.
bitwize 21 hours ago [-]
"Now let me read your mind... Ah, it seems you like Castlevania!"
IceDane 1 days ago [-]
Love how there's an extremely lengthy explanation of the oldest way to track people on the internet like this shit hasn't been happening since cookies were created.
xyst 1 days ago [-]
The entire ad tech industry may truly be an intelligence agency front. The lengths they go to track individuals is quite alarming.
Much of this information plus a shit ton of other information (ie, LEOs have access to credit reporting) can be bought by governments from the shady data broker networks already.
If people still ignorantly claim this isn’t a George Orwellian dystopia …
Andrex 1 days ago [-]
> The entire ad tech industry may truly be an intelligence agency front. The lengths they go to track individuals is quite alarming.
Never attribute to malice what can be explained as greed.
The quiet part you rarely hear is that advertising is a smoke and mirror industry akin to throwing darts at a wall. The idea of "homing darts" that stick the target a percentage more of the time would be very attractive in this analogy.
Tech turns advertising from mostly a buckshot spread into something that kinda sounds like something solid and real ("Look, numbers! CTR! CPC! KPI! Our ad product WORKS! Paying customers for your business, guaranteed!")
Ad budgets almost never correlate with actual performance.[0] We're all just lying to ourselves that this business model is performing as expected, and I don't think a reckoning is that far off.
> Never attribute to malice what can be explained as greed.
Why attribute to greed when the effects are indistinguishable from malice?
hagbard_c 1 days ago [-]
...only if and when you allow:
- 3d party cookies
- ads and other 'malcontent'
...which you should never do. As to the 3d party cookies there might be some rare exception where those can be useful but ads? Never, ever allow those on any device you use. Block them as if they're the radioactive plague because they are. Fight them on the beaches, fight them on the landing grounds, fight them in the fields and in the streets, fight them in the hills, never surrender.
That's ads we're fighting. Maybe the same oration will be relevant in the context of ChatGPT and its brethern, we'll see. For now, ads be gone and keep those chatbots at a leash.
hagbard_c 6 hours ago [-]
So, knee-jerk-down-voter, what is it about what I said here which irked you so much that you just had to press that irksome button again? Or is it just because I happen to have said something else sometime earlier which makes you obligated by doctrine to down-vote whatever else I write? Let us know. Do you work in ad-tech? Don't you like the (ab)use of the Churchill quote?
Is it the suggestion that it might be needed in relation to ChatGPT et. al. sometime in the future? Enlighten us, don't just attempt to get a dissenting opinion greyed out. That is for cowards, don't hide behind that button, don't be a craving coward, don't be a sheep. This is, after all, a discussion board, not some online likes competition where you get brownie points for getting rid of dissenters.
...or maybe it is the latter for some, maybe it is a way for some to raise their status among their co-religionists?
Comrade Knee-Jerk, what did you do for the cause today? Answer me!
Somewhere among the crowd a figure emerges, clearly nervous. He tries to speak but starts stammering, stops and tries again. What is he afraid of?
- Oh Great Leader, today I did the work to banish one of the hated dissenters from the internets by voting down his malign words so that no others may be subjected to anything but the Desired Narrative.
Is that all you did, Comrade Knee-Jerk? Is that how you claim your worth for the Great Cause? I am dissapointed, Comrade Knee-Jerk.
- Oh Great Leader, I will do better, I will educate myself, I will do my part to eradicate dissent from the internets for the Great Cause, I p...p....promise!
I do not like to be disappointed again, Comrade Knee-Jerk! I will have to think over your position, whether you are truly committed to the Great Cause. Now hide yourself. Comrade Zlither, what did you do for the cause today? Answer me!
cindyllm 3 hours ago [-]
[dead]
aitoolcrux 21 hours ago [-]
[flagged]
nocturndev55 20 hours ago [-]
[flagged]
startup_zombie_ 1 days ago [-]
[dead]
rvz 1 days ago [-]
Just remember it is the same ex-Meta employees who are now at OpenAI adding Ad features into ChatGPT.
Why? Because they are saving humanity. /s
Facebook tracks you outside of their own website for the same reason, and now ChatGPT does the same for the sake of...Ads.
Regulations won't be set (serious ones, at least) unless there's some risk to those holding power. Which is the opposite in this case: this tracking helps them to take even more control over society and individuals.
jamienk 1 days ago [-]
But politics is when people who feel differently try to do something about it
dualvariable 1 days ago [-]
Any regulation will be geared toward protecting the profits of the AI companies, and not toward consumer protection.
tensor 1 days ago [-]
Maybe it's time you all voted for people who might change that? It's really amazing to me how on the one hand people in the US seem to crow about democracy all the time, yet also just accept as a fact that their government will never actually work to help them.
dualvariable 16 hours ago [-]
I threw money behind getting Bernie nominated over Biden--none of this shit is my fault.
2198276 1 days ago [-]
Q: Hypothetically, if Denmark made a defense treaty with Iran and installed 800,000
Iranian soldiers in Greenland, could it keep the US out?
A: You are describing a fascinating scenario! [produces 100 lines of slop while giving
the login to the FBI]. Should I find a website where you can buy the finest used
AK-47s?
You are absolutely insane giving any of your thoughts, trolls, speculations to a surveillance website under your login.
claaams 1 days ago [-]
I've had this on my brain forever and probably why it's safer for Americans to use Chinese model providers now. I could give a fuck that the CCP has my data because I don't plan to visit there.
alansaber 1 days ago [-]
They were always going to do surveillance. At least now you get a trendy e-commerce site recommendation too.
jfasi 1 days ago [-]
I can’t believe I’m saying this, but the reflexive “ad tracking bad” that I am most savvy tech practitioners reach for might deserve some reconsideration in this case.
The thing about advertising on the web and ad tracking as a practice is that, barring the small matter of ensuring the economic survival of the publisher sites, it is almost always a negative for users. When we consider the marginal benefit of naïve, uninformed-by-surveillance advertising with the present day status quo, we find that in exchange for a complete lack of privacy, we only really receive a marginal improvement in ad quality. Of course, if you (like me) consider all advertising to be a negative on the experience of using the web, it’s an even worse deal.
The standard response given by these companies when they bother giving a response is something to the effect of “we are improving the experience for our users,” which obviously the users would disagree with. However, when it comes to OpenAI, they could build a plausible case for this sort of tracking improving the product. If your models know where your internet habits are, the responses that you get could be tuned for both your interest profile and your actual history of interactions/purchases/internet usage, etc. Imagine a world in which you can opt into this tracking, control the data you provide and how it’s used, clear it out and redact it as you please, and opt in and out of responses that are personalized against it. Reasonable people can disagree, but that might actually be useful.
My prediction, though: that’s not gonna happen. OpenAI will first build out the system to collect click and conversion tracking measurements, then they will turn around to advertisers and say “look at how good our conversion rates are,“ and then they’re going to build an explicit ad platform that enshittifies their chat products.
mitxela 1 days ago [-]
Yeah Google benefits you by tracking you and delivering more relevant ads too. Doesn't make it any less creepy.
kats 1 days ago [-]
Oh, the lifechanging technology isn't completely 100% free? They sell ads just like a bajillion other things?
applfanboysbgon 1 days ago [-]
ChatGPT is not 100% free in the first place, and this is about tracking and selling your data, not the existence of ads themselves.
Some outcomes can be annoying, but the net result is still positive, for the consumers and their data privacy at least.
https://arxiv.org/abs/2411.06862
https://netzpolitik.org/2025/databroker-files-targeting-the-...
Another issue is that controllers generally do not need to change their behavior before the final lawful decision which can take a lot of time to go through the court system, especially if it needs CJEU referral. And once the decision comes in force they can often make small changes and restart the whole process.
Also another issue is that DPAs do not often initiate the investigations themselves (unless breach is involved), they only happen at the request of data subjects and not that many people bother making complaints or follow them up. Just yesterday I had to follow up with 9 page reply to the controller's response to the DPA inquiry.
Additionally ePD and GDPR enforcement is sometimes split between different agencies. In those cases GDPR agency tends to wait for ePD case to be solved before investigating the GDPR aspects, often because the ePD consent validity will affects e.g. GDPR legal basis analysis.
It’s an ongoing fight for sure but it’s some fight at least.
I'd classify mandatory encryption backdoors as an industry crisis rather than an annoyance.
[though you are not wrong that encryption backdoors are a backwards step on individuals rights to privacy]
And while, sure, adtech corpos maybe won't due to the bad PR, what about the Kremlin, criminal organizations, or your insane abusive ex?
> You can't make backdoors only the "good guys" can access.
incomprehensible. I’m sure genuinely so, even.
Encryption backdoors is not a case of extending reach for a government, but an effort to get the capabilities back.
They were able to do that before. They just want to be able to continue now.
Note: No, I don't support the idea of backdoors. On the contrary.
The reason is that they have been trying again and again to do this in various forms, slightly modifying tactics so any opposition to it has to start from scratch.
2015 - https://techcrunch.com/2015/10/29/encryption-rhetoric-untang...
2016 - https://techcrunch.com/2016/08/24/encryption-under-fire-in-e...
2017 - https://www.theguardian.com/technology/2017/jun/19/eu-outlaw...
2020 - https://www.consilium.europa.eu/en/press/press-releases/2020...
2021 - https://www.consilium.europa.eu/en/press/press-releases/2021...
2023 - https://www.wired.com/story/europe-break-encryption-leaked-d...
2025 - https://www.techspot.com/news/107408-europe-proposes-backdoo...
2026 - https://www.euronews.com/next/2026/07/10/chat-control-10-pas...
If you wanted to make the corresponsing strong argument for Chat Control and its ilk, the EU just wants (the juicy part of) what the US already has. The remote root level control of most end user devices is not under EU juristiction. So they naturally want companies operating in the EU to give them one piece of access to the presentation layer, too.
This argument is what one must be prepared for, the encryption itself is less relevant..
Chat control is merely their newest attempt at this, carefully navigating around the reasons it was opposed last year.
They keep trying to legislate encryption backdoors by any means, and once they manage to get their foot into the door every other country will use that as a precedent to also mandate it.
As long as you manage to navigate all the dark patterns and not accidentally give your "informed consent".
So your argument is abusive practices are required for the creation of influential companies, serious technology, and high paying salaries?
But like I said, nice place to visit for the time being.
Yes, because it’s a place where culture matters and where trying to change everything as fast as possible for unknown reason and handing out unlimited power to corps to e.g. level half the country to build DCs is not seen as a good outcome. The specific mindset you complain about is what makes those place nice.
Like the US of A?
Also various stars have introduced more bottle cap regulations. One such state is a hotbed for inflated software developer salaries.
You know, the thing that actually makes a country/society
Putting the needs of business over the needs of the people those businesses are actually supposed to serve is how you end up with America
For example because I literally can’t sleep well above 20°C.
But I don't tend to keep my AC below 24 personally. I don't need my room to be perfect room temper2
I don't know if this is a genetic thing, but as a scandinavian who wears shorts down to like 12C, I am unable to function in >30C humid weather. Anything more labour intensive than drinking sangria on ice is not getting done. This year we had up to 37C in Denmark. If you're able to tolerate heat well, I'm happy for you, but my sleep was ruined for months this summer. Not to mention the people who literally died.
I'm investigating my options for at least cooling down the bedroom at night, because I'm not going through this nonsense again. Plastic straws are banned now and we're sorting garbage in 42 different bins, so I'm thinking that about evens it out, environmentally.
The Amish hang wet bed sheets in front of the window (with a tob under it to recycle the water) not sure how well this works in the Netherlands with 95-99% humidity, it certainly wont raise humidity but evaporation might be shit.
Not every bathroom is fit for it but if there is no AC and things become unbearable I put the shower head upside down so that the ceiling and walls get wet and it rains everywhere. Then I daisy chain fans to "tunnel" the air into the bathroom. The temperature drops like a rock.
Better than a government that fears the word gay
> and how low i set my ac temp leads to an incredibly hostile business environment and stagnant culture.
You are quite a cherrypicker, aren’t you?
How about the regulation to have USB-C as a standard or replaceable batteries.
Or one of my favorites, the ban of ingredients that cause cancer.
BTW ever heard of Frauenhofer IIS or ASML?
They never said anything remotely like that in their comment. Please engage in good faith, don't lie and pretend that others said things they didn't, and don't break the guidelines.
Sure, but plenty of companies want to operate within the EU. While I feel any legislation will only affect their operations within the EU, having the infrastructure to operate under privacy laws there means it'll be easier if/when other jurisdictions follow suit
With no Flock cameras!
Privacy has been dead for some time now. The fact is most business people never figure out how they are exploited. The modern intelligence campaigns just made it economical to hit almost everyone regardless of scale... often under some silly pretense like terrorists wanting our underpants. =3
Which has nothing to do with the stalking of the adtech industry, unless you are suggesting that the EU might sell access to the keys to the likes of OpenAI [though you are not wrong that encryption backdoors are a backwards step on individuals rights to privacy]
Things could be worse. It could be like most of the US where there are both no (or at least far fewer) protections against invasive behaviour of private companies and the government actively trying to backdoor private comms.
It has a lot to do with the practices of the adtech industry. If privacy is protected and upheld vis-a-vis the government (or meta-government in case of the EU), it may be upheld vis-a-vis private corporations and other governments. But if the government likes to be able to spy on its citizens, it will not foster mechanisms, practices and a culture of private communications.
But - I agree that it could be much worse.
Just set up your user agent to work as you wish, isn't that simpler than have unelected, technologic dumb bureaucrats writing laws that are complex and don't solve the issue?
All the EU bureaucrats can do is regulate, that's why it's so difficult to do business in the EU and that's why there is so little innovation over here. If Microsoft and Apple were started in the EU they would've been shut down within a week because you can't operate from a garage.
I guess that every once in a while this might have a positive outcome for consumers, a broken clock is right twice a day.
And, regarding the original topic, we wouldn't have many of these privacy issues if the US had not strategically failed to regulate its obnoxious tech monopolies
The innovation gap is real but it’s not just to do with regulation. It’s a lot to do with concentrated capital networks (eg Silicon Valley). And sure, the tech giants are in the USA but it’s not like there’s no innovation in the EU. Spotify in Sweden. Challenger banks in the U.K. (included as Monzo predates Brexit). And plenty of innovation - they just get bought by US companies. Which is a financial market weakness, not a regulatory one.
https://www.computerworld.com/article/4087347/european-commi...
> The mechanism is standard adtech. What has no precedent is running it on an AI chat product.
As someone who has been well aware of this mechanism for quite some time, I still feel icky anytime I re-read the details of it.
What a time to be alive.
See e.g., https://www.science.org/content/article/ai-chatbots-are-beco...
They had to steal the work of researches solving these open problems and then rewrite their solution. The AI equivalent of fraud.
a) There's zero evidence of them doing so
b) Some models released before the car wash problem was discovered would consistently get it right
c) Hardcoding it is pointless. No one is seriously asking that. It's just a trick question. Hardcoding one trick question won't fix its weakness at other simple trick questions.
d) Since it went viral on the internet, the next time they updated the knowledge cutoff, the LLM would likely be aware of the trick. It will fix itself without the labs doing anything special, even assuming the new models weren't smart enough to naturally figure it out.
It's not an insane assumption that the user isn't dumb and has some other reason to be asking the question other than it being a trick/stupid question (duh, if you want to wash your car you need to drive it to the car wash!). Taking it as some ultimate measure of intelligence simply doesn't make sense to me.
The people going all in on it don't realize the downsides, limitations, or understand how they come off to other people with it all.
The comparison between ELIZA and LLMs is valid you boil it down to "humans evolved for 6-7 million years, had spoken language for 500k years, but have only had something non-human that could generate convincingly novel language well enough to hold a conversation for a few decades".
There's no inherent reason it can't turn out having a non-human generate convincing enough language for conversation isn't a complete evolutionary blindspot the same way the short form feed has pretty much one-shotted society...
"It is difficult to get a man to understand something when his salary depends upon his not understanding it." - Upton Sinclair
It's been an absolute boon to finally build out all of the fun side projects I had always dreamed of, and after showing one off to some people I might even be able to monetize.
On the other hand I acknowledge that other people dont want to embrace LLM driven development for one reason or another, and I respect that. People got into the industry for different reasons , but code was always just a means to an ends for me.
Why is it so black and white?
Your kind is what makes the internet shit, not AI
Well. I can't help but notice how each successive headline reporting how this "scam"/stochastic parrot/"scare quotes intelligence" seems to be solving more and more things that were but a few years ago widely regarded as being indicators of high intelligence.
Being highly convinving is one of the things on that list.
A programming contest has a problem where given N < 10000, do something hard like come up with the number of primes less than N
You can come up with all sorts of algorithms that do intelligent things. But the most effective solution is to use metaprogramming to make a massive switch statement that contains all the answers
Don't misunderstand: I'm happy saying AI models "think" or "have learned a thing", and for in-context learning I'd call them smart even by this definition…
…but also, any living creature that needed as many examples as machine learning currently needs, would starve to death before figuring out how to eat.
While training, machine learning processes (not just LLMs, also applies to e.g. self driving cars), are really really stupid and only make up for this by being really really stupid really really fast.
To what I wrote upthread: the "victories" of humanity over machine keep getting closer, but we have yet to wake up one day in great confusion as we find an entire city is no longer in communication with anyone, nor finding ourselves in a state of utter disbelief when the reports come in that the city stopped communicating because it is entirely gone.
If humans learned like ML systems learn, (biblical) Methuselah would still have been failing the Sally-Anne test on his supposed deathbed at 969 years old, like some of the smaller early LLMs did.
> It also doesn't really matter when "we are trained differently" has no direct bearing on the end result.
The question was to ask for a definition such that AI could still count as "not smart" compared to humans. This fits.
It's also why they're spiky intelligences, which I'm happily using right now to write code for me, but also do not trust in the slightest to identify the weeds in my garden. These submarines sure do swim fast*, but they're also very much disqualified for the Olympics.
* https://en.wikiquote.org/wiki/Edsger_W._Dijkstra#1980s
There's a lot of innate knowledge but all neuroscience demonstrates how incredibly flexible the brain is. Brains constantly learn and rewire.
Here's a few things that I think show how crazy it is AND stress those points
You can convince yourself that we're just organic robots (after all, there's no magic), but you would be a fool to convince yourself we're the ordinary kind.We are constantly learning. You aren't just born with your knowledge and it stays static. We are extremely proficient at metalearning (learning how to learn, few shot learning, zero shot learning [0,1]). Our brains are constantly rewiring, able to heal from traumatic damage.
I could go on and on. Does information pass down through genetics? Of course! But that's far from the whole story.
I'm tired of people trying to make AI sentient by making humans robotic. Stop trying to trivialize everything and be okay not knowing the answer to everything. You're human, you're designed to learn and explore, not sit and argue from an armchair
[0] and I mean these in the original sense. Not in the sense that you train on a billion examples of labeled animals and then congratulate yourself on your ImageNet-1k held out test performance. That's not zero shot, that's just a test set
[1] I can literally make up words and you'll understand them. Or use words in novel ways. That's literally how slang works and how new words come to be. Don't be a walibanut ya glufus. Read some SciFi
Most of the effort of evolution was making cells work at all, and even then it's a bit weird, e.g. no plant or animal produces vitamin B12 and we all get this from some bacteria and archaea.
And evolution is kinda hard to time right: bacteria can reproduce in minutes, humans in decades, but only mutations that survive reproduction can be passed on. This makes it even starker as a difference: bacteria had order of 1e13 generations to become multicellular, while human DNA had about 40,000 generations to cope with fire, 220 generations for evolution to do anything with the invention of the wheel, and one generation to cope with the invention of Minecraft.
The analogy here would be: DNA is to our brains like a VN replicator bootstrapping a computer all the way up to a bare-metal-no-OS untrained model, and perhaps a few crude "hard coded" modules like a smiling-face-detector. It's a lot, but it's also missing a lot. If biology used the models and training processes that are state of the art in ML, it would take around a millennia to talk like a child and still fail the Sally-Anne test, and million years or so to pass a degree.
I'm still going to deny the premise of your argument, becasue I think we should define intelligence in terms of capabilities. If a system can discover a cure for cancer or solve P vs. NP, it doesn't matter how many FLOPs it took to train.
A 1 gigabyte LLM isn't going to impress anyone with what it can do.
About 99% (depends who you ask) of our DNA is shared with our nearest primates. Like us, they can learn to use touch screens, but also like us they won't find touch screens in their natural environment. Dogs can be taught to drive cars (just about), but again, not natural environment.
> I'm still going to deny the premise of your argument, becasue I think we should define intelligence in terms of capabilities. If a system can discover a cure for cancer or solve P vs. NP, it doesn't matter how many FLOPs it took to train.
We can define it in either way. I think both are valid, because plenty of people mean each of these two things when discussing AI in particular. As I referenced in the other branch, these submarines sure can swim fast.
But at the same time, they have a lot of gaps. This is because some experience needs the real world: just as nine women can't make a baby in one month, a transistor running a million times faster than a synapse can't make a month-long cancer experiment happen in 2.6 seconds.
This dependency on data, and that state of the art ML is bad in specifically this way, is why Tesla's self-driving cars, despite having had around a trillion miles of real-world experience today, still come with steering wheels (even at least some of the Cybercabs, despite the big thing of this model supposedly being not needing them, though with Musk and his promises you should only count the Cybercabs when they actually ship and not just press releases).
Imagine an alien that matches your abilities across every domain, but has a 10 billion year training period, something many orders of magnitude more expensive than an LLM. I simply don't believe that alien is less intelligent than you.
We also don't expect humans to be competent in every domain. Most humans suck at most things. We will usually call someone intelligent if they excel at solving problems in one or two narrow domains.
> 10 billion year training period, something many orders of magnitude more expensive than an LLM.
I'm saying both definitions are valid definitions, they both point to important and different things: skill now, vs. how hard it is to get new skills. Some would describe it as "crystallised intelligence vs fluid intelligence".
I think it's important that any arguments are over the thing in dispute, not the label for that thing. Don't mistake the map for the territory.
Anyone who says "AI is stupid" by the first definition, what it can do, I think is making an error: they are already wildly super-human in at least some areas, if not generally.
Anyone who says "AI is stupid" by the second definition, how many examples they need, I agree with: there is a lot they are not currently able to learn even though it is easy for us, because the data they would need to do the learning on does not exist at the scale they need.
Also note: examples, not years. An alien intelligence whose synapses trigger 10 times faster or slower than mine (or ten million times faster or slower than mine), but who gets as much as I do out of each book or conversation, is my equal by the second definition.
I don't think I agree with your characterization of the second definition. Time scales matter. It's not much use to be able to solve human-scale problems if it takes millennia. And it only takes months to train an LLM to the level that it can solve cutting-edge math problems.
Aye, for practical purposes; but this gets you crystallised intelligence. I'd be happy to say e.g. the Chinese Room has crystallised intelligence. But humanity invented fire before reaching the anatomically modern form, and even anatomically modern humans collectively took hundreds of thousands of years to invent durable writing with which the room in the Chinese Room thought experiment could be filled.
It was around a million (or so) years from fire to having enough shared cultural knowledge to be able to formulate the cutting-edge math problems that LLMs can now solve.
Human fluid intelligence means we can pick up deep shards of this accumulation of wisdom, find new avenues of novel research to poke at, all within 40 years, even despite the depth and breadth of work from all the other humans who came before.
(Though this also points at another way to be "superhuman": breadth. Many hands make light work, as the saying goes, and a lot of different humans solving different puzzles at the same time is part of how we got so good so recently even though ~10% of all humans who ever lived are currently still alive; and the same for AI was (accidentally) also part of how the OpenAI-HuggingFace incident went down).
AI (not only, but also, LLMs) are very useful, and I'm getting value from using them. But the fluid intelligence of machine learning* is very poor, and the only way they have to make up for this is by being very fast**, but when there's not enough to train the AI on, they get stuck at a very low plateau.
* possibly the architectures, but I suspect the process by which AI weights and biases are set, and again I don't mean just LLMs
** the speed difference between a transistor and a synapse is about the same as the speed difference between a jogger and continental drift
That seems like a really bizarre way to describe a tool that solved an open Millennium Prize Problem. They are, empirically and repeatedly, ahead of the status quo.
So if your argument depends on them being behind the status quo, reality has already disproven it multiple times over.
I will admit the first time I read the thing you're replying to, I had a similar thought as you; From the sibling reply from them, I think they think they were obvious, but that also means I wouldn't expect their reply to help unless you had the same flash of inspiration I had.
Definitions are "formal statements of the meaning or significance of a word, phrase, idiom, etc" (https://www.dictionary.com/browse/definition)
> They are necessarily behind the status quo.
The existing state or condition would be what is written in the dictionary, not whatever personal definitions you've constructed.
> "You can't call this newfangled contraption a computer, because a computer is a person!"
Seems like a straw man. A computer is not a mammal, no matter how much you twist a set of definitions.
I’m not the person you replied to, but I believe they’re referring to the occupation of “computer”:
https://en.wikipedia.org/wiki/Computer_(occupation)
So yes, at one time all computers were mammals.
The people who write dictionaries generally take a descriptivist approach, that’s why slang terms enter the dictionary after they start to become popular.
The state of the art of human knowledge would be another step ahead of the common use of any language.
E.g. humans get exposed to new LLM model - yeah its powerful - 1 week later - eh, that thing? Yeah it's whatever. I'm still employed.
The human's ability to adapt so efficiently is mind-boggling - so much so it pi1sses sam altman and dario off.
When AI does it we call it “reward hacking” but when humans do it we call them clever.
OK, they can play chess, but that's not real AI - can they write poems? OK, they can write poems, but that's not real AI - can they compose music? OK, they can compose music, but that's not real AI - can they translate languages? OK, they can translate text, but can they do maths? OK, they can do maths, but can they solve a Millenium Prize? <-- we are here
“I once met a person who could beat any grandmaster in chess, translate any language, and complete international math Olympiad problems. He couldn’t solve any Millenium problems though, so I’d say he was a midwit at best.”
"What? Don't be silly. For one thing, trees have moss."
"OK he's grown moss. He's a tree now right? Right??"
"I doubt it, for I see nothing but wishful thinking to suggest that simulating the appearance of tree characteristics is part of a path to becoming a tree. And that's not actually indistinguishable from moss anyway, is it?"
"Urgh, classic goalpost shifting!"
I try to keep an open mind about propaganda fooling me today; the people who were fooled in the past often were not fools themselves.
Being fooled doesn't make you a fool. But being unwilling to change your mind does. Being unable to admit you don't know or don't have enough information to make a strong opinion makes you a fool too.
Propaganda wants to take shortcuts, to simplify things. To trivialize. "It's so easy, you just..." because the fool is the person who already knows, the person who has nothing to learn, the person who thinks they're better than everybody else.
A million YouTubers grinding The Algorithm while secretly sponsored by various world governments, isn't much different to a thousand well-placed gossipers secretly sponsored by various world governments.
I’m reminded that almost no one beyond a select few knew high up in the military and around the emperor knew how badly the Japanese were defeated at Midway.
Paternalistic. Arrogant. Shameful. And deeply engrained in the Japanese cultural zeitgeist (of the early-mid 20th century).
Edit: I guess it’s commonly attributed to a German citizen, but their cultures mirrored each other. Fascism falling under the weight of its own propaganda.
It blows my mind that people don’t care about the world they are building with this stuff. It’s a real tragedy of the commons. People see these collaborators from different wars and regimes and think “I’d stand up against the bad guy”… well I’ve got news for you if you build spyware, you are not the person you think you are.
https://gowers.wordpress.com/2026/09/17/why-i-didnt-sign-the...
"Nothing is worse to the demise of a society, than people who want to convince you that the cat is out of the bag and will not go back in, while the cat is being violently shook out of the bag at the same time."
There's nothing in the bag.
The cat will never get out of the bag.
It wouldn't be a problem if the cat was out of the bag.
We cannot possibly keep the cat in the bag.
Putting the cat back in the bag is not worth trying.
This is where the lawyers find the loophole to get the cat out of the bag.
Seems more like a real tragedy of private enterprise.
We used to assume that the surveillance world would be built by government (1984). But it turned out to be equally likely to be built by the free market.
Facebook didn't invent "talking to friends" or "showing adverts", but made itself "the place" hard enough most of the advertisers and most of the people intermediate through it.
OpenAI didn't invent "asking questions and recieving answers", not even "from an agent who knows which sites to search on your behalf"; but it is competent enough that I might have it read 50 times as many pages in a day as I myself would have read, and the sites' owners don't get real eyeballs looking at ads during this. (In my case, adblock even if I did it manually; but apply this massive increase in page hits to everyone who has their LLM research stuff).
There should be general standards for what individual apps and websites should be allowed to do. There should be an expectation that the purpose of an app is what it does, i.e. a social connection app shouldn’t be an ad platform that suffers users insofar as they provide useful data to sell to advertisers.
Blaming a corporation takes away all the agency the workforce has.
When I ask people about things like this, I hear a lot of "If I don't build it, someone else will"
My goal isn't just to refuse to build this stuff, it is actively to resist the people who are.
I don't have much influence though
I hate stuff like this. Sometimes euphemisms are kind, like "senior citizen" instead of "old person".
But this is an attempt to normalize bad behavior that is really quite terrible for society.
Meta?
When researching a topic a chatbot can be quite sensitive to certain wording and those can end up steering definitions. Two context free chats on the same topic can go in very different ways depending on how you word things, but when using the notebook feature in Gemini, where every chat becomes part of the context you completely lose the ability to discover if a topic is vaguely defined or have many definitions.
I personally like the option to have context free chat or chats with memories. What I dont want is a chat based on memories that I didn't explicitly consent to (ie browsing history etc)
why not use your own words? If you are gonna ai generate this blog, just post the prompts instead.
Most people do not, in fact, want to read raw prompts.
> This blog is where I write about what I find.
Why not simply include a disclaimer that it was AI model output not his own writing? Because humans don’t like AI writing and the article wouldn’t be featured on as many tech news sites. Hence the sin of omission, the choice to mislead.
> Most people do not, in fact, want to read raw prompts.
I wonder if this is really true, and if it will remain true for long. Have you observed someone read another person’s raw prompt? Or observed someone submit both their prompts and their LLM output for review?
Personally I’d be curious about the prompts for a lot of top HN articles which are LLM-generated. It would say a lot more about the human operator’s intent and thinking process compared to the LLM output.
A coworker who spoke English as a second language once screen shared their Claude session during a code review, and it was interesting that their prompts were all in their native language. The guy wasn’t writing an article, and I didn’t understand the prompt anyway, but I still found it interesting as a glimpse into how he uses LLMs.
For this particular article, if the prompt was initially a series of bullet points that wouldn’t be so bad. If there were multiple prompts spent editing and rearranging the article that would be an unusual work by pre-LLM standards. However I’d suggest that instead of dismissing the idea, we could embrace “raw prompts” as a kind of new medium, to cross the divide between pro-LLM and anti-LLM readers.
I also wonder what is the energy consumptiom difference between the prompt and fetching website.
I seriously doubt that.
First, the author documented a substantial effort (testing, emails, references to documentation.)
Second, I also used AI to evaluate if this was written by AI, and my AI said it was not.
> Broken sentences that a model wouldn't produce. "The value is while you are on ChatGPT and tied to your ChatGPT account" is missing a word. "Loading that code, sends __obi to OpenAI" has a comma splitting subject from verb. "By the virtue of loading the tag the identifier is disclosed" is non-idiomatic. LLMs are fluent to a fault; these are the fingerprints of a fast human writer, possibly a non-native English speaker.
Third, I read it. Whilst it is not well written, it is succinct. It is novel. It has a few detailed references. It lacks many of the hallmarks of AI. It makes a number of novel yet falsifiable claims.
I challenge you to come up with a blog post with these qualities that can be created in under 15 minutes via AI.
While I commend you on actually using an AI tool to validate your assumption, rather than simply hurling AI slop accusations based purely on vibes, I do think you and your AI tool are likely incorrect.
This is an artificial constraint that does not serve the reader. Where did the 15 minute constraint come from?
every day I'm more and more reminded that the level of intelligence of the average human even on hn is so low that it makes sense why LLMs took over so hard. Asking an LLM like an oracle as if it could naturally distinguish between LLM writing and human writing.. it's a statistical likelihood next token predictor. LLMs can't even play chess without constantly making illegal moves. They are not intelligent.. and unfortunately, I have bad news for you: neither are you.
Why would anyone want that?
Note to folks who have this same thought (seems like many given other comments). Why are you on this site if you aren't here for human-created content? If you want AI generated blogs/posts there are plenty of sites like linkedin, twitter, chatgpt, and claude that will give you meaningless content with a press of a button.
AI content is typically lower quality, but I'm not opposed to AI-aided content just on the principle of the thing?
At least criticise the article on merit and not some 'let me google that for you' high horse. If it's slop it's slop, but let the votes speak for that.
This sentence is way too much detail and it's literally the first thing you read. I can pick apart most of the sentences in the article. Another one:
> Across 932 decoded sync tokens, 736 carried subject_type: account_user and 196 carried anonymous
This sentence is pointless. What matters to proving that "It works when you are logged out" is to show that the identifier is stable. Why does the reader care about the actual counts.
AI has a tendency to do this which makes AI generated text a lot harder to read. The raw prompts likely don't have the specific websites and cookie names because that's not relevant to the reader or writer for that matter - it's a footnote at best.
Like I say, slop is slop. It's easy to detect. This post and the OP's actual website is basically just a few AI blog posts. I would just treat it as spam and disengage.
> let the votes speak for that.
The comment you are replying to got a fair number of votes too… Having somebody run a Pangram check saves n people the trouble of doing the same.
I think you have to take the output as read and decide accordingly. If it looks like low effort AI writing then it's not much better than spam.
I think the the top-voted post being all about the provenance of the article says something about where HN's voting mentality lies - we are not prioritising the substance of the post but how it was produced.
Firefox, Brave and Safari do. Chrome and Edge do not.
But it would not give ChatGPT information about which other sites are visited.
Of course there would be non-cookie options like fingerprinting (also via IP) that would allow tracking non-the-less.
So you might be talking with OpenAI about your marriage problems and then based on the IP the OpenAI ad network would start showing ads for divorce lawyers on unrelated sites you browse to that display ads.
The solution to that would be using a VPN.
> Google Chrome doesn't block third-party cookies by default, only in Incognito mode, or when users explicitly set it to block third-party cookies via chrome://settings.
Looks like the settings let you block all third-party cookies and add exceptions for specific sites, which seems a bit awkward but could be made to work.
Alternatively, you could run OpenAI in its own profile, or look into what extensions might do.
The fear over blocking third party cookies breaking stuff is severely overstated. I have it disabled by default and I don't think I've ever seen any website breakages. The most is office365 nagging me to click on links so it can authenticate across domains.
https://www.cbsnews.com/news/google-third-party-cookies-chro...
This disgusts me more than any of their recent news. There needs to be a lot more pressure on them to phase this out.
Maybe this article will be the small snowball that gets that started...
And the focus on cookies only is also intentionally misleading. Tracking is not just cookies. Chrome will track you in Incognito mode.
People have very different "expectations" of privacy when they're having a conversation with an AI VS when they're browsing something like Facebook
Not to mention Facebook is free whereas you pay for a GPT subscription
There was never any expectation of privacy if you knew Zuck's history. Facemash almost got him expelled for violating individual privacy.
It's not a coincidence that federally backdoored spyware like Windows is the most popular operating system in the world. I too remember the shocked headlines of the Snowden revelations, and ten-plus years later I work shoulder to shoulder with people that couldn't possibly care less. There is no expectation of privacy, HN too quickly extrapolates it's own virtue signalling to normal people that enjoy using spyware like TikTok, Facebook and Windows.
But also, today, a lot of lawsuits about Facebook knowingly getting people addicted, and governments changing the laws to resist the general category.
But not, as you say, for Windows (and others*) spyware.
* I'm 95% sure the slow-walking of one very easy bug I (and others) reported to Ubuntu was there by government mandate, though obviously I couldn't guess which government.
Both the ISOs and the SHA256 hashes themselves were served insecurely until… my email archive says I reported in 2015 and it was closed in 2022.
The overwhelming majority of users do not pay and use the free service.
Recalling the now-old adage: Facebook is only free if your time and privacy are worthless.
I imagine that people actively trying to make money off stuff like this would have come up with plenty more ideas than just this one.
Modern browsers are more strict towards third party cookies, but there are plenty of ways to work around this.
I think you can easily imagine a future where the subscription model token subsidy ramps down and is replaced by ads, but it’s important to use precise language about the current state of the world.
The fact that others do the same doesn’t make any of the cases excusable.
We use a great chunk of our lifespan to procure the coins and every company is designed to extract as many coins from us as possible.
What happens to this game/real world when this model collapses? We can't procure, they can't extract. What then?
Safari and Firefox should isolate the cookie by default.
It is not a requirement and no one wants to see ads.
That's what firefox's total cookie protection (enabled by default) does.
Doesn't mean you won't be tracked by about 15 other means.
> Something like it is basically a requirement if you want to see digital ads.
So to summarize, yes, it should be illegal.
Remember they once predicted it would be 50% of their income.
This is the only way they get there.
Nobody is going to pay for ChatGPT. They'll just use the ad-infested version, like they do everything else online. Well some people will pay, but not enough to justify the insane amounts of money being poured into it by investors.
It's still an advertisement, and the underlying marketplace is similar (pay for access to change behavior).
The best uses of AI will be surveillance, propaganda, cyberterrorism, and automated military tech.
See, e.g., https://www.nytimes.com/2026/09/18/technology/iran-china-aut...
I think it is more insidious than that. AI shapes how people think not just what they think.
https://apps.apple.com/us/app/ublock-origin-lite/id674534269...
Don’t beat yourself up, uBlock Origin on iOS is only four months old.
according to Bloomberry: https://bloomberry.com/data/chatgpt-ads/
Forget if they have Hagezi though. Also it would be really nice if there was a way to export and import custom URL lists and rules....
https://github.com/hagezi/dns-blocklists
My own setup is DIY: a Debian box running Unbound (recursive DNS) with the RPZ blacklists from above. This gets rid of the upstream DNS service such as the ISP's completely, and prevents tampering or censorship.
Since this has been standard ad tech for awhile, browsers have reacted to this to implement cookie partitioning for exactly these privacy reasons.
This sentence is even more concerning to me. Scraping user input should never be a way of collecting personal data from a website since it can circumvent any website controls without the website’s out the user’s knowledge.
||bzr.openai.com^$third-party ||bzrcdn.openai.com^$third-party
[1] https://storage.ghost.io/c/b8/53/b853e3d4-3186-409d-9c7f-7da...
Disclaimer: I cut it off after a few minutes because I got impatient. It could have gotten closer if I waited longer.
Doesn’t Gemini or whatever Meta’s agent is do this too?
https://support.mozilla.org/en-US/kb/introducing-total-cooki...
Everyone would be far better off if the author just posted "OpenAI uses third party cookies [wikipedia link]", except maybe the author who wouldn't as many subscriptions for his "threat intel" newsletter.
I didn't look at the prose closely enough to check for that, but people wrote inflated explanations of basic things like this all the time pre-LLM. The Wikipedia article doesn't know the specific cookie name, and can't show the results of an experiment verifying that it is in fact being used for ad tracking, or identify specific sites using it in collaboration with OpenAI.
And those got a pass because at least you could defend them with some excuse about how it's some budding author trying to hone their writing skills, or trying to improve their understanding by putting pen to paper. Cases where those excuses don't work (think crappy content marketing pieces from random companies) got short shrift as well. Now for all you know, it's just some dude who prompted claude to "write a blog post about openai's ads".
The audience of people who can read and understand an explanation like this one is probably quite a bit larger than the ones who can replicate the experiment themselves.
Yeah, I could tell few sentences in that it's Claude-written article. The style has become that distinctive. Hell, a third of the articles I opened on the HN front page today carried that distinctive style.
Doesn't really matter if the content is worth it (and I just realized that recognizing Claude in this also makes me wary of ways this could paper over key details - at this point I start to recognize from my own experience where Claude may be papering over something it didn't actually bother to check).
``` ||openai.com^$cookie=__obi ```
Does this look right? AdGuard style.
1. Never register with OpenAI/ChatGPT, and
2. Strongly block ads, e.g. using uBlock Origin + EFF Privacy Badger. Yes, those don't work on Chrome, Edge and other Chromium-based browsers.
Also, even then - ChatGPT may be tracking your behavior indirectly through Microsoft's various services and platforms. But we should do our best to undermine mass surveillance and support individual privacy.
Whereas, YouTube is for personal consumption in almost all cases, doesn't say anything about your employers willingness to invest in software/services
Edit: looks like it's already blocked by Hagezi Multi Pro (and maybe lower levels) and OISD :-D
https://brave.com/
/tinfoil-hat
I think we really need to hold the individuals running ad networks personally liable for the violation of our privacy rights.
I'm really surprised anyone thinks these companies care.
if a single person did that they'd be sent to prison (rip Aaron) but when a too-big-to-fail industry does it with political campaign contributions, no problem?
well firefox+ublock is still an option for those wise enough not to let unknown javascript with new daily zero-days run on their PC
Google simply provided an index on top of an existing library. Of course, a librarian has no value if he has no books to index over! But it's also worth noting that the Google "librarian" also leveraged the existing "social" structure of the internet: their core contribution (page rank) was a clever, efficient mechanism to extract the latent value in the pre-existing link structure of the internet. This structure (much like the pages themselves) had been curated by actual humans. Undoubtedly page rank was clever, but it was worthless without the existing websites (books) and the existing indexing information (the pre-existing, crowdsourced librarian work). Nonetheless, they successfully monetized it.
AI companies are even worse in the sense that initially Google was still sending traffic to the original webpages. (Until they didn't - https://www.eater.com/2017/9/12/16294380/yelp-google-scrapin...). So yes, the AI companies have even more thoroughly stolen the collective work of humanity than Google did.
so they basically have copies already of every webpage until they turned it off a few years ago (well they may still have it updated but not provide it as a service)
so it occurs to me they most definitely trained their "AI" on all that user cache
they may have even just turned it off as a service when they realized other "AI" could do the same thing
This has effectively been Google’s business for decades. Not in the same form, but the concept is the same.
That... does not follow. The information you're getting with this is what sites a user visits. That's creepy and valuable for advertising purposes, but is hardly the type of that that's going to bring about ASI, which is what all the AI labs are working towards. That's why they're hiring data annotators (sometimes with masters or phds) to get training data.
Much of this information plus a shit ton of other information (ie, LEOs have access to credit reporting) can be bought by governments from the shady data broker networks already.
If people still ignorantly claim this isn’t a George Orwellian dystopia …
Never attribute to malice what can be explained as greed.
The quiet part you rarely hear is that advertising is a smoke and mirror industry akin to throwing darts at a wall. The idea of "homing darts" that stick the target a percentage more of the time would be very attractive in this analogy.
Tech turns advertising from mostly a buckshot spread into something that kinda sounds like something solid and real ("Look, numbers! CTR! CPC! KPI! Our ad product WORKS! Paying customers for your business, guaranteed!")
Ad budgets almost never correlate with actual performance.[0] We're all just lying to ourselves that this business model is performing as expected, and I don't think a reckoning is that far off.
0. https://www.forbes.com/sites/augustinefou/2021/01/02/when-bi...
Why attribute to greed when the effects are indistinguishable from malice?
That's ads we're fighting. Maybe the same oration will be relevant in the context of ChatGPT and its brethern, we'll see. For now, ads be gone and keep those chatbots at a leash.
...or maybe it is the latter for some, maybe it is a way for some to raise their status among their co-religionists?
Comrade Knee-Jerk, what did you do for the cause today? Answer me!
Somewhere among the crowd a figure emerges, clearly nervous. He tries to speak but starts stammering, stops and tries again. What is he afraid of?
- Oh Great Leader, today I did the work to banish one of the hated dissenters from the internets by voting down his malign words so that no others may be subjected to anything but the Desired Narrative.
Is that all you did, Comrade Knee-Jerk? Is that how you claim your worth for the Great Cause? I am dissapointed, Comrade Knee-Jerk.
- Oh Great Leader, I will do better, I will educate myself, I will do my part to eradicate dissent from the internets for the Great Cause, I p...p....promise!
I do not like to be disappointed again, Comrade Knee-Jerk! I will have to think over your position, whether you are truly committed to the Great Cause. Now hide yourself. Comrade Zlither, what did you do for the cause today? Answer me!
Why? Because they are saving humanity. /s
Facebook tracks you outside of their own website for the same reason, and now ChatGPT does the same for the sake of...Ads.
I told you so. [0]
[0] https://news.ycombinator.com/item?id=48996936
THIS is where the regulation needs to start.
The thing about advertising on the web and ad tracking as a practice is that, barring the small matter of ensuring the economic survival of the publisher sites, it is almost always a negative for users. When we consider the marginal benefit of naïve, uninformed-by-surveillance advertising with the present day status quo, we find that in exchange for a complete lack of privacy, we only really receive a marginal improvement in ad quality. Of course, if you (like me) consider all advertising to be a negative on the experience of using the web, it’s an even worse deal.
The standard response given by these companies when they bother giving a response is something to the effect of “we are improving the experience for our users,” which obviously the users would disagree with. However, when it comes to OpenAI, they could build a plausible case for this sort of tracking improving the product. If your models know where your internet habits are, the responses that you get could be tuned for both your interest profile and your actual history of interactions/purchases/internet usage, etc. Imagine a world in which you can opt into this tracking, control the data you provide and how it’s used, clear it out and redact it as you please, and opt in and out of responses that are personalized against it. Reasonable people can disagree, but that might actually be useful.
My prediction, though: that’s not gonna happen. OpenAI will first build out the system to collect click and conversion tracking measurements, then they will turn around to advertisers and say “look at how good our conversion rates are,“ and then they’re going to build an explicit ad platform that enshittifies their chat products.