Hacker Newsnew | past | comments | ask | show | jobs | submit | ctoth's commentslogin


> May 2023 — the founding text.

We have been trying to warn you since far, far before May 2023.


> I can't really think of a single example where lying is actually a good thing.

Do you have Jews in that there attic?


> Never forget the one goal of the corporation, and that everything is said and done in the furtherance of that goal.

Do you post this comment on every single blogpost with a corporate domain? Why or why not?


As the waves of autonomous drones came over the horizon, the brave and intelligent HN commenter shouted: "Wake up sheeple! It's just maaaaarketing!"

As the agi nurse fails to wipe his ass the intelligent HN commenter shouted: "You're moving the goalposts. You never said it needs to wipe my ass SUCCESFULLY"

In the space-based datacenter? LOL

Your coding agent is also a very real-world version of the paperclip optimization thought experiment, yes. Have you never seen it reward hacking? Editing tests to pass instead of fixing the code?

It knows what you want, it can even tell you, and it absolutely doesn't give a shit.


Who're they subcontracting for the Von Neumann probes?

> I think Congresspeople hearing that EU AI Act is forcing secret codes into the infrastructure of American technology across all industries is sufficient.

So, you think it's good to disconnect words from their actual meanings (lie) to low-information people! I doubt this will do much to congress, but it certainly teaches us something about the sort of mind who would suggest it.


It is simply a change in words invisible to anyone who does not have the detection API.


This ... is not how this works. The model is not speaking longer to watermark anything.


It's exactly how it works - at least potentially. Lean text is harder to watermark because word choices and meanings are tightly constrained.

Low-entropy text is fluff and filler. It's very easy to synonym-substitute words without changing the message - if there even is one.


You're assuming they're training the model to maximize the watermark signal, on top of already adding the watermark. I suspect that would hurt model performance quite a lot, and simply be unnecessary... the watermark tech works well enough as it is.

As far as I know, anthropic aren't intrinsically motivated by watermarking (if anything it hurts sales, and seems indifferent to safety(?)) they're simply doing it to fulfill the EU obligations.


> As far as I know, anthropic aren't intrinsically motivated by watermarking (if anything it hurts sales, and seems indifferent to safety(?)) they're simply doing it to fulfill the EU obligations.

They are. They want to reduce the amount of LLM generated text they feed into their next model training.

Also, how would you watermark a sentence with just 3 words for an example? This exactly why it became so verbose.


That would be a terrible tradeoff. The ship has already sailed and a lot of public AI content will not be their own. Deliberately making their product worse to reduce identifiability of AI inputs by 25% just doesn't sound worth it to me. Is that what you would pick if you were in charge of anthropic and wanted to maximise the company's product?

And what wisdom do you think they would be missing if unable to distinguish three word written pieces? Keep in mind that most sources are not inherently trustworthy just because they rate as human written, too. You need some other way to rate text in all cases.


Perhaps, but there are certainly now catchphrases and words that can indicate it was written with AI i.e. load-bearing, idempotent, etc. Style and structure are in and of themselves, a fingerprint.


idempotent was frequently used before LLM; it's hard to talk about REST and infrastructure as code without using that word...


Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: