Hacker Newsnew | past | comments | ask | show | jobs | submit | warpech's commentslogin

I wonder what’s more valuable in our prompts: the raw data or the feedback system that drives the exchange towards a goal.

For a long time it was clearly the former, but now I think it is the latter.

The models have enough knowledge (orders of magnitude more than a human could ever learn) but are now getting better at what to do with it thanks to learning from the decisions that we make in conversations with AI agents.


I think so too. The value is in the entire conversation. IMO, "domain experts" don't run LLMs blindly and hands free. This does not work for top level work (e.g., mathematical proofs, coding anything more complex than yet another slop game or website). Experts have long sessions where they prompt and guide LLM in response to what it produces. This is the discovery process. And frontier labs definitely train on that.

The billion dollar question is whether this works "out of the distribution". I.e., whether LLMs can only find and use the specific ideas buried in training data, or whether they can learn to apply the "thinking process" to a new problem. IMO this is still unanswered (due to these recent controversies).

But regardless of the answer, it seems we have a planet-scale positive feedback loop here. LLM became good (enough) by training on generally available data (books, internet, github) + RLFH, so experts tried to use them on hard tasks, which required lots of hand holding. These conversations became part of the training data, and the next generation of frontier LLMs were better. So, more experts used them on harder tasks, again requiring hand holding. These conversation became part of the training data... etc.

In a nutshell, top human minds across the world are pouring their skills into LLMs just by using them. This is not "continuous learning", but if you re-train on the most recent sessions every, say, quarter (which seems to be happening?) you get close to that in practice.


Last year we were saying there must be a human-in-the-loop (HitL), but anyone who is the HitL exhibits the “HitL skill” to the agent.

There might be no books about human intuition but we teach it to LLMs by interacting with them


I referred to llm’s as mechanised intuition about a year ago.

I don’t know why but it just ‘sounds right’. It’s the best analogy I can think of.


10000000% Correct.

I’ve been working on a novel project for 1 year.

I now no longer use llm’s - the continual chatter I’ve had has resulted in my insights being found in the training data now.

Get stuffed OAI.

Every large firm will soon enough want its own on-prem servers eventually. Maybe nation’s will get involved and build out their own data centres.

Not a chance in hell I’d trust a tech firm to treat my IP as safe and sound - only a sovereign can ‘promise’ that.


What’s a good way to drive models to use current versions?

Thank you for starting the blog post with reminding what the project is.

So many blogs (and newsletters for that sake) assume the reader knows and remembers what they are about


Umm, no? The parent means Wayback Machine, which is a service of Internet Archive (both at archive.org). Or am I confused?

The archive that bypasses paywalls is archive.today (or archive.is, archive.ph, ...) which is run by some shady person, not a professional organization like Wayback Machine.

What is third party? If I self-host OpenClaw, is that third party?

If I write my own harness, is that third party?

Insane policy

Why bother creating a CLI and OAuth interface if your ToS punishes for using it


It comes down to whether the OAuth token generated by their client is used only in their client.

If you register a new OAuth client and generate access tokens for it to call the API, then you are following the rules.


Wouldn’t that violate the quoted ToS? I guess it would.

"Using the service in connection with products not provided by us." The Agy CLI is provided by them.

And if I am not mistaken, Theo’s T3Code (which was explicitly told was bad), doesn’t even use anything else than AI producer provider command line tools and use them with `-p`.

So? There's still a connection to products not provided by them.

Remember that this is a TOS, so you gotta presume that they will act as malicious as possible within the text.


That line of reasoning has no end. If you use Antigravity on anything other than a Google Chromebook or Pixel, the hardware is a 'product not provided by them'. Is that a TOS violation?

Yes. I'd avoid using services like this.

Funny that you can buy tickets for this train as a part of the countrywide train network system, but they don’t have a right graphic for this kind of train. The website displays an electric unit with no information that this will be a steam engine: https://koleo.pl/connection/c21a973b-4eaa-55bd-80d0-5dd2541a...

I was expecting a new HTTP status code


That is true everywhere. I think operating systems could do a better job at informing about the health of the internet connection in words the user understands. What seems to be the current bottleneck:

Your router does not respond. Turn it off and on.

This comes close: https://www.getapp.com/all-software/a/breakdown/


Honorable mention - https://postgrest.org/


I find it odd that this isn't mentioned anywhere in an article about using postgres for everything.


Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: