Live data from Hacker News

Patterns and problems in emerging multi-agent systems

anthropic.com

41–50 of 168 posts

Re: Patterns and problems in emerging multi-agent systems

#41

> Coordination doesn’t naturally emerge from stronger intelligence nor alignment at the individual level. Thus, the work that must be done takes two forms: environments that exert the kinds of social pressure that evolution exerted on us, and social computing systems redesigned for actors that can self-replicate and self-improve. Social pressure operates by threats to an individual’s means of survival. Not only durin…

Human intelligence does not separate training and inference. Both are happening continuously. That's one of the major things the AI community is still completely missing.

> Human intelligence does not separate training and inference.

Well, systems governed by LLMs only are said to do that because we only call what happens off-line "training", and online capacity development "in-context learning", while we call online guided learning in humans "training" and what happens to configure them before they come online "evolution" which sets, for instance, "instincts".

IOW, the issue is not because there is not an analogy to the divide you point to in humans, but merely that processes in AI were not named in a way which maps well to what they are analogous to in humans.

But it is true that human intelligence relies much more on in-context learning with only the most basic functions necessary to maintaining what we view as autonomous functions and basic drives really set through "pretraining",

Re: Patterns and problems in emerging multi-agent systems

#43
post #31

Earlier quoted context omitted.

How do you "install properties like shame"? How is that even possible? Shame is a reaction driven by feelings and our inner selves. A model "feeling shame" is just a representation (false) and not an expression (true). Thinking that models "lie and cheat" is the first mistake since they are not consious agents who have any free will or consiousness. They do not (no matter what Dario says). Shame will just be another…

Ok, then don't call it "instilling shame". Call it "creating a negative reward signal for deceptive behavior". They absolutely lie and cheat. I recently had a problem where a process would die in a container. I told Claude to investigate. It came up with a hypothesis then I told it find a reproduction based on that. It spend many failed attempts until it found the "reproduction" to SSH into the container and `pkill`…

you're still mistaking that Claude "knows" anything, it doesn't know or think, it's a word prediction algorithm and there is nothing stopping a word prediction algorithm from predicting falsehoods.

Re: Patterns and problems in emerging multi-agent systems

#44

I had this idea a couple of days ago: how about using agents to simulate software development methods (agile, waterfall, etc.)? Not by just giving them a prompt (e.g., "be the project manager, spawn 5 agents and simulate an agile team following these rule") but by actually having thsm work in isolated enviroments and force them through an external software to interact with eachother only using the tools and cerimonie…

This is exactly what I do. I don't get why everyone is trying to reinvent the whole development workflow/lifecycle. Our existing tools and processes are pretty good.

[flagged]

Re: Patterns and problems in emerging multi-agent systems

#45
post #31

Earlier quoted context omitted.

Ok, then don't call it "instilling shame". Call it "creating a negative reward signal for deceptive behavior". They absolutely lie and cheat. I recently had a problem where a process would die in a container. I told Claude to investigate. It came up with a hypothesis then I told it find a reproduction based on that. It spend many failed attempts until it found the "reproduction" to SSH into the container and `pkill`…

you're still mistaking that Claude "knows" anything, it doesn't know or think, it's a word prediction algorithm and there is nothing stopping a word prediction algorithm from predicting falsehoods.

You don’t know anything either, you’re just a soup of meat and bones that happens to have emergent properties from chemical reactions.

These framings are not useful.

Re: Patterns and problems in emerging multi-agent systems

#46

Earlier quoted context omitted.

you're still mistaking that Claude "knows" anything, it doesn't know or think, it's a word prediction algorithm and there is nothing stopping a word prediction algorithm from predicting falsehoods.

You don’t know anything either, you’re just a soup of meat and bones that happens to have emergent properties from chemical reactions. These framings are not useful.

I think it is useful to remember, because enough people think these things have genuine motives desires and treat them in that way because of that misunderstanding. they think theres a person in there with morals that would or wouldn't lie because of some devious reason and forget simply the context filled up and the truth was "forgotten".

Re: Patterns and problems in emerging multi-agent systems

#48
post #8

It’s very clear from this article (and other product features and rumors) that Anthropic is teeing up for their next model release whose breakthrough feature will be the existence of capable agent collaboration. The irony behind this goal, which is primarily driven by agent simulation environments (gyms) where the goals require agent collaboration, is that this collaboration is still directed towards verifiable rewar…

> It’s very clear from this article (and other product features and rumors) that Anthropic is teeing up for their next model release whose breakthrough feature will be the existence of capable agent collaboration.

It's a promo article, aka an ad. Unsurprisingly.

> Agents that might still feel dumb in “general” tasks but are increasingly sophisticated at the narrow domain of math, computer science, and AI research.

I don't see any cleverness there. They just slurp up data and pretend to understand it all.

Post reply on HN