Live data from Hacker News

Potential session/cache leakage between workspace instances or consumer accounts

github.com

61–70 of 151 posts

Re: Potential session/cache leakage between workspace instances or consumer accounts

#61
post #2

Sounds like a hallucination unless proven otherwise, even the leading LLMs can do those from time to time, and they will always appear plausible like that. Also could be the session having a lot previous context, like 800K+, which (I think) makes hallucinations more likely. Relevant comment from the OP which makes a hallucination more likely: > There is one tool call result that includes a string that printed a pathn…

Exactly. If you've never had an LLM (all models) suddenly start spouting nonsense in a completely different language...you haven't been using LLMs that much. They will go absolutely insane some % of the time.

Worth looking at https://www.anthropic.com/engineering/a-postmortem-of-three-...

They can “go insane” but it seems often to be infra related as opposed to anything one would consider hallucination. Smaller models will often get stuck repeating a word or phrase forever but that’s a bit different and nobody would call it hallucination.

Re: Potential session/cache leakage between workspace instances or consumer accounts

#62
post #38

Just add a line in AGENTS.md that says "never talk about Minecraft unless you're explicitly asked" , I'm sure it'll be fine after that.

CLAUDE.md, Anthropic is too exclusive and next level to use a standard idiomatic pattern like AGENTS.md

echo “read @AGENTS.md” > CLAUDE.md

Re: Potential session/cache leakage between workspace instances or consumer accounts

#63

Using a throwaway account for obvious reasons, but I’m very involved in this space using LLMs from multiple providers. I’m aware of at least two instances in which the intermediate infrastructure “swapped” responses, once impacting Claude models and once impacting GPT models, from two different providers. One gave us a proper postmortem in which their API gateway was incorrectly handling HTTP 100 status codes, puttin…

[flagged]

Curious why you feel that way about Dario?

Re: Potential session/cache leakage between workspace instances or consumer accounts

#64

Earlier quoted context omitted.

[flagged]

Curious why you feel that way about Dario?

HN thinks the safety crowd is dumb, and has never seriously engaged with the AI safety space.

HN doesn't believe superintelligence will be a thing; while the AI safety crowd believes they are building it. So the decisionmaking of the safety crowd is incomprehensible to HN.

Re: Potential session/cache leakage between workspace instances or consumer accounts

#66
post #36
post #2

Sounds like a hallucination unless proven otherwise, even the leading LLMs can do those from time to time, and they will always appear plausible like that. Also could be the session having a lot previous context, like 800K+, which (I think) makes hallucinations more likely. Relevant comment from the OP which makes a hallucination more likely: > There is one tool call result that includes a string that printed a pathn…

I realize hallucination has no precise definition but this doesn’t sound at all like anything I’ve ever heard called hallucination. Hallucination is usually plausible wrong answers or made up info that ends up fitting the most likely response (like a manufactured citation) and comes from the way LLMs work at predicting tokens. This example demonstrates completely implausible output, it’s not something that fits with…

One of his tool results mentioned the word minecraft.py, and the response was about Minecraft.

It's a hallucination.

Re: Potential session/cache leakage between workspace instances or consumer accounts

#67

Earlier quoted context omitted.

Curious why you feel that way about Dario?

HN thinks the safety crowd is dumb, and has never seriously engaged with the AI safety space. HN doesn't believe superintelligence will be a thing; while the AI safety crowd believes they are building it. So the decisionmaking of the safety crowd is incomprehensible to HN.

Reductionist. Many of us think they’re all dumb.

Re: Potential session/cache leakage between workspace instances or consumer accounts

#68
post #2

Sounds like a hallucination unless proven otherwise, even the leading LLMs can do those from time to time, and they will always appear plausible like that. Also could be the session having a lot previous context, like 800K+, which (I think) makes hallucinations more likely. Relevant comment from the OP which makes a hallucination more likely: > There is one tool call result that includes a string that printed a pathn…

[dead]

Re: Potential session/cache leakage between workspace instances or consumer accounts

#69
post #37

In order Fable 5 has rejected: "Recipe for red-braised pork, I have pork shoulder" "Write up a framework for MCP patterns I can give to claude code" "explain the biomechanics of motion in c. elegans" (I get this one, I mostly did it to test and it's related to my hobby project) Do we get an extra day of functional Fable 5 because it's down?

Not sure the relevance of this comment, but normally if someone built a classifier that bad they’d be fired. Anthropic obviously thinks they have some monopoly power they can use to foist garbage on consumers, I think they don’t.

If people are complaining about Anthropic (on an only-vaguely related thread) rather than simply switching to a suitable competitor, then Anthropic clearly has some 'monopoly' power over the specific capabilities the complainer wants from them.

Re: Potential session/cache leakage between workspace instances or consumer accounts

#70

Earlier quoted context omitted.

CLAUDE.md, Anthropic is too exclusive and next level to use a standard idiomatic pattern like AGENTS.md

echo “read @AGENTS.md” > CLAUDE.md

Yep that should work 100% of the time.
Post reply on HN