Live data from Hacker News

Grok Bot

x.ai

51–60 of 362 posts

Re: Grok Bot

#51
post #13

Interesting. Unfortunately Musk's personal brand is so poisonous that I would never let him anywhere near my data. I'm curious if big businesses will have similar concerns and avoid tools from SpaceXAI regardless of how they are? I guess lots are already using it by default since the Cursor acquisition.

[flagged]

Yes.

Re: Grok Bot

#53
post #49
post #20

Earlier quoted context omitted.

Prompt injection is my biggest fear. Imho it is almost impossible to make a sort of tool that would be successful in detecting an injection - but maybe some antivirus/antimalware producers work on it… The rest can be solved - separate account for agent spending (soon offered by your bank or neobank), separate email for agents (already exist) etc.

According to Boris Cherny from Anthropic [1], the threat of prompt injection has been largely solved. [1]: https://x.com/bcherny/status/2086520950259118464

"largely solved" as in they the models they trained don't fall for prompt injections as often but not "largely solved" as in the underlying issue is solved at all.

Re: Grok Bot

#54

I think perhaps I won't trust anything that ever gets released by this company, likely in perpetuity.

Anecdotally, same; recently I stopped using Cursor after learning that XAI now owns it.

Re: Grok Bot

#56
post #49
post #20

Earlier quoted context omitted.

Prompt injection is my biggest fear. Imho it is almost impossible to make a sort of tool that would be successful in detecting an injection - but maybe some antivirus/antimalware producers work on it… The rest can be solved - separate account for agent spending (soon offered by your bank or neobank), separate email for agents (already exist) etc.

According to Boris Cherny from Anthropic [1], the threat of prompt injection has been largely solved. [1]: https://x.com/bcherny/status/2086520950259118464

Sounds like “according John McAfee the threat of malware has been largely solved”.

Edit: it's way worse than that: the actual figures says that Mythos “only” falls for prompt injection 2.6% of the times. Maybe for an ML engineer used to work with unreliable tools that sounds impressive, but for security purpose having a system that fails every 40 attempts is outright catastrophic. Imagine if your OS vendor only patched known security holes in a way that still let an attacker go through every 40 attempts…

And we're just talking about known kind of prompt injection that are tracked by benchmarks, not any kind of 0-day vulnerability found by clever attackers.

Re: Grok Bot

#59
post #53
post #49

Earlier quoted context omitted.

According to Boris Cherny from Anthropic [1], the threat of prompt injection has been largely solved. [1]: https://x.com/bcherny/status/2086520950259118464

"largely solved" as in they the models they trained don't fall for prompt injections as often but not "largely solved" as in the underlying issue is solved at all.

Largely solved in “it only happens 2% of the times now”.

Don't worry, you only have a 2% chance of having you bank account drained any time an attacker tries their chance.

Post reply on HN