Otherwise, I think this would still go into my "I don't get it" folder. It seems like the flavor of the day is "agents can talk to each other" and Grok Bot is selling that point along with Claude Code. But I think Pi has had this for a while? Or it was easy to wire it in.
Grok Bot
321–330 of 362 posts
Re: Grok Bot
#322How do they stop providers (like Amazon, etc.) from detecting and blocking these agents if they are running on cloud? I know that openai wasn't able to avoid this which is why they moved to 'computer use' on your local machine. I almost wonder if this is a place where SpaceX, as an internet provider through starlink, has a unique advantage because websites are unable to block their networking as it could be residenti…
I use Hermes locally and it’s constantly hitting bot blocks. Just trying to shop clothes for me it gets blocked.
Re: Grok Bot
#323Re: Grok Bot
#324Re: Grok Bot
#325Re: Grok Bot
#326Earlier quoted context omitted.
According to Boris Cherny from Anthropic [1], the threat of prompt injection has been largely solved. [1]: https://x.com/bcherny/status/2086520950259118464
As someone who just got out of a meeting demonstrating how Copilot running Luna can be breadcrumbed by a single line of text in innocuous package into downloading and installing malware I think Boris Cherny might be wrong.
Re: Grok Bot
#327I feel like we'd be better off if we just stopped at chatbots...why are we so eager to make the internet even more botted
The internet sucks. Yesterday, I had my agent search for openings for The Odyssey that fit my requirements and then book them. It was way better than manually looking at seat maps for 20 different showtimes and going through 10 steps just to buy the tickets.
Then I can just pick which time works for me. I have never in my life had a theater be sold out when it's not an opening night or weekend for extremely big films.
Is this something you actually tried to approach before LLMs?
Re: Grok Bot
#328Earlier quoted context omitted.
The internet sucks. Yesterday, I had my agent search for openings for The Odyssey that fit my requirements and then book them. It was way better than manually looking at seat maps for 20 different showtimes and going through 10 steps just to buy the tickets.
This is functionality currently built into Google search. I google "The Odyssey Showtimes" and it shows me when it's playing at all local theaters. Then I can just pick which time works for me. I have never in my life had a theater be sold out when it's not an opening night or weekend for extremely big films. Is this something you actually tried to approach before LLMs?
Re: Grok Bot
#329Earlier quoted context omitted.
Besides technological progress which has nothing to do with this effect and would have happened anyway, what's the value that we didn't have 30 years ago?
It’s so incredible to me that now we have a chat interface we can ask about anything in any language and get really great answers, something literally considered science fiction a few years ago, and people still act like that is no big deal at all. No value in something like that! Don’t tell me it’s inaccurate, I strongly believe it’s way more accurate than if you could ask an expert in each topic , which of course y…
Because the importance of this is all about perspective. It wasn't like these systems created this information out of thin air. They were trained on something. That means the answers they are giving you have been available for decades. You just needed the know-how to find that information and synthesize the answers yourself. To many of us, it's like going from the old physical card catalogs to a modern digital system that would have seemed like sci-fi to a prior generation too. It's definitely more efficient and easier to use, but people acting like it's revolutionary seem to be suggesting that the old system didn't exist or wasn't usable with a little effort.
Re: Grok Bot
#330Earlier quoted context omitted.
As someone who just got out of a meeting demonstrating how Copilot running Luna can be breadcrumbed by a single line of text in innocuous package into downloading and installing malware I think Boris Cherny might be wrong.
Boris is talking about Anthropic models only of course. In the chart in his tweet Luna has like a 43% chance of being compromised with sufficient attempts. But even the Anthropic models have as high as a 5% chance, which is still far too much IMO.
But you’re right…5 percent is better than what it was but at scale that’s still quite a bit.