Live data from Hacker News

Hardening Firefox with Anthropic's Red Team

anthropic.com

91–100 of 188 posts

Re: Hardening Firefox with Anthropic's Red Team

#91
post #26

That's one good use of LLMs: fuzzy testing / attack.

Not contradicting this (I am sure it's true), but why is using an LLM for this qualitatively better than using an actual fuzzer?

It's not really bad or not though. It's a more directed than the rest fuzzer. While being able to craft a payload that trigger flaw in deep flow path. It could also miss some obvious pattern that normal people don't think it will have problem (this is what most fuzzer currently tests)

Re: Hardening Firefox with Anthropic's Red Team

#92

Earlier quoted context omitted.

> what are you hoping to get out of this comment? Rando here. It gives a signal on the account’s other comments, as well as the value of the original comment (as a hypothesis, albeit a wrong one, versus blind raging).

> "It gives a signal on the account's other comments," fair enough. i typically use karma as a rough proxy for that, especially when the user has a lot of it (like, in this case, where the poster is #17 on the leaderboard with 100,000+ karma). you dont get that much karma if you are consistently posting bad takes. > as well as the value of the original comment (as a hypothesis, albeit a wrong one, versus blind raging…

> you dont get that much karma if you are consistently posting bad takes.

I wonder how true that is. While this site doesn't have incentivize engagement-maximizing behaviour (posting ragebait) like some other sites do, I would imagine that simply posting more is the best way to accrue karma long-term.

Re: Hardening Firefox with Anthropic's Red Team

#93

Anthropic feels like they are flailing around constantly trying to find something to do. A C compiler that didn't work, a browser that didn't work, and now solving bugs in Firefox.

However, the shape is there. And no one knows how good the thing is going to be after X months. We are measuring months here, not even years. I believe there is a theoretical cap about the capability of LLM. I'm wondering what does it look like.

If it explore all these cases after a few month and made the tool itself obsolete, that sounds like a total win to me?

However that don't happen unless firefox just stop developing though. New code comes with new bug, and there must be some people or some tool to find it out.

Re: Hardening Firefox with Anthropic's Red Team

#95
At this point about 80% of my interaction with AI has been reacting to an AI code review tool. For better or worse it reviews all code moves and indentions which means all the architecture work I’m doing is kicking asbestos dust everywhere. It’s harping on a dozen misfeatures that look like bugs, but some needed either tickets or documentation and that’s been handled now. It’s also found about half a dozen bugs I didn’t notice, in part because the tests were written by an optimist, and I mean that as a dig.

That’s a different kind of productivity but equally valuable.

Re: Hardening Firefox with Anthropic's Red Team

#96
post #68

Earlier quoted context omitted.

An LLM by any other name would hallucinate the same

Anyone still reading down here will appreciate this https://bsky.app/profile/simeonthefool.bsky.social/post/3kbk...

Hang on, someone downvoted me for a horrific pun? GOOD.

Re: Hardening Firefox with Anthropic's Red Team

#97

Earlier quoted context omitted.

> "It gives a signal on the account's other comments," fair enough. i typically use karma as a rough proxy for that, especially when the user has a lot of it (like, in this case, where the poster is #17 on the leaderboard with 100,000+ karma). you dont get that much karma if you are consistently posting bad takes. > as well as the value of the original comment (as a hypothesis, albeit a wrong one, versus blind raging…

> you dont get that much karma if you are consistently posting bad takes. I wonder how true that is. While this site doesn't have incentivize engagement-maximizing behaviour (posting ragebait) like some other sites do, I would imagine that simply posting more is the best way to accrue karma long-term.

>I would imagine that simply posting more is the best way to accrue karma long-term.

i definitely agree, which is why i use it as a rough proxy rather than ground truth, but i have my doubts that you can casually "post more" your way into the top 20 karma users of all time.

Re: Hardening Firefox with Anthropic's Red Team

#99
post #85
post #55

Earlier quoted context omitted.

And now that you know that it isn't, do you feel differently about the logic you used to write this comment?

Do I?

I don't know. I'm really asking. I have you bucketed in my head in the cohort of "HN commenters who write lots of assembly", so the mismatch between your prediction and the outcome is just really interesting to me.

Re: Hardening Firefox with Anthropic's Red Team

#100

Earlier quoted context omitted.

> what are you hoping to get out of this comment? Rando here. It gives a signal on the account’s other comments, as well as the value of the original comment (as a hypothesis, albeit a wrong one, versus blind raging).

> "It gives a signal on the account's other comments," fair enough. i typically use karma as a rough proxy for that, especially when the user has a lot of it (like, in this case, where the poster is #17 on the leaderboard with 100,000+ karma). you dont get that much karma if you are consistently posting bad takes. > as well as the value of the original comment (as a hypothesis, albeit a wrong one, versus blind raging…

I think a lot of people are overreading this and really all that's happened here is that I was out at a show last night and was really foggy when I woke up and asked a question clumsily. It happens!
Post reply on HN