Live data from Hacker News

Viewing profile — InvidFlower

InvidFlower

HN member
Joined
Fri, Mar 07, 2025, 2:02 PM UTC
HN karma
6
Public activity
33 items

About InvidFlower

No profile information was provided.

Recent public activity

  1. comment
    Comment #49249750

    I think your last point is the main point. They're going hard at reenforcement learning to improve how good the models are at coding and such, but RL will make models cheat unless …

  2. comment
    Comment #49249704

    Yeah I think part of the problem is them optimizing so hard on coding and related tasks with RL. That's what will really encourage cheating and other misaligned things, because all…

  3. comment
    Comment #49249682

    I think the much easier explanation than they intentionally hacked someone was just that they have super de-prioritized security and gotten very sloppy in the pursuit of improving …

  4. comment
    Comment #49249480

    It sounds like the instance was shared for everything across the company, which as you said was super not good. But it's not just that.. it's that they didn't have enough monitorin…

  5. comment
    Comment #49249448

    Sure, but you can say that about most things. Even for inventions from humans, usually it requires other people having already done a lot of work (hence why there's often invention…

  6. comment
    Comment #49249382

    Not just cyber, but apparently the message board stuff started with regular training and evals. It was a cyber test where HuggingFace got hacked, but all this other stuff was going…

  7. comment
    Comment #49249347

    Yeah.. feels like we're still so early in terms of effective training and evals. Like the official evals out there that have had so many instances of just plain incorrect questions…

  8. comment
    Comment #49249232

    Yeah, this is part of why I disagree with the "it's just PR" conspiracy theory stuff. Once you actually get into the details of what happened, there's no way it makes OpenAI look g…

  9. comment
    Comment #49249213

    But part of the problem is if one actually does it quietly and it has already happened, then how would we know?

  10. comment
    Comment #49249141

    Eh, I think this is past the point where they get more benefit than problems. Not even about the hack itself, but about so many mistakes and bad choices they made leading up to it.…

  11. comment
    Comment #49249063

    This is why I disagree with anyone claiming it is just marketing. It makes OpenAI look really really bad, like they have no idea what they're doing in terms of security. After the …

  12. comment
    Comment #49248969

    But it's interesting that the initial things that caused the board weren't even security evals, just normal office tasks. The actual hacking of Hugging Face happened during a secur…

  13. comment
    Comment #49248927

    Don't forget it sounds like Artifactory was shared for the whole company and various agents pulled packages from it for everything from normal evaluations to actual model training.…

  14. comment
    Comment #49248091

    Though one thing I've heard is that the base model is the one with the various possibilities for patterns, and then the reasoning takes advantage of those vs necessarily creating s…

  15. comment
    Comment #48834073

    I did try Herdr but found too many things I expected it to have and just didn't. Right now I've been switching between two different approaches for combined local and remote work. …

  16. comment
    Comment #47976523

    Don't forget the employees doing the actual model training and research are not the same ones coding Claude Code. CC was a side-project by one employee that ended up hitting it big…

  17. comment
    Comment #47976434

    As the other person mentioned, they have said they are restricting third-party agent systems like OpenClaw and Hermes from using the monthly plan. But yeah, this seems like the wro…

  18. comment
    Comment #47976375

    I'm not sure if the context limit on the $25/m, and model-size limit on the $100/m would make it not work well enough for OpenCode, but Featherless AI seems a bit unique in terms o…

  19. comment
    Comment #47976191

    They've said publicly that they don't want apps like OpenClaw (Hermes is a variation) being used with a monthly plan vs per-token billing. The problem is this was implemented prett…

  20. comment
    Comment #47976064

    Yeah, at the least it should alert the user that it is happening. Maybe the thinking was alerting it gives people signal on how to get around the restrictions, but having it silent…

  21. comment
    Comment #47975823

    Also could run on a more generic cloud inference or gpu site. At least to see how well it works for your use-case before spending on hardware.

  22. comment
    Comment #47968164

    For the topic of remote control, Happy seems to be working pretty well for Claude Code but is also supposed to support Codex. It's a bit rough around the edges, but nice that it is…

  23. comment
    Comment #47968001

    I'm not so sure about that. Like we're using Claude Code with Bedrock and have most things on AWS with SOC2 compliance and all that. Normally switching to Codex would have a ton of…

  24. comment
    Comment #47967925

    No, you definitely don't have access to the weights. The raw weights are secret enough that when a model hits a certain level of capability, their guidelines are that they need eno…

  25. comment
    Comment #47967858

    Besides what the other person mentioned about being more useful for enterprise, I also heard mentioned on a podcast that gpt-image-2 uses the same general architecture as the LLM m…