Live data from Hacker News

Viewing profile — cfcf14

cfcf14

HN member
Joined
Fri, May 04, 2018, 6:20 PM UTC
HN karma
489
Public activity
87 items

About cfcf14

No profile information was provided.

Recent public activity

  1. comment
    Comment #47823973

    I would assume so too, so the costs would not be so substantial to Anthropic.

  2. comment
    Comment #47823597

    I'm curious as to why 4.7 seems obsessed with avoiding any actions that could help the user create or enhance malware. The system prompts seem similar on the matter, so I wonder if…

  3. comment
    Comment #46486981

    This makes me think it would be nice to see some kinda child of modern transformer architecture and neural ODEs. There was such interesting work a few years ago on how neural ode/p…

  4. comment
    Comment #43156835

    The obvious next step here is to see how well this generalises to arbitrary inputs :)

  5. comment
    Comment #41996102

    Did your read the paper? Do you have specific criticisms of their problem statement, methodology, or results? There is a growing body of research indicating that in fact, there _is…

  6. comment
    Comment #41901653

    AI detectors do not work. I have spoken with many people who think that the particular writing style of commercial LLMs (ChatGPT, Gemini, Claude) is the result of some intrinsic ch…

  7. comment
    Comment #41775644

    So uh, things are not looking so good for actual physics these days, I gather?

  8. comment
    Comment #41774475

    L-theanine (200mg) with around 100-150mg of caffeine has an extremely noticeable, positive effect on my ability to focus, feeling of "well-situatedness", and overall calmness. L-th…

  9. comment
    Comment #41235997

    It's not - the FTC released a statement on this very topic a few months ago: https://www.ftc.gov/business-guidance/blog/2024/03/price-fix...

  10. comment
    Comment #41234903

    Police in large American cities are not likely to be of much assistance in this situation. Assuming they attend at all, I would expect them to not understand the nature of the issu…

  11. comment
    Comment #41233746

    After reading the paper, I'm really unsure what the novel contribution is. It feels like they're attempting to rebrand well-understood concepts within various fields (control syste…

  12. comment
    Comment #37060304

    This is really funny - it's bordering on truly absurd, almost incomprehensible madness to consider doing this seriously. I can't think of a single property you'd desire in a contro…

  13. comment
    Comment #37047494

    Was fun while it lasted! Will be interesting watching the internal story of the original lab unfold as it all becomes public eventually.

  14. comment
    Comment #36710161

    Yeah - more or less. I still use it sometimes for trivial stuff like giving me recipe or travel inspiration (using the web search API), but I haven't been using it for any sort of …

  15. comment
    Comment #36326407

    Some strange claims in this post. The reddit post datasets are already 'out there' in the wild, and I'm fairly certain every other major LLM release has used their data. Also - did…

  16. comment
    Comment #36135913

    Yeah, definitely. Combination of expert-system gating (some requests probably get routed to weaker models), distillation (for performance/cost), and RLHF lobotomization.

  17. comment
    Comment #35989013

    It's turtles all the way down, except for the final turtle, which is Fortran...

  18. comment
    Comment #35232225

    Lilian Weng's blog is my go-to example for an extremely high quality tech blog, it's truly remarkable how consistently excellent each post is. The only downside is the sadness I fe…

  19. comment
    Comment #35179745

    This is 100% related to (suspected) fraud, anti money laundering, or other types of financial/political sanctions. You may be 100% innocent, but they will never disclose any inform…

  20. comment
    Comment #35114132

    Absolutely not.

  21. comment
    Comment #35067229

    You have to prompt it correctly, non-instruction-aligned models don't behave like agent simulators by default.

  22. comment
    Comment #34934967

    Amazing post, agree with everything you've said. I've always felt that the problems with advanced MCMC methods (HMC, RM-MC, etc) are even more painful when one looks at approximate…

  23. comment
    Comment #34805832

    I wonder whether Bing has been tuned via RLHF to have this personality (over the boring one of ChatGPT); perhaps Microsoft felt it would drive engagement and hype. Alternately - ma…

  24. comment
    Comment #34332046

    This was a really reasonable and interesting post by Stephen. I'm excited to see what the integration between an associative based model like GPT and a symbolic one like WA might b…

  25. comment
    Comment #33328988

    It being from Schmidhuber's lab makes it dramatically more credible, in my views. They've been practically a decade ahead of everybody for ages now from a theoretical point of view…