Live data from Hacker News

Viewing profile — mlin4589

mlin4589

HN member
Joined
Mon, Oct 10, 2022, 2:48 AM UTC
HN karma
12
Public activity
7 items

About mlin4589

No profile information was provided.

Recent public activity

  1. comment
    Comment #43919127

    The reality, I suspect is that internally models are likely modeling these alignment features such as refusals as a secondary filter. In fact, for many models you can remove refusa…

  2. comment
    Comment #43919062

    Calibration (in a binary context) basically means that the confidence of a model/score matches the probability that a particular label is positive or not. For instance, a calibrate…

  3. comment
    Comment #43912267

    Good question! We do know from OpenAI's system card from GPT-4 that the post-trained RLHF model is significantly less calibrated compared to the pre-trained model, so it's a matter…

  4. job
    Intrinsic (YC W23) Is Hiring

    We're building real-time abuse prevention systems. Hiring Full-stack, frontend, and MLEs across product and PoCs. https://withintrinsic.com/careers

  5. comment
  6. story
  7. comment