Live data from Hacker News

Viewing profile — jafitc

jafitc

HN member
Joined
Wed, Apr 30, 2014, 8:31 PM UTC
HN karma
781
Public activity
102 items

About jafitc

No profile information was provided.

Recent public activity

  1. comment
    Comment #48883580

    "thanks to lobbying". if you use "for" it sounds like you are adressing the OP and he did the lobbying that you don't like.

  2. comment
    Comment #47809255

    bigger change here might not be model quality, but debuggability. once you hide the reasoning, remove the knobs, and let the model choose its own effort, it gets much harder to tel…

  3. comment
    Comment #47531042

    subprime mortgages sprinkled on top of prime ones, treated as prime ones. because they were printing money. subprime code sprinkled on the backbone of software we use everyday. bec…

  4. comment
    Comment #47531008

    that's a great story!

  5. comment
    Comment #42050579

    I think you should consider trimming that file. Exclude movies with very low number of rating or potentially very low scores too. The long tail reduction would be significant

  6. comment
    Comment #39060238

    "People are really bad at understanding just how big LLM's actually are. I think this is partly why they belittle them as 'just' next-word predictors" https://nitter.net/jam3scampb…

  7. comment
    Comment #39020555

    Deepinfra Mixtral is $0.27 / M tokens as per their website

  8. story
    ChatGPT Website Updated With Long-Term Memory Feature (RAG)

    Multiple reports (no release note yet): - https://twitter.com/gblazex/status/1744914340203913537 - https://twitter.com/jasonaholloway/status/1744909952349868534 - https://www.reddi…

  9. story
  10. comment
    Comment #38890717

    Important to note that this model excels in reasoning capabilities. But it was on purpose not trained on the big “web crawled” datasets to not learn how to build bombs etc, or be n…

  11. comment
    Comment #38890668

    Do you think the ISIS is bound by the words “non-commercial” in a license file when they have the source anyway? It was available even before this, all they changed is that law abi…

  12. story
  13. comment
    Comment #38661000

    This "vibe" check that it's even better than GPT-4 Turbo is not what its Elo rating shows on the Chatbot Arena based on not 1 but thousands of user votes. GPT-4 (Turbo) is in a lea…

  14. comment
    Comment #38660138

    This is based on users choosing the better from 2 models at a time, and calculating an ELO rating from who-beats-who. BYOT - bring your own tests style. Gives a better picture of r…

  15. story
  16. comment
    Comment #38642186

    from the announcement tweet: https://twitter.com/rasbt/status/1735293149965062476 --- So, we've been quietly building something new for running AI experiments and deploying models …

  17. story
  18. story
  19. comment
    Comment #38590863

    Important note: Bing balanced mode (default) uses GPT 3.5 Only Precise and Creative modes use GPT-4 https://twitter.com/emollick/status/1732495030143549541 Also see: An Opinionated…

  20. comment
    Comment #38560978

    All I can say is it’s really fast

  21. comment
    Comment #38555422

    It’ll never be completely gone. But you’ll need it in less and less everyday scenarios and time goes on Just like we need to write less and less assembly by hand

  22. comment
    Comment #38555403

    OpenAI provides “instruct” version of their models (Not optimized for chat)

  23. comment
    Comment #38555397

    Isn’t that the first sentence?

  24. comment
    Comment #38555377

    Brain is just neurons and synapses at the end of the day. The whole universe might just be a stochastic swirl of milk in a shaken up mug of coffee. Looking at something under a mic…

  25. comment
    Comment #38555339

    The fact that the makers of such LLM make a post about it shows that they have incentive to cater to even these kind of use cases