Live data from Hacker News

Viewing profile — sourabh03agr

sourabh03agr

HN member
Joined
Sun, Sep 13, 2020, 4:31 PM UTC
HN karma
75
Public activity
75 items

About sourabh03agr

No profile information was provided.

Recent public activity

  1. comment
    Comment #39770976

    I have experienced such issues with GPT-4 returning the wrong syntax, especially when asking questions about complex operations in Polars. We are actually running performance monit…

  2. comment
    Comment #39767485

    Hello HN! I am happy to share the Monitoring reports we have been running for the past few months to identify regression in popular LLMs like GPT-4-turbo, Claude-2, etc. There have…

  3. story
  4. story
  5. comment
    Comment #39605958

    Congrats on the launch! Do you need Github permissions to answer questions on open-source repos as well?

  6. story
  7. story
  8. story
  9. story
    Is Gemini a case of a biased tuning dataset or model overfitting?

    A quick answer is both, and here is why. Many of us may have seen posts where Google's Gemini model inadvertently generated images with an incorrect skin tone for some well-known i…

  10. comment
    Comment #39461323

    Great read! On being visible on social media efforts, how do you measure the ROI of your efforts, like how much time do you spend vs how much revenue does it bring?

  11. comment
    Comment #39461241

    Nice idea! From the perspective of someone who is in need of such a solution, I have explored similar apps before, and one of the biggest issues was that I don't look for events re…

  12. comment
    Comment #39461186

    That's fair, but a lot of use cases require strict information retrieval and don't want the LLM to get creative. I am of the opinion that having an LLM which is always factually co…

  13. story
    How does one detect hallucinations?

    Hallucinations are an interesting artifact of LLMs where the model tends to make up facts or generate outputs that are not factually correct. There are two broad approaches for det…

  14. story
  15. comment
    Comment #39352704

    Nice work! I see you have an evaluation module - what all are you evaluating for? Primarily Question-answer accuracy via Exact Match?

  16. story
  17. story
  18. story
    Ask HN: Including irrelevant documents improve RAG accuracy by 30%?

    I was very intrigued after reading "The Power of Noise: Redefining Retrieval for RAG Systems" paper (https://arxiv.org/pdf/2401.14887.pdf), where they explored the impact of releva…

  19. story
  20. story
  21. story
  22. comment
    Comment #39143070

    Guess Indian is far better than US and Japan when it comes to producing plastic waste. A lot of Indian states have plastic bans, drastically reducing the amount of plastic wasted v…

  23. story
  24. comment
    Comment #39120483

    Love the product, a happy user here!

  25. story