Live data from Hacker News

Viewing profile — saurabh20n

saurabh20n

HN member
Joined
Mon, Oct 13, 2014, 6:46 PM UTC
HN karma
582
Public activity
151 items

About saurabh20n

Building foundational code models @ Essential AI.

Founder, Synthetic Minds YC S'18 - Program synthesis for desktop automation https://warpdrive.co

Founder, 20n YC W'15 - Program synthesis for synthetic biology: http://20n.com

Postdoc (UC Berkeley): Program synthesis for cell designs to engineer cells using synthetic biology.

PhD (University of Maryland): Program synthesis.

Personal page: http://www.saurabh-srivastava.com

Recent public activity

  1. story
  2. story
  3. comment
    Comment #40082602

    Looks like you’re one of the authors. It would be nice if you could post if the actual data matches your reconstruction—now that you have it in hand. Would help us not worry about …

  4. comment
    Comment #38476711

    Discussion of the 72B model happening here: https://news.ycombinator.com/item?id=38475501

  5. story
  6. comment
    Comment #38476502

    Summary from https://arxiv.org/pdf/2309.16609.pdf --- (q: how does one format lists on HN?) * qwen-{1.8B,7B,14B}: * 3 trillion tokens; start with BPE tiktoken, cl100k base vocab, a…

  7. comment
    Comment #38422632

    actual title: “Prompting Frameworks for Large Language Models: A Survey” LLM frameworks might imply stack for building models (pertaining, fine tuning, inference etc)

  8. comment
    Comment #35658411

    Notes from quick read of paper at https://arxiv.org/abs/2302.10866 . Title of popsci is overreaching, this is a drop-in subquadratic replacement for attention. Could be promising, …

  9. story
  10. story
  11. story
  12. story
  13. story
  14. story
  15. story
  16. comment
    Comment #35155103

    Congrats on the launch. I think you should share some technical details for a more substantial pitch. You are using the OSS BigCode effort and "The Stack" [1, 2] (as you say in ano…

  17. comment
    Comment #34927099

    Quick notes from first glance at paper https://research.facebook.com/publications/llama-open-and-ef... : * All variants were trained on 1T - 1.4T tokens; which is a good compared t…

  18. comment
    Comment #34599836

    The last author's tweet thread and replies have some interesting tidbits: https://twitter.com/Eric_Wallace_/status/1620449934863642624 * "We propose to extract memorized images by …

  19. story
  20. comment
    Comment #32323867

    For the curious, here are direct links: * Initialization was done 42 days ago: https://etherscan.io/tx/0x53fd92771d2084a9bf39a6477015ef53b7... -- "Click to see More" and notice "In…

  21. comment
    Comment #29946943

    https://ericpony.github.io/z3py-tutorial/guide-examples.htm should be a quick start.

  22. comment
    Comment #29946795

    For this problem, an enumerative solver may be more optimal (and faster); where optimality is finding the word with the least number of guesses: There are ~158k 5-letter words. Sta…

  23. story
  24. story
  25. job