Viewing profile — saurabh20n
saurabh20n
HN member- Joined
- Mon, Oct 13, 2014, 6:46 PM UTC
- HN karma
- 582
- Public activity
- 151 items
- HN profile
- View on Hacker News ↗
About saurabh20n
Founder, Synthetic Minds YC S'18 - Program synthesis for desktop automation https://warpdrive.co
Founder, 20n YC W'15 - Program synthesis for synthetic biology: http://20n.com
Postdoc (UC Berkeley): Program synthesis for cell designs to engineer cells using synthetic biology.
PhD (University of Maryland): Program synthesis.
Personal page: http://www.saurabh-srivastava.com
Recent public activity
- story
- story
-
comment
Comment #40082602
Looks like you’re one of the authors. It would be nice if you could post if the actual data matches your reconstruction—now that you have it in hand. Would help us not worry about …
-
comment
Comment #38476711
Discussion of the 72B model happening here: https://news.ycombinator.com/item?id=38475501
- story
-
comment
Comment #38476502
Summary from https://arxiv.org/pdf/2309.16609.pdf --- (q: how does one format lists on HN?) * qwen-{1.8B,7B,14B}: * 3 trillion tokens; start with BPE tiktoken, cl100k base vocab, a…
-
comment
Comment #38422632
actual title: “Prompting Frameworks for Large Language Models: A Survey” LLM frameworks might imply stack for building models (pertaining, fine tuning, inference etc)
-
comment
Comment #35658411
Notes from quick read of paper at https://arxiv.org/abs/2302.10866 . Title of popsci is overreaching, this is a drop-in subquadratic replacement for attention. Could be promising, …
- story
- story
- story
- story
- story
- story
- story
-
comment
Comment #35155103
Congrats on the launch. I think you should share some technical details for a more substantial pitch. You are using the OSS BigCode effort and "The Stack" [1, 2] (as you say in ano…
-
comment
Comment #34927099
Quick notes from first glance at paper https://research.facebook.com/publications/llama-open-and-ef... : * All variants were trained on 1T - 1.4T tokens; which is a good compared t…
-
comment
Comment #34599836
The last author's tweet thread and replies have some interesting tidbits: https://twitter.com/Eric_Wallace_/status/1620449934863642624 * "We propose to extract memorized images by …
- story
-
comment
Comment #32323867
For the curious, here are direct links: * Initialization was done 42 days ago: https://etherscan.io/tx/0x53fd92771d2084a9bf39a6477015ef53b7... -- "Click to see More" and notice "In…
-
comment
Comment #29946943
https://ericpony.github.io/z3py-tutorial/guide-examples.htm should be a quick start.
-
comment
Comment #29946795
For this problem, an enumerative solver may be more optimal (and faster); where optimality is finding the word with the least number of guesses: There are ~158k 5-letter words. Sta…
- story
- story
- job