Viewing profile — jafitc
jafitc
HN member- Joined
- Wed, Apr 30, 2014, 8:31 PM UTC
- HN karma
- 781
- Public activity
- 102 items
- HN profile
- View on Hacker News ↗
About jafitc
No profile information was provided.
Recent public activity
-
comment
Comment #48883580
"thanks to lobbying". if you use "for" it sounds like you are adressing the OP and he did the lobbying that you don't like.
-
comment
Comment #47809255
bigger change here might not be model quality, but debuggability. once you hide the reasoning, remove the knobs, and let the model choose its own effort, it gets much harder to tel…
-
comment
Comment #47531042
subprime mortgages sprinkled on top of prime ones, treated as prime ones. because they were printing money. subprime code sprinkled on the backbone of software we use everyday. bec…
-
comment
Comment #47531008
that's a great story!
-
comment
Comment #42050579
I think you should consider trimming that file. Exclude movies with very low number of rating or potentially very low scores too. The long tail reduction would be significant
-
comment
Comment #39060238
"People are really bad at understanding just how big LLM's actually are. I think this is partly why they belittle them as 'just' next-word predictors" https://nitter.net/jam3scampb…
-
comment
Comment #39020555
Deepinfra Mixtral is $0.27 / M tokens as per their website
-
story
ChatGPT Website Updated With Long-Term Memory Feature (RAG)
Multiple reports (no release note yet): - https://twitter.com/gblazex/status/1744914340203913537 - https://twitter.com/jasonaholloway/status/1744909952349868534 - https://www.reddi…
- story
-
comment
Comment #38890717
Important to note that this model excels in reasoning capabilities. But it was on purpose not trained on the big “web crawled” datasets to not learn how to build bombs etc, or be n…
-
comment
Comment #38890668
Do you think the ISIS is bound by the words “non-commercial” in a license file when they have the source anyway? It was available even before this, all they changed is that law abi…
- story
-
comment
Comment #38661000
This "vibe" check that it's even better than GPT-4 Turbo is not what its Elo rating shows on the Chatbot Arena based on not 1 but thousands of user votes. GPT-4 (Turbo) is in a lea…
-
comment
Comment #38660138
This is based on users choosing the better from 2 models at a time, and calculating an ELO rating from who-beats-who. BYOT - bring your own tests style. Gives a better picture of r…
- story
-
comment
Comment #38642186
from the announcement tweet: https://twitter.com/rasbt/status/1735293149965062476 --- So, we've been quietly building something new for running AI experiments and deploying models …
- story
- story
-
comment
Comment #38590863
Important note: Bing balanced mode (default) uses GPT 3.5 Only Precise and Creative modes use GPT-4 https://twitter.com/emollick/status/1732495030143549541 Also see: An Opinionated…
-
comment
Comment #38560978
All I can say is it’s really fast
-
comment
Comment #38555422
It’ll never be completely gone. But you’ll need it in less and less everyday scenarios and time goes on Just like we need to write less and less assembly by hand
-
comment
Comment #38555403
OpenAI provides “instruct” version of their models (Not optimized for chat)
-
comment
Comment #38555397
Isn’t that the first sentence?
-
comment
Comment #38555377
Brain is just neurons and synapses at the end of the day. The whole universe might just be a stochastic swirl of milk in a shaken up mug of coffee. Looking at something under a mic…
-
comment
Comment #38555339
The fact that the makers of such LLM make a post about it shows that they have incentive to cater to even these kind of use cases