Viewing profile — rar00
rar00
HN member- Joined
- Tue, Nov 26, 2024, 10:11 PM UTC
- HN karma
- 41
- Public activity
- 36 items
- HN profile
- View on Hacker News ↗
About rar00
Recent public activity
-
comment
Comment #46709017
> the researchers noticed a group of neurons that consistently signaled the wrong response. Instead of fading as learning improved, these neurons grew stronger, and occasionally ev…
-
comment
Comment #46069884
Really cool! An affective basis of homeostatic drive seems promising. Have you performed any basic evaluation / test of your approach? I'm also curious if there was any deliberatio…
-
comment
Comment #46050274
typo? Rounding it up to 2 billion, 30% means 600 million per year
-
comment
Comment #46048234
hmm, can't tell if complete bullshit or a work of genius. On the one hand, the approach overlaps a lot with my thinking, and has some original tweaks (like the emotionally valenced…
-
comment
Comment #45424494
I started reading two recent neuroscience books (elusive cures and natural neuroscience) that while have different goals both highlight the utility of systems neuroscience. In elus…
-
comment
Comment #45138437
"As [Essential AI Labs (founded in 2021 by Vaswani)] changes focus, Vaswani is asking investors for at least $150 million." expected, yet still funny. Noting that their initial aim…
-
comment
Comment #45138361
https://archive.ph/pn5Tt
-
comment
Comment #44984102
I know people are pushing back, taking "only" literally, but from a reasonable perspective what causes LLMs (technically their outputs) to give that impression is indeed the crux o…
-
comment
Comment #44427850
That's an orthogonal concern IMO. Running a batch in Europe is about tapping into another source of opportunities. There are plenty of founders that won't or aren't able to attend …
-
comment
Comment #44426231
disagree, there are a few organisations exploring novel paths. It's just that throwing new data at an "old" algorithm is much easier and has been a winning strategy. And, also, the…
-
comment
Comment #44376582
yep, even with greedy sampling and fixed system state, numerical instability is sufficient to make output sequences diverge when processing the same exact input
-
comment
Comment #44376558
This argument works better for state space models. A transformer would still steps context one token at a time, not maintain an internal 1e18 state.
-
comment
Comment #44355940
getting a billionaire to finance your startup with 9-figures for no equity... :D
-
comment
Comment #44275768
It is a sensible position. YC and VCs are backing businesses, not charitable causes or research initiatives. It is the founders' responsibility to liaise the two sides in order to …
-
comment
Comment #44251856
the robot arm demonstration video jumps at the 00:28s mark...
-
comment
Comment #44246929
Model weights already available in HF, code to be released shortly in GitHub ( https://github.com/nvidia-cosmos/cosmos-predict2 )
- story
- story
-
comment
Comment #42818203
aligns with (or is based on) Demis Hassabis' assessment from yesterday on missing cognitive capabilities for AGI: long-term memory, reasoning, hierarchical planning. He then goes o…
-
comment
Comment #42817131
Says that new cognitive capabilities are needed to attain AGI: reasoning, hierarchical planning, long-term memory, inventive creativity. Interesting that LeCun also mentions the sa…
- story
- story
- story
-
comment
Comment #42756670
> However, we have a verbal agreement that these materials will not be used in model training. That'll do it... clearly no incentive to do otherwise. There should be some form of a…
-
comment
Comment #42756609
nice way to spoil the fight result to people. Because, of course, an athlete gaining more than its adversary is HN-worthy news...