Viewing profile — marci
marci
HN member- Joined
- Tue, Oct 10, 2017, 12:24 PM UTC
- HN karma
- 515
- Public activity
- 345 items
- HN profile
- View on Hacker News ↗
About marci
No profile information was provided.
Recent public activity
-
comment
Comment #49253948
delta.chat is email https://delta.chat/en/help#can-i-use-a-classic-email-address...
-
comment
Comment #49239769
Then maybe what you want is email? Sprinkled with a little bit of https://delta.chat ? And potentially a dash of iroh https://news.ycombinator.com/item?id=48542480 https://www.iroh…
-
comment
Comment #49170836
Nothing to understand. Straight up hallucination. I could have sworn I read that they used a novel architecture where the model is dense but you could select specific layers or som…
-
comment
Comment #49165389
Seems like what Apple's going for with afm3. Their latest model that will be embedded in macOS 27 is a quantized dense 20B that only select between 1 to 4B at inference, based on t…
-
comment
Comment #49154827
https://xcancel.com/Alibaba_Qwen/status/2084100707423289643#...
-
comment
Comment #49060235
Makes me wonnder... how much compute/storage there's in all the satellites currently in LEO combined.
-
comment
Comment #48715964
This was a preview release. They haven't finish training. The Pro contains more knowledge but it probably takes longer training than flash for the smarts to kick in.
-
comment
Comment #48684616
"That’s where EMO comes in. We show that EMO – a 1B-active, 14B-total-parameter (8-expert active, 128-expert total) MoE trained on 1 trillion tokens – supports selective expert use…
-
comment
Comment #48683229
But with Apple's AFM 3 architecture, we might end up with huge SOTA adjacent on devices with limited RAM. They use a technique where you only load between 1B and 4B of a 20B dense …
-
comment
Comment #48537516
Don't worry. Most people spend most of their compute time on a phone, where you're ability to filter ads is way more enshitified.
-
comment
Comment #48280190
I wonder where's the line between using a font and copyright/trademark infringement.
-
comment
Comment #48145371
"That’s where EMO comes in. We show that EMO – a 1B-active, 14B-total-parameter (8-expert active, 128-expert total) MoE trained on 1 trillion tokens – supports selective expert use…
-
comment
Comment #48066173
Did they modify their post? I can't see who claimed that consumer hardware will be able to build most things?
-
comment
Comment #47573916
when you sign an app with your personal dev account.
-
comment
Comment #47569596
That's just a regular rounndabout. I thought you were talking about this: https://www.youtube.com/watch?v=6OGvj7GZSIo
-
comment
Comment #47503939
Also everything from scratch by allen.ai. Weights, datasets, code, multiple checkpoints... I like their FlexOlmo concept.
-
comment
Comment #47456468
I think it's mostly because most cpus that can run a gpu already have parts dedicated as h264 encoder, way more efficient energy wise and speed wise.
- story
-
comment
Comment #47251153
Unfortunately, the most extreme is that it's the new normal that now, there's >0 chance that someone, whether they are a US citizen or not apparently, child or adult, can end up in…
-
comment
Comment #47051306
Imagine, a llm trained on the best thrillers, spy stories, politics, history, manipulation techniques, psychology, sociology, sci-fi... I wonder where it got the idea for deception…
- comment
-
comment
Comment #46883220
Their issue with the mac was the sound of fans spinning. I doubt a dedicated gpu will resolved that.
-
comment
Comment #46679360
"finetune" Not "Train from scratch"
-
comment
Comment #46679266
It is so, so long... I barely reached the middle before my brain just "Nope." They are talking about this kind of battery replacement: https://www.ifixit.com/Guide/Fairphone+3+Batt…
-
comment
Comment #46679024
If your family members ever had to mount an ikea furniture or equivalent, they'll probably have an as easy or easier time replacing a part on a fairphone. Especially for the batter…