Live data from Hacker News

Viewing profile — rajman187

rajman187

HN member
Joined
Tue, Nov 07, 2017, 10:23 AM UTC
HN karma
355
Public activity
166 items

About rajman187

No profile information was provided.

Recent public activity

  1. comment
    Comment #46746899

    i think it's worth revisiting this in a short while because, by and large, how the engineering craft has been for the last 40+ years is no longer the correct paradigm. it takes Cla…

  2. comment
    Comment #46451179

    Re: cerebras, they filed a S1 [1] last year when attempting to go public. It showed something like a $60M+ loss for the first 6 months of 2024. The IPO didn’t happen because the CE…

  3. comment
    Comment #45922634

    > You could cut your MongoDB costs by 100% by not using it ;) Came here to say exactly this

  4. comment
    Comment #45856309

    They’ve filed a S1 [1] last year when attempting to go public. It showed something like a $60M+ loss for the first 6 months of 2024. The IPO didn’t happen because the CEO’s past in…

  5. comment
    Comment #45794953

    If I could upvote this more than once I certainly would

  6. comment
    Comment #45391965

    My main keyboard has been a 34-key split Ferris. I usually have either a trackpad between the halves if I’m using a Mac or an ergonomic Logitech if on my Linux desktop. Not having …

  7. comment
    Comment #44968970

    Yeah the org structure is one thing, the missions are another. Yann adds some clarity here https://www.linkedin.com/posts/yann-lecun_were-excited-to-ha...

  8. comment
    Comment #44907319

    This has nothing to do with the newly appointed fellow nor Meta Superintelligence Labs, but rather work from FAIR that would have gone through a lengthy review process before seein…

  9. comment
    Comment #44858772

    An intuitive treatment of RLHF, TRPO, PPO, GRPO, DPO and RLAIF

  10. story
  11. story
  12. comment
    Comment #44254201

    That’s why you have encoders as well as decoders. For example, another model from Meta does this for translations; they have encoders and decoders into a single embedding space tha…

  13. comment
    Comment #43395816

    Not a lawyer but would assume downloading material from libgen is, in the vast majority of cases, illegal because it's a breach of copyright or similar. That’s gotten Meta in quite…

  14. comment
    Comment #43378883

    Well there was the case of an employee leaving due to his perceived moral issues around the use of copyrighted material in the training dataset [1] [1] https://www.pbs.org/newshour…

  15. comment
    Comment #43360884

    > other than a bit of open source (PyTorch and React are nice, I guess) Not to detract from your main point but I think this misses a lot of contributions, eg Cassandra, Hive, Pres…

  16. comment
    Comment #43123663

    > It was clearly valuable from day 1 I’m not sure that’s the case even if in retrospect we can clearly argue this In 1998, Paul Krugman, winner of the Nobel memorial prize in econo…

  17. comment
    Comment #42413600

    It originates in Yann LeCunn’s paper from 2022 [1], the term AMI being district from AGI. However, the A has changed over the past few years from autonomous to advanced and even au…

  18. comment
    Comment #41500817

    From the documentation [1] > The mission of Sail is to unify stream processing, batch processing, and compute-intensive (AI) workloads. Currently, Sail features a drop-in replaceme…

  19. comment
    Comment #41411435

    MTIA will be for inference initially. Another to add to the list is wafer maker Cerebras https://www.forbes.com/sites/craigsmith/2024/08/27/cerebras-...

  20. comment
    Comment #40708759

    Meta is already working on this [1], not sure it can replace NVIDIA for training large models within that time frame however. The ecosystem around their chips is what gives a huge …

  21. comment
    Comment #40003362

    The title and opening is perhaps giving people reason to infer something which Lecun isn’t arguing, that AI isn’t going to reach such a level of intelligence. Indeed, he’s publishe…

  22. comment
    Comment #39933273

    I would have liked to see a more generic implementation that isn't necessarily tied to NVIDIA, while I agree that's a much greater ask than a 7-person team can probably take on, th…

  23. comment
    Comment #38716211

    In a world of finite time and resources, wouldn't it be more useful to improve services on the lines themselves? The Elizabeth line had the highest rate of cancelations in the enti…

  24. comment
    Comment #37647302

    While I agree the public transit in the US is abysmal at best, I'm not sure the UK is a good measure anymore. Rail strikes and engineering work are happening with such frequency th…

  25. comment
    Comment #37328931

    Several years ago Walmart dramatically sped up their online store's performance by storing images as blobs in their distributed Cassandra cluster. https://medium.com/walmartglobalt…