Live data from Hacker News

Viewing profile — iliane5

iliane5

HN member
Joined
Fri, Dec 02, 2022, 8:54 PM UTC
HN karma
85
Public activity
44 items

About iliane5

No profile information was provided.

Recent public activity

  1. comment
    Comment #39904176

    Just wanted to say thank you for your work and attention to detail, it's immensely valuable and we're all very grateful for it.

  2. comment
    Comment #39392563

    Watching an entirely generated video of someone painting is crazy. I can't wait to play with this but I can't even imagine how expensive it must be. They're training in full resolu…

  3. comment
    Comment #37862166

    I think it's mostly the scale. Once you have a consistent user base and tons of GPUs, batching inference/training across your cluster allows you to process requests much faster and…

  4. comment
    Comment #36311490

    What I was saying is that because you need to go out of your way to make sure it's tokenized properly, I wouldn't be surprised if there are enough non properly tokenized examples i…

  5. comment
    Comment #36303927

    > LLMs are not particularly good at arithmetic, counting syllables, or recognizing haikus I suspect most of this is due to tokenization making it difficult to generalize these conc…

  6. comment
    Comment #36158541

    AFAIK it's pretty standard practice not to expose the "raw" LLM directly to the user. You need a "sanity loop" where user input and the output of the LLM is checked by another LLM …

  7. comment
    Comment #36104736

    100% agree. However, seeing how excited Palantir is with their war assistant LLM , the US testing autonomous fighter jets a few months ago, etc. I think there's a decent chance tha…

  8. comment
    Comment #35981413

    [flagged]

  9. comment
    Comment #35963625

    I don't think we need sentient AI for it to be autonomous. LLMs are powerful cognitive engines and weak knowledge engines. Cognition on its own does not allow them to be autonomous…

  10. comment
    Comment #35962286

    > Why is building what amounts to a calculator/spreadsheet/CAD program for language somehow a Rubicon that cannot be crossed? We've already crossed it and I believe we should go fu…

  11. comment
    Comment #35962127

    There's no denying this is regulatory capture by OpenAI to secure their (gigantic) bag and that the "AI will kill us all" meme is not based in reality and plays on the fact that th…

  12. comment
    Comment #35961948

    The human brain works around a lot of limiting biological functions. The necessary architecture to fully mimic a human brain on a computer might not look anything like the actual h…

  13. comment
    Comment #35961691

    > I don't want every to know how to make a bomb. This information is not created inside the LLMs, it's part of their training data. If someone is motivated enough, I'm sure they'd …

  14. comment
    Comment #35961613

    > Why is it so hard to hear this perspective? Like, genuinely curious. Because people have different definition of what intelligence is. Recreating the human brain in a computer wo…

  15. comment
    Comment #35666019

    Agreed, there is way too much hype about the actual capabilities of the LLaMa models. However, instruction tuning alone makes Alpaca much more usable than the the base model and to…

  16. comment
    Comment #35645928

    I'm sure they're tweaking lots of things under the hood, especially now that they have 100M+ users. It could be bigger (30B?, maybe 65B) as coming down from 175B gives quite a lot …

  17. comment
    Comment #35642294

    GPT-3.5 is much worse at "complex" cognitive tasks than Davinci (175B), which seem to indicate that it's a smaller model. It's also much faster than Davinci and costs the same as C…

  18. comment
    Comment #35580501

    It’s not only 10x cheaper, it’s also way faster at inference and not as smart as Davinci. IMO the only logical answer is that the model is just smaller.

  19. comment
    Comment #35575731

    I bet they’re not saying how big of a model GPT-4 is because it’s actually much smaller we would expect. ChatGPT is IMO a heavily fine-tuned Curie sized model (same price via API +…

  20. comment
    Comment #35476764

    I think as soon as text2video gets really good (like midjourney level), there’s gonna be so much AI generated content that unless it’s all extremely good, human made content will b…

  21. comment
    Comment #35476724

    > LLMs architected and trained the way they are now can never approach human reasoning capability Not sure if you’ve played with GPT-4 but honestly it’s getting there. If you take …

  22. comment
    Comment #35476612

    Absolutely. What’s fascinating is that they’re getting such good understanding of many things through just text. Multimodal models that can process text, images, sounds, video, etc…

  23. comment
    Comment #35473369

    There's no denying LLMs are anything but sentient however is sentience really needed for intelligence? I feel like if we can have machines that are X% smarter than a human could ev…

  24. comment
    Comment #35473255

    > But it's not better than almighty human intelligence, it _is_ human intelligence, because it was trained on a mass of some of the best human intelligence in all recorded history …

  25. comment
    Comment #35473134

    Well it's like birds and airplanes. Do airplanes "fly" in the same sense that birds do? Of course not, birds flap their wings and airplanes need to be built, fueled and flown by hu…