Live data from Hacker News

Viewing profile — futureshock

futureshock

HN member
Joined
Sun, Mar 08, 2015, 10:21 AM UTC
HN karma
14,271
Public activity
2,709 items

About futureshock

No profile information was provided.

Recent public activity

  1. comment
    Comment #49146193

    Judging by the Seedance 2.5 demos today, I’d say it’s not that many orders of magnitude away now.

  2. comment
    Comment #49110431

    Could have been worse really. It had an open internet connection. At least it didn’t take the researchers family hostage.

  3. comment
    Comment #49024907

    I still think there’s something to be said here for generality. This does not appear to been designed as a cyber pen test tool with specialized harness. From what I understand they…

  4. comment
    Comment #48867836

    I think a lot of this has to do with the post-training these models normally get. They are designed to answer basic questions with straightforward and short summary answers. They h…

  5. comment
    Comment #48850623

    There has been a lot of chatter ever since the Mythos scores had been release that SWEbench pro had major contamination and that Mythos had memorized many questions that lacked the…

  6. comment
    Comment #48741479

    I think this is black and white thinking. Fable and US AI is not unique technology. It’s just marginally better than open source tech at 10 times the price. You can swap out the mo…

  7. comment
    Comment #48434042

    Yes I was the exact same. I got curious during the GPT-3 release and went over to AI Dungeon. It was just running GPT-2. Hmm wow interesting. This felt new! Then I subscribed so I …

  8. comment
    Comment #48162602

    It is plausible, the model would just need to be trained on a lot of stereoscopic data.

  9. comment
    Comment #48162591

    World in this context means that these videos are interactive, just like a video game. In the linked examples you can see the keyboard and mouse inputs. The model is trained to mai…

  10. comment
    Comment #48147459

    I think this is interesting because it collides my intuition from the pre-adtech world with the post. Surely collecting telemetry on nearly every mile you drive could never be a se…

  11. comment
    Comment #48019328

    Reducing the network latency helps with this exactly. OpenAI can make better timed decisions when to begin responding so it'll feel less like an interruption. I've also seen some r…

  12. comment
    Comment #47698038

    “A human being should be able to change a diaper, plan an invasion, butcher a hog, conn a ship, design a building, write a sonnet, balance accounts, build a wall, set a bone, comfo…

  13. comment
    Comment #47522605

    Well yes, that is exactly the point! The very purpose of the ARC AGI benchmarks is to find a pure reasoning task that humans are very good at and AI is very bad at. Companies then …

  14. comment
    Comment #47522530

    The evidence is that humans are able to win these games. AGI is usually defined as the ability to do any intellectual task about as well as a highly competent human could. The poin…

  15. comment
    Comment #46994725

    I think step 4 is the agent swarm. Manager model gets the prompt and spins up a swarm of looping subagents, maybe assigns them different approaches or subtasks, then reviews result…

  16. comment
    Comment #46724301

    My workaround is I use SMS 2 factor for banking and use my Google Voice number.

  17. comment
    Comment #46724283

    I think this is clearly the way forward for Apple. The rest is just UX and refinement. I recently set up a Shortcut on my Apple Watch that lets me bypass Siri and talk directly to …

  18. comment
    Comment #46642472

    This is a personal item size bag for under the seat. The max size on Ryanair is 24 liters. You are thinking of the cabin bag which is more like 44 liters. This Decathlon bag is gre…

  19. comment
    Comment #46639157

    I like this question because I come at it from a very different lifestyle. I’m a digital nomad and I have mostly lived out of a backpack and carry on for the past 10 years. My phil…

  20. comment
    Comment #46373041

    It’s best to think about this as angular resolution. Even a very small screen could take up an optimal amount of your field of view if held close. You get the max benefit from a 4k…

  21. comment
    Comment #46368966

    Thats a waste of image quality for most people. You have to sit very close to a 4k display to be able to perceive the full resolution. On PC you could be 2 feet from a huge gaming …

  22. comment
    Comment #46110649

    The higher token output is not by accident. Certain kinds of logical reasoning problems are solved by longer thinking output. Thinking chain output is usually kept to a reasonable …

  23. comment
    Comment #46038391

    A really great way to get an idea of the relative cost and performance of these models at their various thinking budgets is to look at the ARC-AGI-2 leaderboard. Opus 4.5 stacks up…

  24. comment
    Comment #45957341

    I really love this piece! I relate to it but it also doesn’t describe me. I’m far more intuitive than this person, though still agree that insights have driven a leveling up of how…

  25. comment
    Comment #45627656

    I’ll borrow ideas from investing: financial independence, diversification and optionality. If you have enough money you can free yourself from the labor market, but you are still d…