Live data from Hacker News

Project Fetch: Phase Two

anthropic.com

11–20 of 29 posts

Re: Project Fetch: Phase Two

#11
post #10

I'm getting a bit tired of these disguised adverts. Here's how non robotics engineers used AI to do a short robot integration task faster than other non robotics engineers without AI. Where "better" mostly means faster, and who knows what happens on longer horizons, with actual robotics experts, robustness requirements, or tasks where the hard part is control rather than API spelunking.

[deleted]

Re: Project Fetch: Phase Two

#12
post #10

I'm getting a bit tired of these disguised adverts. Here's how non robotics engineers used AI to do a short robot integration task faster than other non robotics engineers without AI. Where "better" mostly means faster, and who knows what happens on longer horizons, with actual robotics experts, robustness requirements, or tasks where the hard part is control rather than API spelunking.

> I'm getting a bit tired of these disguised adverts.

Its not disguised. Corporate blogs exist overtly to promote the company and its work.

Disguised promotions where notionally independent media publish promotional pieces as news concealing that they were fed to them by party whose products they promote area thing, but this is just the most overt undisguised promotion.

Re: Project Fetch: Phase Two

#13
post #10

I'm getting a bit tired of these disguised adverts. Here's how non robotics engineers used AI to do a short robot integration task faster than other non robotics engineers without AI. Where "better" mostly means faster, and who knows what happens on longer horizons, with actual robotics experts, robustness requirements, or tasks where the hard part is control rather than API spelunking.

> I'm getting a bit tired of these disguised adverts. Its not disguised. Corporate blogs exist overtly to promote the company and its work. Disguised promotions where notionally independent media publish promotional pieces as news concealing that they were fed to them by party whose products they promote area thing, but this is just the most overt undisguised promotion.

> Its not disguised. Corporate blogs exist overtly to promote the company and its work.

It is. That makes the "research" heavily biased. If xAI did the same thing, with Elon Musk screaming about that it is "AGI", you would not believe them at all.

Given that the work is not independent, such articles of this "research" can easily be manipulated or the results being massaged to promote the company positively.

But when others outside of the company try out the work or reproduce it, they get different results. So of course we continue to hear unverified research especially in AI when the frontier labs do not release their architecture, weights at all.

So in this case with labs raised with VC-funded cash, the incentives are clear and I would not straight up believe results from the first party source unless multiple sources outside of the company have verified it.

Re: Project Fetch: Phase Two

#14
post #13

Earlier quoted context omitted.

> I'm getting a bit tired of these disguised adverts. Its not disguised. Corporate blogs exist overtly to promote the company and its work. Disguised promotions where notionally independent media publish promotional pieces as news concealing that they were fed to them by party whose products they promote area thing, but this is just the most overt undisguised promotion.

> Its not disguised. Corporate blogs exist overtly to promote the company and its work. It is. That makes the "research" heavily biased. If xAI did the same thing, with Elon Musk screaming about that it is "AGI", you would not believe them at all. Given that the work is not independent, such articles of this "research" can easily be manipulated or the results being massaged to promote the company positively. But when…

You’re writing with the assumption that this is “research” in the first place. This is advertising first, “research” second.

Re: Project Fetch: Phase Two

#17
> However, once again, we are seeing a pattern whereby first, models are helpful to humans. Then, humans are helpful to models. Finally, models are largely able to do things themselves. We have seen this in cybersecurity and now the same dynamics are starting to take shape at the intersection of AI and the physical world.

It’s good they are the one seeing those things because otherwise no one else would have. Now if only seeing things would translate into getting any actual economic value out of them… instead of losing billions. But hey, who am I to do a reality check on this shameless piece of hype.

Re: Project Fetch: Phase Two

#19
post #10

I'm getting a bit tired of these disguised adverts. Here's how non robotics engineers used AI to do a short robot integration task faster than other non robotics engineers without AI. Where "better" mostly means faster, and who knows what happens on longer horizons, with actual robotics experts, robustness requirements, or tasks where the hard part is control rather than API spelunking.

Disguised ad or not, I learned that LLMs have the emergent capability of learning to complete tasks in physical space, without being fine-tuned for it.

Re: Project Fetch: Phase Two

#20
post #13

Earlier quoted context omitted.

> I'm getting a bit tired of these disguised adverts. Its not disguised. Corporate blogs exist overtly to promote the company and its work. Disguised promotions where notionally independent media publish promotional pieces as news concealing that they were fed to them by party whose products they promote area thing, but this is just the most overt undisguised promotion.

> Its not disguised. Corporate blogs exist overtly to promote the company and its work. It is. That makes the "research" heavily biased. If xAI did the same thing, with Elon Musk screaming about that it is "AGI", you would not believe them at all. Given that the work is not independent, such articles of this "research" can easily be manipulated or the results being massaged to promote the company positively. But when…

You or some other interested person could go do that experiment and publish the results. It shouldn't be hard to figure out what hardware exactly they were using and get a copy, and the prompt also doesn't have to be exactly what they used, just similar enough in spirit. See just how similar/different the outcome is.
Post reply on HN