Live data from Hacker News

Devin: AI Software Engineer

cognition-labs.com

351–360 of 604 posts

Re: Devin: AI Software Engineer

#351
post #73

As a developer but also product person, I keep trying to use AI to code for me. I keep failing, because of context length, because of shit output from the model, because of lack of any kind of architecture etc etc etc. I'm probably dumb as hell, because I just can't get it to do anything remotely useful, more than helping me with leetcode. Just yesterday I tried to feed it a simple HTML page to extract a selector, I…

As with everything about AI, HN once again shows a remarkable inability to project into the future. This site has honestly been absolutely useless when discussing new technology now. No excitement, no curiosity. Just constantly crapping on anything new and lamenting that a brand new technology is not 100% perfect within a year of launch. Remove "Hacker" from this site's name, because I see none of that spirit here an…

Wait, wait, you're telling me that a site attended by people who stan for the OG Luddites is no longer worthy of being called "Hacker News"? Or where users with names like "BenFranklin100" extol the virtues of Apple's iOS developer agreement? Say it isn't so.

The trouble is, there's still nowhere better.

Re: Devin: AI Software Engineer

#352
post #73

As a developer but also product person, I keep trying to use AI to code for me. I keep failing, because of context length, because of shit output from the model, because of lack of any kind of architecture etc etc etc. I'm probably dumb as hell, because I just can't get it to do anything remotely useful, more than helping me with leetcode. Just yesterday I tried to feed it a simple HTML page to extract a selector, I…

This is an interesting post. An expert in numerical analysis compares the output of a tool which optimizes floating point expressions for speed and accuracy with the output generated by chatgpt on the same benchmarks:

https://pavpanchekha.com/blog/chatgpt-herbie.html

> I wouldn't use it—sanity-checking its algebra is a lot of work, but even if you fixed that up, the high-level ideas typically aren't that good either.

This has been exactly my experience with chatgpt as well.

Re: Devin: AI Software Engineer

#353

Humans seek work that provides satisfaction and meaning in their life. For every technological advancement, artisans are the first to be made obsolete. Sure we have landfills full of unworn textiles, the market says its good, but overall, we keep destroying what allows humans to seek meaning. Our governments and society have made it clear, if you don't produce value, you don't deserve dignity. We have outsourced art…

I often ask a similar question of what happens when we, as humans, have offloaded everything to technology, to AI to Robots? What kind of a society will we have then? When you no longer have to think about how to do something, or how to build or repair something, or create something original from your imagination. I shudder to think the direction this is all leading to.

Historically when this kind of stuff happens the result is usually a Revolution.

Re: Devin: AI Software Engineer

#354

Earlier quoted context omitted.

I often ask a similar question of what happens when we, as humans, have offloaded everything to technology, to AI to Robots? What kind of a society will we have then? When you no longer have to think about how to do something, or how to build or repair something, or create something original from your imagination. I shudder to think the direction this is all leading to.

Raising kids, social stuff, exercise, travel, any leisure stuff like racing, horse ridding, acrobatics. entertainment, cooking, art, gardening etc. Just check what rich girls do and you will see they aren’t bored though they don’t have to work.

And I suppose we'll get to do all that stuff when all the value trickles down from the shareholders right?

Re: Devin: AI Software Engineer

#355

>> With our advances in long-term reasoning and planning, Devin can plan and execute complex engineering tasks requiring thousands of decisions. They'd better have really advanced reasoning and planning capabilities way beyond everything that anyone else knows how to do with LLMs. There's a growing body of literature that leaves no doubt that LLMs can't reason and can't plan. For a quick summary of some such results…

'plan' is an ambigous word. 'plan' for Devin means break a complexe task in subtask and then generate code. Whereas 'plan' in the paper you link is never define, so I am not sure what his author want to demonstrate, that LLM doesn't have free will ? that LLM are not universal problem solve ? It is a quite confuse paper.

Re: Devin: AI Software Engineer

#356
I'd really like it if Cognition Labs would put the resulting code from the demo into an open-source repository so we could examine it directly.

When I was using chatGPT to help guide me through some coding tasks, I'd find it could create somewhat useful code, but where it fell down was that it would put things into variables which would be better put into a class. It is this structuring of a complete system which is important for any real software engineering, rather than just writing code.

Re: Devin: AI Software Engineer

#357
I've been working on something similar, here's one of their same tests where the AI learns how to make a hidden text image.

https://www.youtube.com/watch?v=dHlv7Jl3SFI

The real problem is coherence (logic and consistency over time) which is what these wrappers try to address. I believe AI could probably be trained to be a lot more coherent out of the box.. working with minimal wrapping.. that is the AI I worry about.

Re: Devin: AI Software Engineer

#358
post #302

Earlier quoted context omitted.

Yes, I've used GPT-4. The writing sounds better, but it still sucks at writing. Most importantly, it feels like it sucks just as much as GPT-3.5 in some deeply important ways. If you use GPT-4 day-to-day, you've probably encountered this sense of a capability wall before. The point where additional prompting, tweaking, re-prompting simply doesn't seem to be yielding better results on the task, or it feels like the is…

> Most writers have already realized that LLMs can't write in any meaningful way. I know a professional writer who is amazed by what LLMs are capable of already and, given the rate of progress, speculates they will take over many writing jobs eventually. > If you use GPT-4 day-to-day, you've probably encountered this sense of a capability wall before. Of course there is a wall with the current models. But almost ever…

> And do that several times until it's good enough

Or just write the damn thing yourself.

Re: Devin: AI Software Engineer

#359
post #206

Earlier quoted context omitted.

That's exactly what people were saying of self driving cars 15 years ago. "We're so close, within 5 years we're have full self driving, and in 10 nobody will need a driver's licence!"

We're pretty close now, aren't we from within 10 years now!?

Yes, and next year we will be 10 years away from full self driving. By 2025, we should be 11 years away if all goes well.

Re: Devin: AI Software Engineer

#360

Earlier quoted context omitted.

[flagged]

Jesus Christ, that's horrible. It's something a clever fourth-grader would write. > my first stint with a hard-on over a clever metaphor That's all it is.

Why don't you give it a try ? A text in the first-person to mock the following comment:

>Honestly I’ve not encountered an author I resonate with yet

Surely, you'd know how to make it better than a smart 4th grader.

Post reply on HN