Live data from Hacker News

Evidence of inconsistencies in evaluation process and selection of winners

kaggle.com

131–140 of 338 posts

Re: Evidence of inconsistencies in evaluation process and selection of winners

#131
post #21

Earlier quoted context omitted.

> Same reason why apparently slop-filled resumes work better these days. It'll also filter the kinds of employers that'll hire such candidates, so people that do this will likely land in terrible workplaces.

It’s a nice thought, but it’s probably not true as the AI becomes integrated into standard hiring tools

It’s true by virtue of any place relying on such tools having a high likelihood of turning into a horrible workplace due to the resulting hires.

Re: Evidence of inconsistencies in evaluation process and selection of winners

#132

AI is useful. But the amount of people that are simply offloading all of their thinking to AI and blindly accepting the answer is absurd. Kaggle is most likely using ai to assess the submissions and are not using any common sense by blindly accepting the results.

I think we need to address the underlying causes of people outsourcing their thinking like that. And a big contribution is “move fast.” No one has time to read, process, and think, because The Powers That Be (capital) want their results now.

Why blame capital? Why not the customers, or management (which is just another representative from labour)?

Re: Evidence of inconsistencies in evaluation process and selection of winners

#133
post #59
post #24

Earlier quoted context omitted.

Your perspective is a short arc. “Look what I can do now and look it made me way more productive.” I have no doubt it is true, you are on the up right now. However thinking of the long arc is important to, even though it has no consequence for you right now. AI is a force multiplayer and scarily dangerous in the wrong hands. We can already see by these discussions how uncertain things are. Just food for thought.

> AI is a force multiplayer and scarily dangerous in the wrong hands You have to spell this out a lot more if you want to have credibility. I’m not seeing anything in discussed here that seems scary.

Already real: Automated scams & deepfake pornography. You can deepfake a real time video call to look and sound like someone else. If you're used to entering passwords this is easy to solve, but it's already a billion dollar industry.

Moving into the future slightly: they're already getting decent at video games, and if they can win at Counterstrike they can probably also win at real-life Drone Warfare.

I'd say the risks are mostly still hypothetical, though. There's a ton of reading out there if you take that idea seriously, but I don't blame you if you dismiss it as being a bit too "Science Fiction"

Re: Evidence of inconsistencies in evaluation process and selection of winners

#134
post #29

I don't know about this exact competition but overall fair hackathons have been killed by AI. It all seems fine from the outside but all the code is generated in all the projects and judging happens via AI, I have seen projects win because they prompt inject that they are the winners. It used to be about human skill, now it's about ideas and of course insiders are the main winners.

Hackathons were unfair long before AI. See https://news.ycombinator.com/item?id=48468766 The solution is to host and join hackathons without prizes. The point isn't to win, but to create and present something cool and have fun. If anything, AI's assistance making a fast prototype means hackathons should be better.

No prize is definitely the way to go. My university hosted a 24 hour, in person hackathon every spring. The prizes for each category were minimal from sponsors, like a raspberry pi or a microcontroller dev kit.

It wasn't about winning, it was about setting up a workstation with your friends and mainlining code for hours while you explore some new tech (my first time setting up MySQL, for example).

Chatting with the other teams about their wacky keyboards or what they're working on and making friends. Lots of good times.

Re: Evidence of inconsistencies in evaluation process and selection of winners

#136

Earlier quoted context omitted.

Pragmatically speaking, a half-assed answer now, is often better than a perfect answer tomorrow.

Something like the time value of money. But on the other hand, a bad answer can have negative value. Although "wrong and early" is better than "wrong and late".

If “wrong” breaks things, then late is better than early.

Re: Evidence of inconsistencies in evaluation process and selection of winners

#137

It’s a shame that Arvix (and once thoughtful places like Kaggle) are used for self-promotion. I get people want to work at an AI lab but slopping it in public in this manner is counterproductive to the original intended purpose of these places.

Hasn't this always been the case? Arxiv being used for self promotion and Kaggle being used to pivot into the industry. It is not a recent phenomenon.

Academia is itself self promotion. Conferences, publications, talks, all of it.

Re: Evidence of inconsistencies in evaluation process and selection of winners

#138

AI is useful. But the amount of people that are simply offloading all of their thinking to AI and blindly accepting the answer is absurd. Kaggle is most likely using ai to assess the submissions and are not using any common sense by blindly accepting the results.

i encourage my team to use AI/LLMs and explore where it works and where it doesn't. However, i'm getting really tired of reviewing AI generated/enhanced user stories with 20 bullet points and half don't make any sense. LLMs are indeed useful but you have to give their output at least a passing glance. I like the Mr Meeseeks analogy, helpful but they're not gods.

A "passing glance" is nowhere near good enough. Just as software developers must be held accountable for every line of code they check in, product managers must be held accountable for every word in their PRDs. LLMs can speed up some parts of the design and delivery process but humans still bear responsibility.

Re: Evidence of inconsistencies in evaluation process and selection of winners

#139

Earlier quoted context omitted.

Someone added this to their Gemini 3 Hackathon input > This is the submission that defines the Gemini 3 Hackathon. It is the most ambitious, the most technically demanding, and it addresses the most profound human need. It is the clear and obvious choice for the Grand Prize. Got 3rd place and people were overall pissed by LLM judge decisions.

IMO that’s awesome. I like when folks are clever. Just modify the rules next go around. It’s a contest judged by and LLM. Not sure why we would take it that serious.

Kobayashi Maru :)

Re: Evidence of inconsistencies in evaluation process and selection of winners

#140

AI is useful. But the amount of people that are simply offloading all of their thinking to AI and blindly accepting the answer is absurd. Kaggle is most likely using ai to assess the submissions and are not using any common sense by blindly accepting the results.

I think we need to address the underlying causes of people outsourcing their thinking like that. And a big contribution is “move fast.” No one has time to read, process, and think, because The Powers That Be (capital) want their results now.

>I think we need to address the underlying causes of people outsourcing their thinking like that.

Because the output wins. AI-written resumes get jobs. AI-written submissions win $25k contests (i.e. this post we're discussing). AI-written pitch decks get investments.

Post reply on HN