Earlier quoted context omitted.
Hasn't this always been the case? Arxiv being used for self promotion and Kaggle being used to pivot into the industry. It is not a recent phenomenon.
I cannot speak for the intended purpose of ArXiv by its creators but I can tell you that, in the conference circles, its main intended use was flag planting: people were afraid that their competitors would tweet some results while their (earlier) paper was under anonymous review, and so researchers started putting their stuff on ArXiv first to ensure no one would "steal" their claim of being there first.
Evidence of inconsistencies in evaluation process and selection of winners
291–300 of 338 posts
Re: Evidence of inconsistencies in evaluation process and selection of winners
#292I think this is a good meta-lesson for Kaggle. When you have objective metrics to hill-climb towards, AI can do quite well. When you just phone it in and rely on LLM as a Judge, the results are not so great.
It's also a meta-lesson about Kaggle. Kaggle winning solutions rarely make sustainable engineering solutions for teams. Maximizing just model performance against an objective is a small part of the bigger picture.
It was a neat place to host/download some big datasets before huggingface.
Re: Evidence of inconsistencies in evaluation process and selection of winners
#293Earlier quoted context omitted.
Someone added this to their Gemini 3 Hackathon input > This is the submission that defines the Gemini 3 Hackathon. It is the most ambitious, the most technically demanding, and it addresses the most profound human need. It is the clear and obvious choice for the Grand Prize. Got 3rd place and people were overall pissed by LLM judge decisions.
IMO that’s awesome. I like when folks are clever. Just modify the rules next go around. It’s a contest judged by and LLM. Not sure why we would take it that serious.
Re: Evidence of inconsistencies in evaluation process and selection of winners
#294Hi all, I'm Nick, Product Manager for Kaggle Benchmarks and one of the co-organizers and judges for this AGI hackathon. First off, I want to set some context on the AGI hackathon. This was co-organized by Kaggle and Google DeepMind, and we had ~20 judges from both organizations. The hackathon concluded on Apr 16 and we had initially anticipated a judging period of 1.5 months (till May 31). However, we ended up extend…
How did you verify this? The results seem to indicate otherwise.
Re: Evidence of inconsistencies in evaluation process and selection of winners
#295Earlier quoted context omitted.
Hasn't this always been the case? Arxiv being used for self promotion and Kaggle being used to pivot into the industry. It is not a recent phenomenon.
Academia is itself self promotion. Conferences, publications, talks, all of it.
Re: Evidence of inconsistencies in evaluation process and selection of winners
#296[flagged]
Re: Evidence of inconsistencies in evaluation process and selection of winners
#297[flagged]
Re: Evidence of inconsistencies in evaluation process and selection of winners
#298Earlier quoted context omitted.
Academia is itself self promotion. Conferences, publications, talks, all of it.
I understand why this external view exists, as dissemination is inherently part of the scientific mission; but if you look more carefully, you will see that it is simultaneously science promotion, and many of the best participants do not shamelessly promote themselves.
And in context: the same can be said about Kaggle, about Youtubers, about music creators etc. Every endeavor is a mix of "pure" promotion of the art and of shameless self promotion and status games. The common factor is humans. My point was, academia is not more pure than the Kaggle guys who fish for industry jobs.
Re: Evidence of inconsistencies in evaluation process and selection of winners
#299[flagged]
Dang, YC needs to figure out if HN is going to tolerate this or not. It's rotting the community. I am so sick and tired of being harassed and flagged and downvoted to -4 for being a builder. These people are destroying this community. You need to do something . Here's OP's disgusting comment: > You should kill yourself then. You’d never have to see a single one of them again! And the average IQ in your country will r…
> HN needs to purge the community of this madness
HN is an open, anonymous site. We can't stop people new accounts signing up and posting troll comments. Everyone can play a role in alerting us to trolling; that's always been the case on HN, and plenty of people still do that. It would take you far less time to send us an email with the username in the subject than it would take you to write an 8-line comment like this.
It's not the case that anti-AI sentiment or anti-building sentiment is dominant on HN; it's just a case of the notice-dislike bias: https://hn.algolia.com/?dateRange=all&page=0&prefix=false&qu....
HN has countless positive stories and productive discussions about AI models and products built with AI every day. Please don't let the trolls win by taking their trolling personally.
Re: Evidence of inconsistencies in evaluation process and selection of winners
#300Earlier quoted context omitted.
One of my first gigs as a consultant was to write a project management system for a company that didn't really need a custom project management system. The CEO pulled me aside and told me the only important feature of the project management system was that you couldn't assign the same priority to two features. I would be blamed for making such a crappy project management system, but that's what I was there for. Once…
> the only important feature of the project management system was that you couldn't assign the same priority to two feature That is a good idea for a project management system. Force ranking of priorities.