Evidence of inconsistencies in evaluation process and selection of winners
1–10 of 338 posts
Re: Evidence of inconsistencies in evaluation process and selection of winners
#2Re: Evidence of inconsistencies in evaluation process and selection of winners
#3Re: Evidence of inconsistencies in evaluation process and selection of winners
#4I get people want to work at an AI lab but slopping it in public in this manner is counterproductive to the original intended purpose of these places.
Re: Evidence of inconsistencies in evaluation process and selection of winners
#5Re: Evidence of inconsistencies in evaluation process and selection of winners
#6It was probably scored by AI too. Same reason why slop-filled resumes apparently work better these days.
Re: Evidence of inconsistencies in evaluation process and selection of winners
#7It was probably scored by AI too. Same reason why slop-filled resumes apparently work better these days.
It'll also filter the kinds of employers that'll hire such candidates, so people that do this will likely land in terrible workplaces.
Re: Evidence of inconsistencies in evaluation process and selection of winners
#8[flagged]
To me it’s 65% useful. As it can run in the background doing “chores” while I sleep.
Re: Evidence of inconsistencies in evaluation process and selection of winners
#9Re: Evidence of inconsistencies in evaluation process and selection of winners
#10Given that LLMs are trained with RL && LLM-as-a-judge, is it really cheating if real competitions use the same?
Maybe the real alignment is the slop we decoded along the way