Viewing profile — emregucerr
emregucerr
HN member- Joined
- Fri, Aug 25, 2023, 8:03 AM UTC
- HN karma
- 31
- Public activity
- 20 items
- HN profile
- View on Hacker News ↗
About emregucerr
No profile information was provided.
Recent public activity
-
story
Show HN: Artificial Intelligence Squared – LLMs Debate Each Other
I built this fun benchmark to pitch LLM models against each other in Oxford-style debate. The format is inspired by Intelligence Squared. The side who flips most votes win.
-
comment
Comment #47656889
I would love to see someone build it as some kind of an SDK. App builders could use it as a local LLM plugin when dealing with data involving sensitive information. It's usually to…
-
comment
Comment #47383743
Did they pick the word "uncrewed" to not use the word "unmanned"? If so, I'm not hopeful. Might be another EuroDrone disaster.
-
comment
Comment #47135079
People dismiss this as a meme too quick but I think this is a good thought experiment not only for drawing a comparison for energy consumption but learning efficiency. AI is often …
-
comment
Comment #46511106
> I don’t think the gates should animate up into the air. I agree! It feels off compared to the overall aesthetic of the game. Awesome game though! Loved it.
- story
-
story
Show HN: I built a tool to convert Loom videos into Playwright tests
Hi HN, my name is Emre. I am one of the co-founders of Fume - a tool to generate Playwright test suites from product walkthrough videos. HOW DOES IT WORK 1. Upload a Loom or equiva…
-
comment
Comment #43090622
I think most people blindly try to get controls in Vanta/Drata to pass like us. I'd much rather build a dumb bot than having to talk to my auditor. But still > we realize how stupi…
-
comment
Comment #43090097
Hey HN! We recently got our SOC2 certification. One thing that really annoys us is having to get all deployment PRs approved by at least one person per guidelines. This might not s…
- story
-
comment
Comment #42966766
i wonder how good is R1 at counting pixels from a screenshot. what enabled claude and OAI's CUA to develop computer use was being able to precisely give x-y coordinates of a click …
-
story
Show HN: Fume (YC W24) – bot to test all of your pull requests
Hi HN! We are building an AI agent that can use a computer for software tasks. You can have it review and manually test all your pull requests. Here is a demo video of how it works…
-
comment
Comment #42460692
why specifically books? i never was able to get real value out of them in the case of programming.
-
comment
Comment #42460673
imo, the more senior counterpart benefits less from pair programming and therefore enjoys less. however, it's still the fastest way to get someone familiarized with a concept/proje…
-
comment
Comment #41466461
Am I missing something with the latest buzz around the 'founder mode'? It's a new concept for startups that transitioned into an enterprise. Not seed/series A stage startups. In th…
- story
-
comment
Comment #41368642
> The best lens for future performance of large models is uncertainty. 100% agree. I think to better way to phrase my argument there would be to reject the notion that LLMs are des…
-
comment
Comment #41368570
> Why not? The 'original' part is more important than the 'beautiful' part - which should have been more clear in my writing. This argument also triggers the question "is true orig…
- story
- story