Viewing profile — meffmadd
meffmadd
HN member- Joined
- Fri, Apr 11, 2025, 2:33 PM UTC
- HN karma
- 31
- Public activity
- 23 items
- HN profile
- View on Hacker News ↗
About meffmadd
No profile information was provided.
Recent public activity
-
comment
Comment #49242052
Of course, but that was not part of my workflow and I found it quite strange that this was simply not possible especially when docker-compose can easily do this
-
comment
Comment #49240288
I tried Docker Sandboxes but last time I checked you could not configure custom volume mounts, making more complex setups impossible. For work I need two directories for context fo…
-
comment
Comment #48818840
This doesn’t hold up imo. Your phone analogy fits the description of a harness. That is the thing you interact with. The LLM is more like the processor in that phone. As long as it…
-
comment
Comment #48356122
That is very true but I was surprised by how clear the “signal” was. Only Gemini really confidently solved all levels. But yeah the goal is now to include harder levels as well!
-
comment
Comment #48355759
Yes that was the post that inspired me to build this. While I did implement a more comprehensive harness with path finding tools etc. the models themselves have improved significan…
-
comment
Comment #48355302
I found LLMs to be surprisingly good at puzzle games like Baba Is You: https://meffmadd.github.io/samplesurium/posts/baba_is_agent/
-
comment
Comment #48179527
…or basically the plot of twin peaks
- story
-
comment
Comment #47590130
I think this view assumes no human will/should ever read the code. This is considered bad practice because someone else will not understand the code as well whether written by a hu…
-
comment
Comment #47482297
Genuine question to people more knowledgeable: Why are politicians/technocrats doing this? Also generally speaking e.g. in relation to chat control and so on. Do they think this is…
-
comment
Comment #47163579
I am amazed that the IDM is able to produce enough high quality annotations for the downstream FDM to work, even matching the ground truth contractor annotations!
-
comment
Comment #47153498
As an EU citizen I really hope we can gain some meaningful distance to the US asap. I hope my leaders feel the same. And if everything works out I think this will be great for the …
-
comment
Comment #47136404
I really wonder what motivates a seven year old to persistently work on that „one thing“ and not get distracted/bored. I guess he knew he was special?
-
comment
Comment #47035374
Have you ever used an open model for a bit? I am not saying they are not benchmaxxing but they really do work well and are only getting better.
-
comment
Comment #47005375
Here is what I don’t get tho: you have UX designers/engineers creating a new interface. What do you tell them? Just to do whatever? They probably spent months designing the new int…
-
comment
Comment #46977849
It will be tough to run on our 4x H200 node… I wish they stayed around the 350B range. MLA will reduce KV cache usage but I don’t think the reduction will be significant enough.
-
story
Ask HN: What is the most complicated Algorithm you came up with yourself?
Hi HN, what was the most complex and sophisticated algorithm you came up with yourself while working on a problem?
-
comment
Comment #46280398
From what I have heard „Bodycam“ uses scans of actual locations for its maps.
-
comment
Comment #46164456
Yeah true, but also this is a bit like saying the lock screen of your phone should not become a "one stop shop" for all push notifications. I actually do not own an Apple TV but I …
-
comment
Comment #46160887
Wow that is quite anti-consumer! Surely a monopoly on streaming will help them realize this. /s
- story
-
comment
Comment #45485487
As an EU citizen hosting LLMs for researchers and staff at the university I work at, this is hits home. Without Chinese models we could not do what we do right now. IMO, in the EU …
-
story
Show HN: Aqueduct AI Gateway – A Self-Hostable AI Gateway Without the "Taxes"
Hey HN! I’m excited to share Aqueduct, an open-source AI gateway you can self-host to manage LLM traffic without hitting paywalls for basic features. I'm part of a small team at TU…