Viewing profile — Leynos
Leynos
HN member- Joined
- Mon, Sep 27, 2010, 7:36 AM UTC
- HN karma
- 997
- Public activity
- 549 items
- HN profile
- View on Hacker News ↗
About Leynos
https://github.com/leynos
meet.hn/city/55.9533456,-3.1883749/City-of-Edinburgh
Recent public activity
-
comment
Comment #49154894
Is this using the server side compaction of the responses API?
-
comment
Comment #49124845
1st line of defence, use something like ponytail to enforce brevity. Use property testing and behavioural testing on top of unit testing. Enforce readability standards so you will …
-
comment
Comment #49124716
If you ask Fable or 5.6 Sol to improve performance, it will generally know to build a benchmark and create a test corpus. I'm not sure where the contrary suggestion is coming from.…
-
comment
Comment #49060619
The problem with Empire of AI is that it takes such a scattergun approach and doesn't really build a coherent thesis. It is also very difficult to draw clear directional informatio…
-
comment
Comment #49057752
Zerover for life
-
comment
Comment #48975483
[flagged]
-
comment
Comment #48889282
I'd prefer a tag to the mounds of "this looks like it is AI generated, I can tell from the pixels and from having seen quite a few AIs in my time" comments. That way the people who…
-
comment
Comment #48889255
You'd need the whole edit tree along with all the prompts used along the way, which most people are not yet set up to capture.
-
comment
Comment #48888779
For normal building work I use Opus to plan and GPT 5.6 Terra to build. The point is, these are not normal constrained building tasks. Perhaps I should have had more faith in Opus'…
-
comment
Comment #48883867
The sort of thing Fable and Sol excel at are long horizon tasks. The sort of thing I have been using them for is migrating large numbers of repositories to new tooling simultaneous…
-
comment
Comment #48879837
Things that I reckon will become a lot more important from a developer's perspective over the next year: - Shaping work so it is more decomposable, legible, verifiable and understa…
-
comment
Comment #48876615
Yeah, I went through a period after 4.7 launched of not using Claude for code work at all because of the condescending refusals. (Kept using it for planning and design work). Still…
-
comment
Comment #48874271
While there are some coding focused models (composer, for example), the majority of frontier models are pitched as general purpose. The coding harnesses for Claude and GPT are even…
-
comment
Comment #48869870
From a purely utilitarian standpoint, direct to cell feels like a good thing to me. Large swathes of Scotland don't even have sufficient mobile connection to send a text message (s…
-
comment
Comment #48845635
The prompted response is far from the finished piece of writing. You'd probably want to share the full edit tree and include subsequent refinement prompts in the commit messages.
-
comment
Comment #48843193
Yes, that was the point. It made unsafe behaviour visible in a way that could be addressed. I hadn't heard any reports of it being dysfunctional.
-
comment
Comment #48806171
As a reader (not as someone who is posting the articles), the AI prose generally doesn't bother me. I'm usually more concerned about what the article says than how it says it.
-
comment
Comment #48657082
Samsung are back up 5% today on news of a planned buyback.
-
comment
Comment #48607443
Alerts on test fixtures, so suspect it is doing nothing new.
-
comment
Comment #48598010
You can see the validation approach they used here: https://github.com/adamraudonis/prylint/blob/main/harness/ch...
-
comment
Comment #48597918
Currently, there are things pylint does that ruff doesn't. To use these, I was running pylint on pypy to get it running at a reasonable speed. Having pylint reimplemented in Rust s…
-
comment
Comment #48572980
I said "outside of situations where it is required by contract", which I believe would include a CLA.
-
comment
Comment #48570562
Which model was used for the benchmark results shown on your GitHub README.md?
-
comment
Comment #48556986
Context: https://www.businessinsider.com/what-is-le-chaton-fat-mistra...
-
comment
Comment #48452168
I quite like my mechanical spider from Wild Wild West and the coffee it makes with a 50% success rate