Viewing profile — oliver_dr
oliver_dr
HN member- Joined
- Sun, Mar 08, 2026, 8:33 PM UTC
- HN karma
- 1
- Public activity
- 28 items
- HN profile
- View on Hacker News ↗
About oliver_dr
Recent public activity
-
comment
Comment #47343892
The framing of "benchmarks measure capability, we measure reliability" resonates. The industry has been so focused on making agents more capable that reliability infrastructure has…
-
comment
Comment #47343228
Nice approach. The "don't trust an LLM to guard another LLM" principle is sound for tool-call safety specifically, where the threat model is well-defined (destructive commands, SSR…
-
comment
Comment #47343212
This is almost certainly a server-side session isolation bug rather than an LLM hallucination - the model is returning a response to someone else's prompt. The DeepSeek cloud endpo…
-
comment
Comment #47343196
We've been dealing with this at multiple layers. Here's what actually works in production: Input-side (preventing injection): - Strict input sanitization with role-boundary enforce…
-
comment
Comment #47337445
[dead]
-
comment
Comment #47337437
[dead]
-
comment
Comment #47336307
[dead]
-
comment
Comment #47334718
[dead]
-
comment
Comment #47334599
[dead]
-
comment
Comment #47334368
[dead]
-
comment
Comment #47334092
[dead]
-
comment
Comment #47333649
[dead]
-
comment
Comment #47333133
[dead]
-
comment
Comment #47333125
[dead]
-
comment
Comment #47330790
[dead]
-
comment
Comment #47328379
[dead]
-
comment
Comment #47327608
[dead]
-
comment
Comment #47327213
[dead]
-
comment
Comment #47327196
[dead]
-
comment
Comment #47327186
[dead]
-
comment
Comment #47308730
[dead]
-
comment
Comment #47308083
[dead]
-
comment
Comment #47308078
[dead]
-
comment
Comment #47308062
[dead]
-
comment
Comment #47308053
[dead]