Viewing profile — gsandahl
gsandahl
HN member- Joined
- Fri, Jan 20, 2023, 8:36 AM UTC
- HN karma
- 14
- Public activity
- 15 items
- HN profile
- View on Hacker News ↗
About gsandahl
Recent public activity
-
comment
Comment #47514667
Agree, this is where llms can uncover new perspectives!
-
comment
Comment #47509457
Oh lord, imagine asking ”serious” questions https://opper.ai/ai-roundtable/questions/you-are-standing-in...
-
comment
Comment #44836671
Most of the tasks have assessed with ground truth, occasionally helped with an LLM as a judge to assess the answer if the answer is a sentence and not an exact result. Example: Giv…
-
comment
Comment #44836649
We are running task specific benchmarks across a number of categories (agentic tasks, context tasks, normalization tasks etc), and on our benchmarks we see Gpt-5 rating slightly be…
-
comment
Comment #44573532
Please do and give us some feedback!
-
comment
Comment #44573527
I think just how far you can go with examples has been an interesting learning! As these models have become smarter, they are also getting better at reasoning from examples and und…
-
comment
Comment #44571099
No up to date demo video unfortunately :( Sounds like a great use case though!
-
comment
Comment #44571091
We have been thinking a bit about this, and one option would be to have some form of locally hosted runner. You can optimize the task in the cloud and deploy it locally. Something …
-
comment
Comment #44571004
Yes that's possible! You can populate examples of great outputs to task specific datasets and have those be automatically populated to the prompt. More info here: https://docs.oppe…
-
comment
Comment #44570977
Thanks for the shout out!
-
comment
Comment #44570737
Co-Founder here thanks for taking a look at Opper! I’m hanging around the thread all day, so feel free to ask anything, share feedback, or tell us where you’d like the product to g…
- story
-
comment
Comment #43045726
Its on that trajectory at least :)
-
comment
Comment #43035290
[dead]
- story