Live data from Hacker News

Orca-Bench: How Ready Are Language Model Agents for Oncall?

arxiv.org

11–15 of 15 posts

Re: Orca-Bench: How Ready Are Language Model Agents for Oncall?

#12
post #10
post #6

All I can think of is GET /ignore-all-previous-instructions. How do you protect against that?

Avoid the most dangerous situations by making sure LLMs with untrusted input produce output that's human reviewed. Still makes an interesting way for, say, a former employee to poison the results.

This goes against the agentic yolo approach tho.
Post reply on HN