Viewing profile — Rutledge
Rutledge
HN member- Joined
- Thu, Mar 15, 2012, 7:27 AM UTC
- HN karma
- 184
- Public activity
- 42 items
- HN profile
- View on Hacker News ↗
About Rutledge
No profile information was provided.
Recent public activity
-
comment
Comment #48779162
Would love to see the 'contract'
-
comment
Comment #47607127
Scorecard AI | Founding Engineer | San Francisco (Onsite) | Full-Time | https://scorecard.io Scorecard builds simulation environments and reward models that frontier AI labs and en…
- story
- story
-
comment
Comment #46470676
Scorecard | Founding Engineers & GTM | San Francisco | ONSITE | Full-time | scorecard.io Scorecard is the simulation platform for self-improving AI agents. We help teams encode exp…
-
comment
Comment #45802190
Scorecard | Founding Engineer, Founding UX Designer, Founding GTM | SF, CA ONSITE | Full-time Scorecard is building the leading platform for testing, evaluating, and monitoring AI …
-
story
Show HN: Scorecard – Evaluate LLMs like Waymo simulates cars
Hey HN! I built self-driving sim and eval at Waymo. Now I’m building Scorecard to bring that approach to agent eval: reproducible, automated scoring for AI. Scorecard lets you: - R…
-
comment
Comment #45422770
I call them 'CLI agents'!
-
comment
Comment #44374187
Here's the image from Wayback: https://web.archive.org/web/20250625051706/https://blog.goog... The biggest diffs from Claude code (the current champion): 1. Generous free tier (60 …
-
comment
Comment #44139494
Aannnnndd X is down x) Here's the LI: https://www.linkedin.com/posts/scorecard-ai_introducing-scor...
-
comment
Comment #44139476
Hi HN- we're excited to launch the first remote MCP server for claude.ai and cursor for LLM evaluation. Would love your thoughts and feedback :)
- story
-
comment
Comment #43297283
Here's the repo: https://github.com/agntcy and docs: https://docs.agntcy.org/pages/abstract.html
- story
-
comment
Comment #43191630
This initiative is designed to be community-driven, so we're looking forward to your feedback on what agent benchmarking needs exist in your domains. While starting with legal AI, …
- story
-
comment
Comment #43098964
Yes quite helpful- thanks for explaining and will try it out!
-
comment
Comment #43076798
The concurrent request handling seems great for our AI eval workloads, where we're waiting for LLM API calls and DB operations but curious how Vercel handles potential noisy neighb…
- comment
-
comment
Comment #42443156
This is great :) and pretty impressive that it was possible in coda!
-
comment
Comment #42443008
Post from the Coda blog: https://coda.io/blog/about-coda/grammarly-acquires-coda
-
comment
Comment #42442986
New chapter in the AI arms race
- story
-
comment
Comment #39506093
+1 on data labeling platform: https://web.archive.org/web/20230403164757/https://feather.o... It's been around and used since 2022. It's an site for SME to write code data: https:/…
-
comment
Comment #38935569
ChatGPT now learns about users with a RAG system. This is the first step towards an OpenAI assistant: https://help.openai.com/en/articles/8590148-memory-in-chatgp...