Live data from Hacker News

Viewing profile — Rutledge

Rutledge

HN member
Joined
Thu, Mar 15, 2012, 7:27 AM UTC
HN karma
184
Public activity
42 items

About Rutledge

No profile information was provided.

Recent public activity

  1. comment
    Comment #48779162

    Would love to see the 'contract'

  2. comment
    Comment #47607127

    Scorecard AI | Founding Engineer | San Francisco (Onsite) | Full-Time | https://scorecard.io Scorecard builds simulation environments and reward models that frontier AI labs and en…

  3. story
  4. story
  5. comment
    Comment #46470676

    Scorecard | Founding Engineers & GTM | San Francisco | ONSITE | Full-time | scorecard.io Scorecard is the simulation platform for self-improving AI agents. We help teams encode exp…

  6. comment
    Comment #45802190

    Scorecard | Founding Engineer, Founding UX Designer, Founding GTM | SF, CA ONSITE | Full-time Scorecard is building the leading platform for testing, evaluating, and monitoring AI …

  7. story
    Show HN: Scorecard – Evaluate LLMs like Waymo simulates cars

    Hey HN! I built self-driving sim and eval at Waymo. Now I’m building Scorecard to bring that approach to agent eval: reproducible, automated scoring for AI. Scorecard lets you: - R…

  8. comment
    Comment #45422770

    I call them 'CLI agents'!

  9. comment
    Comment #44374187

    Here's the image from Wayback: https://web.archive.org/web/20250625051706/https://blog.goog... The biggest diffs from Claude code (the current champion): 1. Generous free tier (60 …

  10. comment
    Comment #44139494

    Aannnnndd X is down x) Here's the LI: https://www.linkedin.com/posts/scorecard-ai_introducing-scor...

  11. comment
    Comment #44139476

    Hi HN- we're excited to launch the first remote MCP server for claude.ai and cursor for LLM evaluation. Would love your thoughts and feedback :)

  12. story
  13. comment
    Comment #43297283

    Here's the repo: https://github.com/agntcy and docs: https://docs.agntcy.org/pages/abstract.html

  14. story
  15. comment
    Comment #43191630

    This initiative is designed to be community-driven, so we're looking forward to your feedback on what agent benchmarking needs exist in your domains. While starting with legal AI, …

  16. story
  17. comment
    Comment #43098964

    Yes quite helpful- thanks for explaining and will try it out!

  18. comment
    Comment #43076798

    The concurrent request handling seems great for our AI eval workloads, where we're waiting for LLM API calls and DB operations but curious how Vercel handles potential noisy neighb…

  19. comment
  20. comment
    Comment #42443156

    This is great :) and pretty impressive that it was possible in coda!

  21. comment
    Comment #42443008

    Post from the Coda blog: https://coda.io/blog/about-coda/grammarly-acquires-coda

  22. comment
    Comment #42442986

    New chapter in the AI arms race

  23. story
  24. comment
    Comment #39506093

    +1 on data labeling platform: https://web.archive.org/web/20230403164757/https://feather.o... It's been around and used since 2022. It's an site for SME to write code data: https:/…

  25. comment
    Comment #38935569

    ChatGPT now learns about users with a RAG system. This is the first step towards an OpenAI assistant: https://help.openai.com/en/articles/8590148-memory-in-chatgp...