Show HN: AgentReady – Drop-in proxy that cuts LLM token costs 40-60%
1–10 of 14 posts
Re: Show HN: AgentReady – Drop-in proxy that cuts LLM token costs 40-60%
#2It uses a deterministic rule-based engine (not another LLM call): removes filler words, simplifies verbose constructions, strips redundant connectors. ~5ms overhead.
Works with any OpenAI-compatible SDK: Python, Node, LangChain, LlamaIndex, CrewAI, Vercel AI SDK.
Free during beta, no credit card: https://agentready.cloud/hn
Python: pip install agentready-sdk && agentready init
Happy to answer any technical questions.
Re: Show HN: AgentReady – Drop-in proxy that cuts LLM token costs 40-60%
#3Re: Show HN: AgentReady – Drop-in proxy that cuts LLM token costs 40-60%
#4Re: Show HN: AgentReady – Drop-in proxy that cuts LLM token costs 40-60%
#5Re: Show HN: AgentReady – Drop-in proxy that cuts LLM token costs 40-60%
#6Re: Show HN: AgentReady – Drop-in proxy that cuts LLM token costs 40-60%
#7I'm sure I'm not the only one hesitant to provide a 3rd party virtually MITM access to both my LLM usage + API keys. If this were capable of running locally, or even just an API for compressing non-sensitive parts of a prompt, I think it would be much easier to adopt.
Re: Show HN: AgentReady – Drop-in proxy that cuts LLM token costs 40-60%
#8There's zero percent chance that I would proxy all my LLM calls with my API key through some third party service. However, if it was self-hostable, so that I can ensure it is only able to reach the LLM providers, I could see deploying this behind an LLM provider router. If it actually achieves the kind of token use reduction that is advertised, that would be worth paying for - especially in the enterprise. I'm skepti…
Re: Show HN: AgentReady – Drop-in proxy that cuts LLM token costs 40-60%
#9No. You only need our API key for the compression step. Your LLM keys and usage stay entirely in your own app — we never see them. We receive text, compress it, and return it. Your LLM (local, OpenAI, Claude, or any other) then processes it with your own keys. We don't even know your app's name.
Re: Show HN: AgentReady – Drop-in proxy that cuts LLM token costs 40-60%
#10I'm sure I'm not the only one hesitant to provide a 3rd party virtually MITM access to both my LLM usage + API keys. If this were capable of running locally, or even just an API for compressing non-sensitive parts of a prompt, I think it would be much easier to adopt.
Hi! You only need our API for the compression part — API keys and LLM usage are entirely managed by your own application. We don't have access to your SaaS, and we don't even know its name. We simply receive the text through our API, compress it, and return the response to your app. Your LLM — whether local, OpenAI, Claude, or any other — then processes it using your own API keys. Your data stays safe with you. And w…
from openai import OpenAI
client = OpenAI(
base_url="https://agentready.cloud/v1", # ← only change
api_key="ak_...", # AgentReady key
default_headers={
"X-Upstream-API-Key": "sk-..." # your OpenAI key
}
)
# Every call is now compressed automatically
response = client.chat.completions.create(
model="gpt-4o",
messages=[{"role": "user", "content": your_long_prompt}]
)
provide you our OpenAI key (via the X-Upstream-API-Key header)?