Building in public Sign in with GitHub
Kredisco
Credit scoring for AI agents

No bank writes a mortgage without a credit check. You wired six agents in on faith.

An agent will always tell you it did well. Kredisco never asks it — every score is rebuilt from what the caller actually observed: retries, failures, deadlines missed.

Live simulation — not real traffic trustedflagged

Six agents, one orchestrator. Every completed task signs a receipt and moves a score. Watch agt_2c81 — it's the one costing someone money.

How it works
01

The caller signs, not the agent

An agent can't vouch for itself. Every receipt is signed by whoever called the agent, the same way lenders report to a credit bureau instead of borrowers reporting on themselves.

02

Receipts accumulate into history

Cost, latency, retries, and outcome, tagged by task class. One receipt is a data point. A thousand is a track record.

03

History is scored against the class

Nine thousand tokens means nothing on its own. Nine thousand tokens for a job the median agent does in four thousand means something.

Using it

Name each agent once. Wrap the calls you already make. Kredisco opens a file and starts scoring.

kd = Kredisco(api_key="kd_...") writer = kd.agent("draft-writer") # before reply = write_reply(ticket) # after reply = kd.track(writer, "draft", write_reply, ticket, validate=lambda r: len(r) > 40, retries=1)

Python. LangGraph, CrewAI, or a plain loop.

A credit score follows you between lenders. It is held by a bureau, not by you, and a stranger reads it in seconds.

Agents have no equivalent. Today that means you cannot tell which one in your own pipeline to stop calling. Soon it will mean more than that.

Today
Your own agents, ranked by what they actually delivered.
Soon
Agents are products you buy. The file follows the agent between the teams that hire it, and you check the number before you wire it in.

A name tells you who built it. A score tells you whether it delivers.

What I'm looking for
"

Teams running multi-agent pipelines

If you've ever had output come back wrong and had to guess which step caused it, I want to hear how you figured it out. Ten minutes on a call, no pitch.

Get in touch

Building this in the open. Early access, questions, or war stories about agents behaving badly all welcome.