AgentCompass: A Unified Evaluation Infrastructure for Agent Capabilities
As Large Language Models (LLMs) evolve into autonomous agents, the need for unified evaluation infrastructure becomes critical. However, current evaluation pipelines remain highly fragmented and tightly coupled, hindering reproducibility an
https://arxiv.org/abs/2607.13705v2 ↗Thesis fit
Good fit
Within your typical scope; diligence still required.
In your usual scope
Idea match
Moderate
How close the company’s idea is to your thesis statement
Sector
agents, llm, ai
Overlap with sectors you care about
Geography
Unknown
Location unknown — scores 0
Your thesis: “We back exceptional technical founders building AI-first products and infrastructure, deploying $100K checks within 24 hours.”
Founder → stable
Traction → stable
Idea vs market ↓ declining
▸ Add / edit details
Correct facts used on the next screening or memo.
People
K
Kai Chen
1.0
low confidence
Z
Zichen Ding
1.0
low confidence
J
Jiaye Ge
1.0
low confidence
S
Shufan Jiang
1.0
low confidence
M
Mo Li
1.0
low confidence
Q
Qingqiu Li
1.0
low confidence
Z
Zehao Li
1.0
low confidence
Z
Zonglin Li
1.0
low confidence
T
Tiaohao Liang
1.0
low confidence
S
Shudong Liu
1.0
low confidence
Z
Zerun Ma
1.0
low confidence
Z
Zixing Shang
1.0
low confidence
Activity & evidence
Similar baseline plays (YC · idea space)
Doublezero
· Active
Platform to build, use, and monetize fully autonomous agents
founders not scraped yet