OmniaBench: Benchmarking General AI Agents Across Diverse Scenarios
Large language models are increasingly evolving from text generators into general agents capable of understanding user requests, invoking external tools, and completing complex tasks through interaction. However, existing agent benchmarks o
https://arxiv.org/abs/2607.14989v1 ↗Thesis fit
Good fit
Within your typical scope; diligence still required.
In your usual scope
Idea match
Light
How close the company’s idea is to your thesis statement
Sector
agents, ai
Overlap with sectors you care about
Geography
Unknown
Location unknown — scores 0
Your thesis: “We back exceptional technical founders building AI-first products and infrastructure, deploying $100K checks within 24 hours.”
Founder → stable
Traction → stable
Idea vs market ↓ declining
▸ Add / edit details
Correct facts used on the next screening or memo.
People
C
Chengyu Shen
1.0
low confidence
Y
Yujie Fu
1.0
low confidence
G
Gangtao Xin
2.0
low confidence
Y
Yanheng Hou
2.0
low confidence
W
Wenlong Fei
2.0
low confidence
G
Guojie Zhu
2.0
low confidence
J
Jiawei Li
2.0
low confidence
H
Hongcheng Gao
1.0
low confidence
R
Runming He
2.0
low confidence
Z
Zhen Hao Wong
2.0
low confidence
M
Meiyi Qiang
2.0
low confidence
H
Hao Liang
2.0
low confidence
C
C-K Shen
1.6
low confidence
Y
Y Fu
1.2
low confidence
Activity & evidence
Similar baseline plays (YC · idea space)
Mastra
· Active
The Javascript framework for building AI agents, from the Gatsby devs
founders not scraped yet
Sourcebot
· Active
Helping humans and AI agents understand massive codebases
founders not scraped yet