Research Built to Be Used

We build benchmarks, datasets, and infrastructure to advance machine learning research.

We are a research organization building open benchmarks, datasets, and infrastructure. Our work explores machine learning, autonomous systems, and how computational tools can advance scientific research.


Benchmarks

We build benchmarks to measure how AI agents perform on real research tasks.

Datasets

1.1M enriched papers. 129K research repositories. 778K code functions. The raw material for studying how AI systems interact with real scientific work.

Infrastructure

Runtimes for structured agent workloads. Orchestration topologies for recursive improvement loops. The scaffolding to run experiments at scale.

Algorithmic Research Group builds tools and infrastructure for research. Benchmarks for evaluating autonomous agents. Datasets for studying machine learning systems. Runtimes for running agent workloads at scale.