# harbor environment/evals Projects at AI Tinkerers

> Canonical HTML: https://aitinkerers.org/technologies/harbor-environment-evals
> Markdown URL: https://aitinkerers.org/technologies/harbor-environment-evals.md
> Technology record last updated: 2026-06-03T09:29:35Z
> Generated: 2026-09-20T22:36:41Z

Harbor is an open-source framework for evaluating, optimizing, and training AI agents in sandboxed container environments.

Built by the creators of Terminal-Bench, Harbor standardizes how developers test and optimize AI agents (such as Claude Code and OpenHands) inside secure, isolated containers. The platform solves the infrastructure headache of agent evaluation by orchestrating parallel trials across cloud runtimes like Daytona and Modal. By decoupling tasks, environments, and agents into modular configurations, Harbor allows teams to run complex benchmarks (including SWE-Bench Verified and Terminal-Bench 2.0), generate high-quality trajectories for reinforcement learning, and implement robust CI/CD pipelines for agentic workflows.

- Official technology site: https://harborframework.com
- Public AI Tinkerers demos and talks: 0
- Result page: 1 of 1

## Recent Public Talks and Demos

No public projects are currently indexed for this technology.
