Harbor Adapters and Harbor-Index provide a unified evaluation infrastructure and curated meta-dataset for large-scale agentic evaluation
Read the original at arxiv.org→arXiv:2609.04298v1 Announce Type: new Abstract: Evaluating agents on the growing number of agentic benchmarks is challenging because they often require complex environments and agent integrations. We introduce...
Original headline: "Harbor Adapters and Harbor-Index: Infrastructure and a Curated Meta-Dataset for Large-Scale Agentic Evaluation"
Coverage timeline
- Sep 7, 04:00 UTC arXiv cs.AI lead source Harbor Adapters and Harbor-Index: Infrastructure and a Curated Meta-Dataset for Large-Scale Agentic Evaluation