Skip to content
AI Landscape

About

Benchmarks multimodal agents on open-ended tasks in real computer environments

See something off? Suggest an edit

Related in Agent & General Benchmarks