About Multi-Environment Web Challenge
Definition and scoring
- Organisation
- MiniMax / benchmark maintainers
- Category
- Agents
- Version
- 2026
- Direction
- higher is better
- Ranking use
- Reference
- Contamination risk
- Unknown
A benchmark that evaluates AI agents on multi-environment web challenges, testing navigation and task completion across diverse live web environments. Every genuine source row stays tied to its exact model label, configuration, benchmark version and evaluation system. The summary chart shows one best compatible score per canonical product; Score 2.0 admits only explicitly mapped, frozen protocols and keeps incompatible configurations separate.
Open benchmark source ↗