# Multilingual agent release-gate worksheet

> Record task coverage, languages, modalities, tools, side effects, human review, safety, rollback, and approval before an agent reaches users or production systems.

Published: 2026-08-20. Updated: 2026-08-20. Last verified: 2026-08-20. Review interval: 365 days. Disclosure: none.

## Instructions

- List markets, language varieties, scripts, modalities, and risk tiers
- Test full traces including tool calls and side effects
- Set thresholds before final evaluation
- Keep results separate by language and task
- Require named approval, rollback, incident capture, and review date
- Block a critical cell that lacks evidence

## Gate rule



```
releaseReady = allCriticalCellsHaveEvidence
  && allRiskTierThresholdsPass
  && rollbackTested
  && incidentOwnerNamed
  && no unresolved high-severity language or modality defect
```

## Download

[Download the agent release gate](https://admas.net/resources/templates/downloads/multilingual-agent-release-gate.md)

Version 1.0. License: CC BY 4.0.

## Frequently asked questions

### Can one score cover every market?

No. Keep evidence for each material language, variety, script, modality, task, and risk tier. Aggregate summaries must not hide a failing market.

### What should be attached?

Link versioned test sets, traces, rubric and reviewer qualifications, separate results, defects, rollback evidence, approvals, and the exact system versions evaluated.

## Methodology

Admas designed this plain-language working template for transparent language-service delivery. The downloadable file contains no macros, remote formulas, tracking links, or hidden fields. Adapt it to the project, jurisdiction, client requirements, tax rules, language pair, service, and professional advice relevant to you.

## Source register

- [Web Content Accessibility Guidelines 2.2](https://www.w3.org/TR/WCAG22/) — World Wide Web Consortium; verified 2026-08-20.
- [AI Risk Management Framework: Generative AI Profile](https://www.nist.gov/publications/artificial-intelligence-risk-management-framework-generative-artificial-intelligence) — U.S. National Institute of Standards and Technology; verified 2026-08-20.
- [MAPS: A Multilingual Benchmark for Agent Performance and Security](https://aclanthology.org/2026.findings-eacl.42/) — ACL Anthology; verified 2026-08-20.
- [Creative Commons Attribution 4.0 International](https://creativecommons.org/licenses/by/4.0/) — Creative Commons; verified 2026-08-20.
