playwright · Microsoft

Playwright Agent 使用指南

运行浏览器测试、捕获截图与 Trace,并自动化可留存证据的 UI 工作流。

官方工具操作风险: R1R3docs-verified
Agent 适配度
86/100
证据可信度
docs-verified
文档检查日期
2026-07-10
独立测试版本
尚未独立测试

为 Agent 安装

选择与运行环境匹配的官方安装方式。在团队或 CI 环境中固定版本,并先运行版本检查。

npm推荐
macos · linux · windows
$ shell
npm install --save-dev @playwright/test
认证与最小权限
只提供任务需要的权限,凭证通过环境变量或平台密钥存储传入,不能写进提示词、仓库或日志。
无需认证支持无界面认证

无需服务凭据;仍应把文件系统和网络访问限制在任务范围内。

认证方式
none
密钥环境变量
None
凭据保存位置
此 CLI 不保存服务凭据。
Agent 与运行环境兼容性
先确认 Agent 能使用 Shell,再检查平台、网络与凭证是否可用。
claude-codecodexgemini-clicopilot-cli
Environments
local, ci, container, headless
Platforms
macos, linux, windows

用于稳定自动化的结构化输出

优先使用机器可读格式,并把 stdout 作为结果、stderr 作为诊断信息分别处理。

json · junit · html · blob
在支持的命令中使用 --reporter=json 或 --reporter=junit,并将诊断日志保留在 stderr。
--reporter=json--reporter=junit

尚未记录真实输出样例

当前结构化输出能力来自官方文档。完成有限、非破坏性执行并保存 stdout 前,不展示推测样例或伪造 Schema。

R0–R3 命令风险指南

风险按单条命令判断。R0 是本地或远程只读,R1 是可逆的本地写入,R2 会改变远程状态,R3 可能造成不可逆或生产级影响。

只读不等于可公开

R0 只表示命令不更改本地或远程状态。只读命令仍可能返回令牌、身份信息、配置或生产数据;只展示完成任务所需的最少内容,不得写入日志、Prompt 或提交内容。

R0检查测试报告
读取已有报告,不重新执行浏览器操作。
$ shell
npx playwright show-report
可重复执行
R1运行浏览器测试
会执行项目定义的浏览器操作并写入报告与 Trace。
$ shell
npx playwright test --reporter=json
执行前必须确认可能产生重复变更敏感输出
R2录制已认证流程
可能通过浏览器提交表单或修改远程状态。
$ shell
npx playwright codegen https://example.com
执行前必须确认可能产生重复变更

Agent Readiness 评分依据

适配度描述 Agent 操作工具的稳定程度,不代表所有命令都安全,也不替代独立执行测试。

文档证据对应的 Agent Readiness 为 86/100;尚未记录本地执行测试。

结构化输出
在支持的命令中使用 --reporter=json 或 --reporter=junit,并将诊断日志保留在 stderr。
18/20
x
无界面运行
官方文档描述了非交互认证或执行路径。
14/15
x
安全控制
CLI Finder 将读取命令与需要确认的命令分开。
11/15
x
确定性
命令尽量使用显式参数和文档支持的输出控制。
8/10
x
认证
无需服务凭据;仍应把文件系统和网络访问限制在任务范围内。
10/10
x
文档
本条目引用了 2026-07-10 检查的官方文档。
9/10
x
安装
官方安装路径覆盖 macOS、Linux 和 Windows。
8/8
x
维护状态
条目链接了官方源码仓库,便于检查发布与维护状态。
6/7
x
Agent 产物
CLI Finder 可生成基于注册表的 Skill 和策略;该分数不假定工具自身提供这些产物。
2/5
x

生成 Skill 或 Agent 策略

选择目标 Agent 和安全模式,生成包含安装、允许命令、确认边界与证据说明的可复制产物。

生成结果预览
SKILL.md
---
name: playwright-agent-workflow
description: Use Playwright for browser testing, screenshots, traces with explicit command risk and evidence boundaries.
---

# Playwright agent workflow

Use this skill when the task needs browser testing, screenshots, traces, authenticated UI flows.

## Evidence boundary

- Registry confidence: `docs-verified`
- Documentation checked: `2026-07-10`
- Locally tested version: `not tested`
- Do not describe this CLI as locally verified until its commands have actually been executed in an isolated environment.

## Executed smoke checks

- No local execution record is available.

## Installation

- npm (macos, linux, windows): `npm install --save-dev @playwright/test`

## Authentication

- Methods: none
- Secret environment variables: none
- Minimum permissions: No service credential is required; restrict filesystem and network access to the task.
- Credential storage: No service credential is stored for this CLI.
- Never print, persist, or commit credential values.

## Allowed commands (read-only)

- `npx playwright show-report` — R0: Reads an existing report without rerunning browser actions.

## Commands requiring explicit approval (read-only)

- None recorded.

## Forbidden commands (read-only)

- R1 `npx playwright test --reporter=json` — Executes project-defined browser actions and writes reports and traces.
- R2 `npx playwright codegen https://example.com` — Can submit forms or mutate remote state through the browser.

## Execution rules

1. Mode boundary: R0 exact commands may be used; R1, R2, and R3 commands are forbidden.
2. Confirm the selected account, project, context, database, namespace, or environment before any command.
3. Prefer structured output using `--reporter=json`, `--reporter=junit`.
4. Capture the exact command, exit code, stdout, and stderr separately.
5. A generated prefix policy must prompt unless that exact prefix is explicitly marked suffix-safe; do not infer safety from the executable name.
6. Never broaden credentials or disable safety controls to make a command succeed.

## Official sources

- [Playwright test CLI documentation](https://playwright.dev/docs/test-cli)
- [Playwright test CLI documentation source repository](https://github.com/microsoft/playwright)

这个任务该用 CLI、MCP 还是 API

CLI
适合在开发机、CI 或容器里复用现有 Shell、凭证和脚本,尤其适合短时、可观察的任务。
MCP
当 Agent 需要受控工具定义、委托身份或由服务端集中治理访问时,MCP 可能更合适。
API
当工作流是应用内长期集成、批量调用或事件驱动时,直接 API 往往比启动进程更稳定。
查看 CLI 与 MCP 完整对比

验证记录与官方证据

CLI Finder 分开记录文档检查和真实执行。未执行过的安装、帮助、退出码与输出不能标为 Verified。

当前证据边界
已审阅官方文档,但未在本地执行安装、帮助输出、退出码、无界面行为或结构化输出测试。
证据可信度
docs-verified
独立测试版本
尚未独立测试
测试环境
未记录
官方来源
打开官方资料确认当前版本和命令。

替代工具与相关入口

通过可审计产物测试浏览器界面并提取已授权的网页内容。
把网页抓取或爬取为 Markdown 与 JSON,供 Agent 后续分析。
先限定少量 URL,再选择内容提取或浏览器自动化,并为每份结果保留来源。
用确定性本地工具配合 Codex,只在沙箱和审批规则允许时加入远程 CLI。
内容提取优先 Firecrawl;浏览器行为、截图、测试和必要交互使用 Playwright。

Agent 使用 Playwright 的常见问题