Token导航 LogoToken导航TokenDH.com
效率操作浏览器clawhub未标认证来源可访问clear审计提醒

page-agent寻呼 Agent

Agent Skill

page-agent 用于处理浏览器自动化、网页检查和页面信息提取,适合在 OpenClaw 中需要让 Agent 打开页面、读取网页或验证前端流程时使用。可结合来源仓库、安装命令和原始 README 继续核验具体用法。安装前建议确认权限范围、维护状态,以及是否会触发联网、命令执行或文件读写。

总安装

29,376

周安装

1,198

GitHub Stars

公开资料未说明

下载量

9,504
OpenClaw

安装说明

本站只整理中文说明和来源信息,不托管安装包,也不代用户安装。

GitHub

来源数

2

许可证

MIT-0

最后核验

2026-05-01

来源状态

来源可访问

安装方式

通过对话安装

复制提示词发给支持本地命令或 Skills 的 AI 助手,先确认命令和权限,再让它执行。

请帮我安装这个 Agent Skill:page-agent(寻呼 Agent)
来源仓库:https://github.com/dongdongbear/page-agent
安装命令:
openclaw skills install page-agent
安装前请先检查当前环境是否支持对应 CLI,并向我确认将要执行的命令、安装目录、联网范围和文件读写权限;确认后再执行。

命令行安装

复制命令到本机终端执行。该命令会通过 OpenClaw 从第三方来源获取 Skill;本站只展示命令,不托管安装包,也不自动执行。

ClawHubOpenClaw
openclaw skills install page-agent

简介

增强浏览器DOM操作的页面控制器工具。

  • 支持精确的网页元素提取和交互检测。page-agent 属于效率类 Skill,可作为该场景下的辅助能力补充。
  • 提供页面状态监控和动态内容捕获能力。
  • 适用于前端测试和网页数据抓取场景。适用宿主包括 OpenClaw,接入前应确认版本、权限和运行环境要求。
  • 建议注意目标网站的反爬虫防护措施。

SKILL.md

name
page-agent
license
MIT
description
Enhanced browser DOM manipulation using PageAgent's page-controller. Injects into any web page to provide precise DOM extraction, interactive element detection (cursor:pointer heuristic), and robust interaction (full event chain simulation, React-compatible input). Use when you need to operate on web pages with precision — clicking, typing, scrolling, form filling, or reading page structure. Combines with frontend-design skill for full design→code→browser-operate workflow.

PageAgent Browser Enhancement Skill

Injects alibaba/page-agent v1.5.6 PageController into web pages via the browser tool's evaluate action. Gives you superior DOM manipulation compared to basic browser actions.

Key Advantages Over Basic Browser Tool

  1. cursor:pointer heuristic — detects clickable elements even without semantic tags
  2. Full event chain — mouseenter→mouseover→mousedown→focus→mouseup→click (not just .click())
  3. React/Vue compatible input — uses native value setter to bypass framework interception
  4. contenteditable support — proper beforeinput/input event dispatch
  5. Indexed elements[N]<tag> format for precise LLM-directed operations
  6. Incremental change detection*[N] marks new elements since last step

Usage Flow

Step 1: Inject PageController into the page

Use the CDP injection script (handles the 72KB library injection):

node ~/.openclaw/workspace/skills/page-agent/scripts/inject-cdp.mjs <TARGET_ID>

Where TARGET_ID is from browser(action="open", ...). The script injects both page-controller-global.js and inject.js via CDP WebSocket, outputting ✅ injected on success.

Step 2: Get page state (DOM extraction)

// Returns { url, title, header, content, footer }
// content is the LLM-readable simplified HTML with indexed interactive elements
const state = await window.__PA__.getState();
return JSON.stringify({ url: state.url, title: state.title, content: state.content, footer: state.footer });

The content field looks like:

[0]<a aria-label=首页 />
[1]<div >PageAgent />
[2]<button role=button>快速开始 />
[3]<input placeholder=搜索... type=text />

Step 3: Perform actions by index

// Click element at index 2
await window.__PA__.click(2);

// Type text into input at index 3
await window.__PA__.input(3, "hello world");

// Select dropdown option
await window.__PA__.select(5, "Option A");

// Scroll down 1 page
await window.__PA__.scroll(true, 1);

// Scroll specific element
await window.__PA__.scrollElement(4, true, 1);

Step 4: Re-read state after actions

After each action, call getState() again to see the updated DOM. Look for *[N] markers which indicate newly appeared elements.

Practical Workflow: Design → Code → Operate

  1. Design: Use frontend-design skill to create the page
  2. Serve: Start a local dev server (npx serve or framework dev server)
  3. Open: browser(action="open", targetUrl="http://localhost:3000")
  4. Inject: Load PageController into the page (Step 1 above)
  5. Inspect: Get DOM state to understand current page structure
  6. Operate: Click, type, scroll to test and interact with the page
  7. Iterate: Modify code based on what you observe, re-inject, repeat

Tips

  • Always re-inject after page navigation (SPA route changes are fine, full reloads need re-inject)
  • The content output is token-efficient — use it instead of screenshots when possible
  • For long pages, use scroll + getState to see content below the fold
  • Clean up highlights with window.__PA__.cleanUp() before taking screenshots
  • Use profile="openclaw" for the isolated browser, or profile="chrome" for the Chrome extension relay

Files

  • scripts/page-controller.js — PageController library (72KB, from @page-agent/page-controller@1.5.6)
  • scripts/inject.js — Helper wrapper that creates window.__PA__ API

适合场景

01

OpenClaw 用户查找和安装 Skill 时

02

用户想查找某类 Agent Skill 时

03

需要根据任务场景推荐可安装能力包时

04

需要对比不同来源的安装命令和来源信息时

能力概览

能力 1

按任务关键词查找相关 Skills

能力 2

展示可复制的安装命令

能力 3

保留来源站点、仓库和原始说明,方便继续核验

能力 4

补充不同宿主或平台的使用分布数据

能力 5

展示第三方安全扫描或审计结果

安装后应在对应宿主中按原始 README 的触发条件使用;具体调用方式请以来源页面和 README 为准。

平台分布

OpenClaw

90.73%
按下载量换算8,623

安全审计

VirusTotal

可疑

ClawScan

可疑

Static analysis

可疑

权限和风险

操作浏览器

该 Skill 可能涉及浏览器控制能力,使用时可能读取或操作网页内容,需要在受控环境中确认权限边界。

安装前确认

本站仅展示第三方公开信息,不托管安装包,不提供自动安装或运行环境。安装前应自行审查源码、依赖和命令行为。来源安全扫描存在 warning/failed 结果,不能写成本站确认安全。当前只有一个来源,正式发布前建议补源仓库或其他目录站核验。

来源信息

继续浏览同类 Skills