01020304

GUIDE · DEVELOPMENT

Agent Skills for a safer development loop

A development stack should make review, testing, browser evidence, and delivery clearer—not hide more work behind a larger toolchain. Start with the weakest part of your current loop.

This guide brings together source-linked Skills for coding, testing, browser workflows, and delivery. It is intentionally smaller than the full development directory.

SELECTION ORDER

How to choose

Treat each Skill as one part of an engineering loop. A stronger loop has a visible scope, an inspectable output, and a clear stop condition.

  1. 01

    Fix one bottleneck first: plan, implementation, test, browser evidence, or release work

  2. 02

    Prefer outputs you can inspect, such as tests, diffs, screenshots, or a written plan

  3. 03

    Keep write and command permissions proportionate to the task

  4. 04

    Check the source revision and first risk before combining Skills

REVIEWABLE CANDIDATES

Start with these candidates

Sorted from the same published SkillSignal metadata score used throughout the directory. Scores are signals, not safety approval.

Browser Automation96

agent-browser

Browser automation built for repeatable agent workflows.

Agent fit
Development workflow
Main limitation
The checked-in SKILL.md is a discovery stub; full instructions come from the installed CLI version.
First risk
Browser-side actions · High
Code Review & Testing96

frontend-design

Design constraints that push generated interfaces beyond generic AI UI.

Agent fit
Development workflow
Main limitation
Aesthetic judgments remain subjective and still need stakeholder review.
First risk
Style over function · Medium
Code Review & Testing94

mcp-builder

Plan, implement, and evaluate high-quality MCP servers in Python or TypeScript.

Agent fit
Development workflow
Main limitation
The guide cannot replace current MCP SDK or target-API documentation.
First risk
Overpowered external-service tools · High
Code Review & Testing94

skill-creator

Create, improve, and benchmark Agent Skills with trigger tests and comparative evaluation.

Agent fit
Development workflow
Main limitation
Meaningful benchmarks require representative prompts, repeat runs, and careful human judgment.
First risk
Untrusted evaluation prompts and artifacts · High
Code Review & Testing94

web-design-guidelines

Review interfaces against practical web design and accessibility rules.

Agent fit
Development workflow
Main limitation
It depends on network access and a live guideline document whose contents can change over time.
First risk
Rule drift or context-free findings · Medium
Code Review & Testing93

anti-ui-slop

Product-specific UI design contracts grounded in real interface evidence and a hard finish gate.

Agent fit
Development workflow
Main limitation
Requires access to the repository and ideally external UIZZE references or user-provided screenshots.
First risk
Copying external interface patterns too literally · Medium

DECISION BOUNDARY

What this guide does not prove

A Skill that creates tests, screenshots, or plans does not prove production behavior, accessibility, security, or release readiness. Human review remains part of the loop.

KEEP EXPLORING

Need a more exact match

Open the directory to filter by Agent, operating system, price, and permission preference. The ranking is deterministic and does not call an LLM.

Open the Skill matcher