A development stack should make review, testing, browser evidence, and delivery clearer—not hide more work behind a larger toolchain. Start with the weakest part of your current loop.
This guide brings together source-linked Skills for coding, testing, browser workflows, and delivery. It is intentionally smaller than the full development directory.
Treat each Skill as one part of an engineering loop. A stronger loop has a visible scope, an inspectable output, and a clear stop condition.
01
Fix one bottleneck first: plan, implementation, test, browser evidence, or release work
02
Prefer outputs you can inspect, such as tests, diffs, screenshots, or a written plan
03
Keep write and command permissions proportionate to the task
04
Check the source revision and first risk before combining Skills
REVIEWABLE CANDIDATES
Start with these candidates
Sorted from the same published SkillSignal metadata score used throughout the directory. Scores are signals, not safety approval.
Browser Automation96
agent-browser
Browser automation built for repeatable agent workflows.
Agent fit
Development workflow
Main limitation
The checked-in SKILL.md is a discovery stub; full instructions come from the installed CLI version.
First risk
Browser-side actions · High
Code Review & Testing96
frontend-design
Design constraints that push generated interfaces beyond generic AI UI.
Agent fit
Development workflow
Main limitation
Aesthetic judgments remain subjective and still need stakeholder review.
First risk
Style over function · Medium
Code Review & Testing94
mcp-builder
Plan, implement, and evaluate high-quality MCP servers in Python or TypeScript.
Agent fit
Development workflow
Main limitation
The guide cannot replace current MCP SDK or target-API documentation.
First risk
Overpowered external-service tools · High
Code Review & Testing94
skill-creator
Create, improve, and benchmark Agent Skills with trigger tests and comparative evaluation.
Agent fit
Development workflow
Main limitation
Meaningful benchmarks require representative prompts, repeat runs, and careful human judgment.
First risk
Untrusted evaluation prompts and artifacts · High
Code Review & Testing94
web-design-guidelines
Review interfaces against practical web design and accessibility rules.
Agent fit
Development workflow
Main limitation
It depends on network access and a live guideline document whose contents can change over time.
First risk
Rule drift or context-free findings · Medium
Code Review & Testing93
anti-ui-slop
Product-specific UI design contracts grounded in real interface evidence and a hard finish gate.
Agent fit
Development workflow
Main limitation
Requires access to the repository and ideally external UIZZE references or user-provided screenshots.
First risk
Copying external interface patterns too literally · Medium
DECISION BOUNDARY
What this guide does not prove
A Skill that creates tests, screenshots, or plans does not prove production behavior, accessibility, security, or release readiness. Human review remains part of the loop.
KEEP EXPLORING
Need a more exact match
Open the directory to filter by Agent, operating system, price, and permission preference. The ranking is deterministic and does not call an LLM.