Agent Building skills

Browse reusable Agent Skills, each with a clear purpose and practical guidance.

Cursor Skill (.mdc) Authoring

Author effective Cursor rules in .cursor/rules/*.mdc - YAML frontmatter (description, globs, alwaysApply), the four rule types, scoped QA rules, subagents, and structure that actually steers the model.

174 repo starsObserved in 1 repos
Agent Building

dv-connect

One-step setup for a Dataverse environment — installs tools, authenticates, registers the MCP server, and writes `.env`. Use when starting a new project, switching environments, fixing authentication, or troubleshooting an MCP connection that won't come up.

174 repo starsObserved in 1 repos
Agent Building

dv-overview

Tool routing and cross-cutting rules for Dataverse work — which skill applies to which task, environment-confirmation, and pull-to-repo. Use when the user mentions Dataverse, Dynamics 365, Power Platform, or CRM; this skill picks the specialist (dv-connect / dv-data / dv-metadata / dv-query / dv-solution / dv-admin / dv-security) for the request.

174 repo starsObserved in 1 repos
Agent Building

OpenAI Evals Trace Grading

Grade LLM and agent traces with OpenAI Evals - build datasets, configure string/python/model graders, run eval suites, and gate agent behavior changes in CI.

174 repo starsObserved in 1 repos
Agent Building

Playwright CLI Agent Loop

Teach AI coding agents to use the Playwright CLI and debug loop efficiently with last-failed runs, locator probing, trace evidence, and safe healing.

174 repo starsObserved in 1 repos
Agent Building

Playwright Multi-Tab & Window Handling

Teaches the agent to handle popups, new tabs, and multiple browser windows in Playwright using waitForEvent('page'), context pages, and reliable tab switching for OAuth and target=_blank links.

174 repo starsObserved in 1 repos
Agent Building

Prompt Testing

Comprehensive prompt testing and LLM output evaluation skill covering hallucination detection, response quality scoring, regression testing for prompts, A/B testing, and building evaluation pipelines for AI-powered applications.

174 repo starsObserved in 1 repos
Agent Building

Promptfoo LLM Red Teaming

Evaluate and red-team LLM applications with promptfoo, declarative YAML evals, assertions, model comparisons, and automated adversarial scans for prompt injection, jailbreaks, PII leaks, and unsafe outputs in CI.

174 repo starsObserved in 1 repos
Agent Building

qa-agent-claude

Turn Claude Code into an autonomous QA agent — an explore, generate, run, heal, report loop that maps the app, writes tests for real user journeys, executes them, self-heals broken locators, and reports coverage. Build a QA agent skill for Claude Code.

174 repo starsObserved in 1 repos
Agent Building

QA Agent for Claude Code

Turn Claude Code into an autonomous QA agent — an explore, generate, run, heal, report loop that maps the app, writes tests for real user journeys, executes them, self-heals broken locators, and reports coverage. Build a QA agent skill for Claude Code.

174 repo starsObserved in 1 repos
Agent Building