Skip to main content

Browse the directory

Showing 30 of 66 resources for "testing"
Saved
Active

Source-backed filter active — add entries to compare trust side by side.

Trust snapshot

Trust signals across 40 of 66 results

Claimed
0%(0/40)

2 trust signals differ in this sample: Source provenance, Submitter

Signals differ on Source provenance, Submitter — add entries to compare before you install.

Rollout signal scan

2 rollout risk signals in current results

Biggest gaps: metadata review, package integrity. 0 entries have 2+ required gaps.

12 scanned

Install payload

Install payload is broadly covered in current results.

good

100% (12/12)

Adoption queue

Browse adoption queue · balanced

10/66 visible results are in hold tier and need mitigation before adoption.

ready 0caution 56hold 10

.NET Agent Skills

1 blockers: Metadata review

caution

50/100

Request metadata review from maintainers or internal owners.

Collect package checksum or signed artifact information.

skills/dotnet-agent-skills · trust review · confidence 67%

Agent Device MCP Server

1 blockers: Metadata review

caution

50/100

Request metadata review from maintainers or internal owners.

Collect package checksum or signed artifact information.

mcp/agent-device-mcp-server · trust review · confidence 67%

caution

50/100

Request metadata review from maintainers or internal owners.

Collect package checksum or signed artifact information.

agents/angular-repository-contributor-agent · trust review · confidence 67%

caution

50/100

Request metadata review from maintainers or internal owners.

Collect package checksum or signed artifact information.

agents/ansible-repository-contributor-agent · trust review · confidence 67%

caution

50/100

Request metadata review from maintainers or internal owners.

Collect package checksum or signed artifact information.

agents/astro-repository-contributor-agent · trust review · confidence 67%

Decision confidence

Decision confidence scan · balanced

10/66 results are low-confidence and need review before adoption.

high 0medium 56low 10

.NET Agent Skills

Address Metadata review, Package integrity before broader rollout.

medium

54/100

Missing: Metadata reviewMissing: Package integrity

skills/dotnet-agent-skills · trust review

Agent Device MCP Server

Address Metadata review, Package integrity before broader rollout.

medium

54/100

Missing: Metadata reviewMissing: Package integrity

mcp/agent-device-mcp-server · trust review

Astro Repository Contributor Agent for Claude

Address Metadata review, Package integrity before broader rollout.

medium

54/100

Missing: Metadata reviewMissing: Package integrity

agents/astro-repository-contributor-agent · trust review

Freshness distribution

Current results are broadly fresh

Median age 53 days; all 12 scanned entries are within 90 days.

median 53d

Aging

91–180 days

0%

0 entries

Stale

> 180 days

0%

0 entries

Theme distribution

Results center on testing

54% of this view shares the top theme. Leading themes: testing, mcp, skills.

Focused

88 distinct themes across 24 scanned

AWS Labs logo

Official AWS Labs MCP server for serverless development that gives AI assistants contextual guidance plus tools to initialize, build, deploy, and troubleshoot AWS SAM and Lambda-based serverless applications.

Microsoft .NET team skill marketplace for AI coding agents working on .NET, C#, ASP.NET Core, Blazor, MAUI, diagnostics, MSBuild, NuGet, upgrades, tests, AI workflows, RAG pipelines, and C# MCP servers.

Level:expertType:capability-packVerified:validated
Safety ✓ Privacy ✓

Official LiveKit Agent Skills for AI coding agents building low-latency voice AI, LiveKit Agents workflows, handoffs, mandatory tests, and simulation scenario suites.

Level:expertType:capability-packVerified:validated
Safety ✓ Privacy ✓
LiveKit Agents logo

Open-source framework for building realtime voice, video, and multimodal AI agents with LiveKit rooms, STT, LLMs, TTS, job scheduling, telephony, MCP tools, testing, and production deployment paths.

MATLAB logo

Official MathWorks MCP server that lets Claude detect MATLAB installations, inspect toolboxes, analyze MATLAB files, evaluate MATLAB code, run scripts, run tests, and expose reviewed custom MATLAB functions as MCP tools.

Unity Editor MCP bridge that lets AI assistants manage assets, edit scripts, control scenes, run tests, and automate Unity workflows through the Model Context Protocol.

HexStrike AI logo

Offensive security MCP framework that connects AI agents to a large toolkit for authorized penetration testing, vulnerability discovery, CTF, OSINT, and security research workflows.

Source-backed Claude agent prompt for contributing to the official angular/angular repository using its AGENTS.md guidance for pnpm, Bazel test targets, coding standards, commit guidelines, zoneless tests, async stability, and PR handling.

Build Inngest-backed Next.js workflows with event triggers, durable steps, local Dev Server testing, API route serving, retries, concurrency, and production deployment review.

Level:advancedType:generalVerified:validated
Safety ✓ Privacy ✓

Source-backed Claude agent prompt for contributing to the official elastic/kibana repository using its AGENTS.md guidance for Kibana modules, plugin lifecycle boundaries, server plugin lazy loading, TypeScript style, i18n, Scout, Jest, FTR, scoped type checks, and focused validation.

Source-backed Claude agent prompt for contributing to the official mui/material-ui monorepo using its AGENTS.md guidance, pnpm workspace filters, package build and test commands, component conventions, public error-message rules, API docs generation, visual regression and accessibility checks, and pre-PR checklist.

Build and maintain NestJS backend APIs with modules, controllers, providers, dependency injection, configuration, validation pipes, guards, interceptors, exception filters, OpenAPI docs, testing, and production safety review.

Level:advancedType:generalVerified:validated
Safety ✓ Privacy ✓

Connect Claude to BrowserStack for permission-scoped web, app, accessibility, and test automation workflows.

Ragas logo
Ragasby Vibrant Labs · submitted by oktofeesh1

Open-source evaluation framework for testing RAG systems, prompts, agents, workflows, and other LLM application behavior.

Official visual testing and debugging tool for Model Context Protocol servers.

Promptfoo logo

Open-source prompt testing and red-teaming framework for LLM outputs, regressions, evaluations, and security checks.

Safety · Privacy ✓
Giskard logo

AI testing platform for evaluating, scanning, and monitoring machine learning and LLM application quality.

Safety · Privacy ·

Official Dart team Agent Skills for AI coding agents working on Dart unit tests, CLI apps, coverage, runtime errors, mocks, package conflicts, static analysis, Native Assets, FFI, ffigen, and pattern matching.

Level:expertType:capability-packVerified:validated
Safety ✓ Privacy ✓

Official Flutter team Agent Skills for AI coding agents building Flutter apps, fixing layout issues, adding widget and integration tests, creating widget previews, applying layered architecture, routing, localization, JSON serialization, and HTTP workflows.

Level:expertType:capability-packVerified:validated
Safety ✓ Privacy ✓

Community slash command runbook for frontend visual QA using documented Claude Code Chrome integration workflows: enable /chrome, open a local page, read console messages, and follow the design verification checklist from the Chrome integration guide.

Invocation:/frontend-visual-qa <route-or-host>
Safety ✓ Privacy ✓

Community slash command runbook for adding minimal automated tests around a changed module: inspect the git diff, mirror repository test conventions, and draft focused unit or integration tests using Anthropic develop-tests guidance.

Invocation:/targeted-test-generation <file-or-symbol>
Safety ✓ Privacy ✓

Official CircleCI MCP server that lets LLMs query build and test failures, detect flaky tests, check pipeline status, fetch build logs, and validate config in CircleCI through natural language.

agent-device logo

Official MCP server for agent-device, Callstack's device automation CLI for inspecting, controlling, debugging, recording, and collecting evidence from iOS, Android, TV, macOS, Linux, React Native, Expo, Flutter, and native apps.

Cypress Cloud logo

Official remote MCP server for connecting Claude and other AI coding tools to Cypress Cloud runs, failures, flake data, accessibility reports, and UI Coverage results.

Mozilla-maintained MCP server for automating Firefox through WebDriver BiDi, with tools for page navigation, snapshots, UID-based input, screenshots, network requests, console messages, dialogs, history, viewport changes, optional JavaScript evaluation, privileged Firefox contexts, preferences, and WebExtension.

WebDriverIO MCP Server logo

WebDriverIO MCP server that lets Claude automate browsers and mobile apps with navigation, clicks, typing, screenshots, accessibility snapshots, cookies, geolocation, native app sessions, and Appium-style gestures.

BrowserMCP logo

Browser automation MCP server and Chrome extension that lets AI applications control a connected tab in the user's existing browser profile.

Mobile MCP logo

Cross-platform mobile automation MCP server for iOS and Android simulators, emulators, and real devices using accessibility snapshots, screenshots, coordinate taps, app management, and input tools.

OpenAI Evalsby OpenAI · submitted by JSONbored

Open-source framework from OpenAI for evaluating LLM and agent behavior with reusable eval definitions, grading logic, datasets, and regression workflows.

Expert skill for reviewing Playwright trace artifacts, screenshots, action timelines, network events, retries, and CI evidence to classify flaky browser test failures without guessing from logs alone.

Level:expertType:capability-packVerified:validated
Safety ✓ Privacy ✓