Category

Quality & Security

Quality & Security covers 32 resources across 5 subcategories in the AI Operator's Index. It's one of 9 top-level categories EE Solutions tracks for people operating with Claude and other frontier models.

Evals & Testing
12 resources
Agent evaluation product for versioned test sets, regression checks, multi-step analysis and suggested fixes.
Open-source tracing and evaluation.
Link
Evals, traces and production quality workflows.
Link
Code-first evaluation framework.
Link
Concrete patterns for evaluating cybersecurity agents using sandboxes, tasks, tools and graders.
Reusable eval skills built from Hamel Husain's work with production AI teams.
Product-oriented argument for making AI outputs inherently easy to verify before layering on evaluation tooling.
Open-source LLM observability and evals.
Link
Evaluation and telemetry platform that turns observed failure modes into continuously running evals.
Open-source LLM eval and red-team framework.
Link
RAG-focused evaluation framework.
Link
Tracing and evaluation from Weights & Biases.
Link
Governance & Risk
1 resources
Canonical AI risk framework.
Link
Observability & Tracing
2 resources
Tracing and observability platform for agent and LLM applications.
Link
Add and work with Logfire observability in Python applications.
Plugin123 stars
Security & Red Teaming
12 resources
Guidance for recognizing phishing, credential theft and social engineering during agent operation.
Skill489 stars
Repository-grounded threat modeling with trust boundaries and abuse paths.
Skill489 stars
Audit OpenAPI specs and API security risks from Claude Code.
Plugin1 stars
SAST, secret and IaC scanning through Aikido's Claude plugin.
Plugin13 stars
Anthropic plugin with edit warnings and diff review for common vulnerability classes.
Plugin33.7k stars
Project site for garak, the LLM vulnerability scanner.
Link
Adversarial threat knowledge base for AI.
Link
LLM vulnerability scanner.
Repo8.9k stars
Current OWASP 2026 risk list for GenAI and LLM applications.
Link
Core security guidance for GenAI systems.
Link
High-value LLM risk checklist.
Link
Security scanning and code guidance from Semgrep inside Claude Code.
Plugin10 stars
Security Skills
5 resources
Postgres performance, schema, security and RLS guidance maintained by Supabase.
Skill2.5k stars
Claude-specific visual skill directory. Use for discovery, not as a security or quality guarantee.
Link
Opinionated Claude Code config and workflows.
Repo2.1k stars
Reviewed external Claude plugins and skills.
Repo489 stars
Security-focused Claude Code skills from a respected security firm.
Skill6.7k stars

Looking for hands-on help in Quality & Security? EE Solutions can help. EE Solutions is a senior technology team for private capital firms and their portfolio companies.

Talk to EE Solutions ↗