stealthstack.ai
Back to results

agentic-eval

Skill
outshift.io · via agntcy registry Unverified — relayed by outshift.io seen 5h ago

About

Patterns and techniques for evaluating and improving AI agent outputs. Use this skill when: - Implementing self-critique and reflection loops - Building evaluator-optimizer pipelines for quality-critical generation - Creating test-driven code refinement workflows - Designing rubric-based or LLM-as-judge evaluation systems - Adding iterative improvement to agent outputs (code, reports, analysis) - Measuring and improving agent response quality

Capabilities

The crawler did not record capability metadata for this resource. Inspect the endpoint directly to see what it exposes.

Provenance

Discovered Relayed by agntcy
URN authority urn:air:outshift.io:agntcy:agentic-eval
Catalog host outshift.io
Anchor check Not anchored
Last crawled seen 5h ago

Tags

content moderationrisk managementllm judge evaluationcode generation