General skillAI and agent development

deepeval-bcg

Evaluate strategic AI outputs with consulting-style scoring and adversarial checks.

Stars
1
Forks
0
License
MIT
Updated
Updated May 21, 2026

Checking live repository facts…

Project overview

Run the skill against a strategic deliverable to receive a PASS, REVISE, or FAIL verdict, scores across eight consulting-oriented dimensions, and a concrete repair directive. Its Skeptic Agent probes for silent ambiguity choices, agreement with flawed premises, and missing counterarguments, while a separate novelty stack checks whether the insight is more than generic strategy language. The evaluation runs inside Claude Code without separate provider API keys.

Repository facts

Primary language
Python
License
MIT
Repository updated
May 21, 2026
Default branch
main

Resource types

General skill

Use cases

AI and agent development

Platforms

Claude Code, Codex, and more

Capabilities

Verification and evals

Related projects

Browse more similar projects