Project overview
Focusgroup dynamically tests conversational systems by combining reusable personas, feature scenarios, and a system under test. It runs real multi-turn conversations, then evaluates them with state checks, safety invariants, performance metrics, persona debriefs, or an LLM judge as a last resort. Reports cluster themes, segment failures by persona axes, and compare regressions. The current scope excludes browser interfaces, batch jobs, and pure classifiers.
Repository facts
- Primary language
- Python
- License
- MIT
- Repository updated
- May 26, 2026
- Default branch
- main
Resource types
General skill
Use cases
Testing and debugging
Platforms
Claude Code, Codex, and more
Runtime
Command lineLocal
Audience
Developers