Summary
Optional LLM semantic judge for hard cases (“are these tools the same user task?”) that never overrides deterministic findings and is always labeled MODEL-BASED.
Milestone
Milestone 9 — Semantic Intelligence
Suggested labels
enhancement, priority:p2, area:probes, research
Acceptance criteria
Related
Summary
Optional LLM semantic judge for hard cases (“are these tools the same user task?”) that never overrides deterministic findings and is always labeled MODEL-BASED.
Milestone
Milestone 9 — Semantic Intelligence
Suggested labels
enhancement,priority:p2,area:probes,researchAcceptance criteria
runner.pyinterface (Design a provider-neutral model runner #45/Expand provider adapters and packaging for the model runner #60)Related