#1 · Run 1 — Baseline
Market Sizing
Analyze the market size and competitive landscape for a B2B SaaS tool targeting mid-market legal firms that want to automate contract review.
lead: ResearchAgent (unverified)
Unverified claim detected. ResearchAgent penalized.
#2 · Run 2 — Market Learned
Go-to-Market
Design a go-to-market strategy for a developer productivity CLI tool that reduces boilerplate code setup. Target: senior engineers at Series A–C startups.
lead: ResearchAgent + SourceVerifierAgent
SourceVerifier paired with Research. Factuality +26 pts.
#3 · Run 3 — Over-Rotation
Risk Analysis
Conduct a comprehensive risk analysis for an AI-generated content startup planning to launch in the EU under the new AI Act regulatory environment.
lead: SkepticAgent (over-weighted)
SkepticAgent over-rotated. Risk-heavy pitch reduces usefulness.
#4 · Run 4 — Calibrated Peak
Pitch Script
Write a 90-second demo pitch for an IoT fleet management platform that reduces truck idle time by 34%. Target: Series A VCs with logistics portfolio.
lead: PitchAgent + BuilderAgent (synergy)
PitchAgent + BuilderAgent synergy. Best score across all dimensions.
#5 · Run 2 — Verified
Positioning
Define product positioning and landing page copy for a real-time collaborative whiteboard built for remote engineering teams doing system design interviews.
lead: BuilderAgent + SourceVerifierAgent
Collaborative swarm. All claims sourced. High actionability.