HubSpot · Task agent
Sidekick
Sidekick is HubSpot's internal AI code reviewer, built by its Developer Experience AI team. It runs on every pull request in GitHub. A second Judge Agent evaluates each draft comment for succinctness, accuracy, and actionability, and only comments that pass are posted.
- Approach type
- Task agent
- Work
- Code review
- Human involvement
- Drafts reviewed
- Invocation
- Event driven
- Interfaces
- Github
- Deployment stage
- Scaled
- Evidence strength
- Detailed primary
- Entry reviewed
Where people stay involved
- pull request → AI review commentsWork product review · Level 3
Level 3 for pull request → AI review comments; human attention boundary: work-product-review.
- Observation date
- 2026-03
Implementation details
Aviator, an internal Java agent framework; replaced the earlier Claude Code review implementation on Crucible
unknown
Aviator framework for precise tool control
github
Reported results and limitations
The catalog records what the sources report, with the scope and the denominator of every figure. A qualification below limits the figure it sits under.
Reported metrics
Reviews every pull request and cut engineer feedback time by 90%
The source does not report the denominator of this figure.
- Reported by
- HubSpot
- Scope
- Time for engineers to receive code feedback from Sidekick; not overall PR completion time
- Observation date
- 2026-03
Reviews every pull request
- Reported by
- HubSpot
- Scope
- Pull-request coverage after the six-month rollout
- Denominator
- HubSpot pull requests
Engineer feedback time cut by 90%
The source does not report the denominator of this figure.
- Reported by
- HubSpot
- Scope
- Time for engineers to receive code feedback from Sidekick; not overall PR completion time
Over 80% thumbs-up reaction rate on review feedback during the preceding couple of months
- Reported by
- HubSpot
- Scope
- Developer emoji reactions on review comments during the preceding couple of months
- Denominator
- Thumbs-up and thumbs-down reactions; not all developers or all reviews
- Method
- Emoji reactions and replies on review comments
Sources and research details
Citations link to the original publisher. Each source also keeps a preserved copy in the repository, so a changed or removed page stays checkable.
- Automated code review, the 6-month evolutionhttps://product.hubspot.com/blog/automated-code-review-the-6-month-evolution
- Cloud coding agents at HubSpothttps://product.hubspot.com/blog/cloud-coding-agents-at-hubspot
Research details for every claim on this page
- Summary
- Statement type
- Fact
- Provenance
- Reported
- Confidence
- High
- Confidence reason
- A linked first-party source states the claim.
- SupportsAutomated code review, the 6-month evolutionPreserved content.md, lines 14, 24–45, 75–87
- ContextualizesCloud coding agents at HubSpotPreserved content.md, lines 30–57 (earlier Crucible implementation)
- Headline claim
- Statement type
- Metric
- Provenance
- Reported
- Confidence
- Medium
- Confidence reason
- HubSpot reported the figures in its own engineering blog without independent verification.
- Reported by
- HubSpot
- Scope
- Time for engineers to receive code feedback from Sidekick; not overall PR completion time
- Denominator
- Not reported
- Method
- Not reported
- Observation date
- 2026-03
- SupportsAutomated code review, the 6-month evolutionPreserved content.md, lines 14–18
- Harness
- Statement type
- Fact
- Provenance
- Reported
- Confidence
- High
- Confidence reason
- A linked first-party source states the claim.
- SupportsAutomated code review, the 6-month evolutionPreserved content.md, lines 24–45
- Sandbox
- Statement type
- Inference
- Provenance
- Catalog judgment
- Confidence
- Medium
- Confidence reason
- Current review runs on Aviator; its execution isolation is not specified. Crucible Kubernetes workloads describe the predecessor implementation.
- SupportsAutomated code review, the 6-month evolutionPreserved content.md, lines 39–45
- Tool access
- Statement type
- Fact
- Provenance
- Reported
- Confidence
- High
- Confidence reason
- A linked first-party source states the claim.
- Interfaces
- Statement type
- Fact
- Provenance
- Reported
- Confidence
- High
- Confidence reason
- A linked first-party source states the claim.
- Key observation
- Statement type
- Metric
- Provenance
- Reported
- Confidence
- Medium
- Confidence reason
- HubSpot reported the figures in its own engineering blog.
- Reported by
- HubSpot
- Scope
- Pull-request coverage after the six-month rollout
- Denominator
- HubSpot pull requests
- Method
- Not reported
- Observation date
- Not reported
- SupportsAutomated code review, the 6-month evolutionPreserved content.md, lines 14
- Key observation
- Statement type
- Metric
- Provenance
- Reported
- Confidence
- Medium
- Confidence reason
- HubSpot reported the figures in its own engineering blog.
- Reported by
- HubSpot
- Scope
- Time for engineers to receive code feedback from Sidekick; not overall PR completion time
- Denominator
- Not reported
- Method
- Not reported
- Observation date
- Not reported
- SupportsAutomated code review, the 6-month evolutionPreserved content.md, lines 14–18
- Key observation
- Statement type
- Metric
- Provenance
- Reported
- Confidence
- Medium
- Confidence reason
- HubSpot reported the figures in its own engineering blog.
- Reported by
- HubSpot
- Scope
- Developer emoji reactions on review comments during the preceding couple of months
- Denominator
- Thumbs-up and thumbs-down reactions; not all developers or all reviews
- Method
- Emoji reactions and replies on review comments
- Observation date
- Not reported
- SupportsAutomated code review, the 6-month evolutionPreserved content.md, lines 99–116
- Operating model assessment
- Statement type
- Inference
- Provenance
- Catalog judgment
- Confidence
- Medium
- Confidence reason
- The engineering blog describes engineers acting on Sidekick review comments, which locates human attention at work-product review.
- Observation date
- 2026-03