Ch

Chunk sidecars for validating agent-generated code before pushing to CI

Hacker News

Chunk sidecars for validating agent-generated code before pushing to CI

Hi HN! My name is Olaf, I work at CircleCI as a technology advisor in the CTO office, came in through the acquisition of my company Vamp.io (progressive delivery for microservices on k8s) in 2021. Wanted to hear the HN community feedback and thoughts on a project we think could be very interesting when adding AI coding agents to the SDLC and your CI pipelines. Our team at CircleCI built Chunk sidecars after repeatedly running into the same issue internally: by the time our CI catches a failure, the agent has already moved on and most of the useful context is gone. The basic idea of Chunk sidecars is to move fast lightweight validation into the inner development loop. Chunk sidecars runs scoped “microbuilds” inside a lightweight microVM that mirrors your CI environment. It tries to auto-detect your stack and test commands, syncs changes from the agent session, and runs validations before commit/push. A few implementation details that might be interesting: validation hooks trigger automatically during agent stop/evaluation events warm snapshots keep startup times low validations run against environments matching the CI stack instead of local machine state microbuilds only run the relevant slice instead of the entire pipeline In our own experiments we measured: ~27 second average microbuild compute ~5 minutes total billable compute for equivalent full CI runs 3x–5x lower token usage in retry loops The compute comparison is billable compute vs billable compute, not wall clock time. Full CI pipelines were parallelized. The 27s is with warm snapshots — first-time setup takes about 15 minutes. We tested this on our own pipeline, not a large corpus. Larger repos with heavier deps will vary. Under the hood it's currently Firecracker microVMs, running on E2B infrastructure. Current spec: 4 CPU, 8GB RAM (comparable to a Docker large). Things can change in the future depending on feedback and learnings. Short demo video (YT) here: https://circle.ci/4dq9fph Blog post: https://circleci.com/blog/chunk-sidecars/ Chunk CLI GitHub repo: https://github.com/CircleCI-Public/chunk-cli This works with any CircleCI account (including the free one), and integrates with Claude Code, Codex, Cursor, or your own agents. The project is open source and also has features that work without CircleCI connected. Simply install the Chunk CLI and run "chunk init" and the sidecar auto-detects your stack and test commands. Would love all feedback, especially from people already experimenting with agentic workflows. We're especially curious whether others are seeing the same CI failure rate pattern and "widening gap" between inner and outer dev/SDLC loop with agent-generated code?

Share card

Actual performance

1points
2comments
Did not reach leaderboard

Launch Intel predictions

Analyze your own launch →
Product HuntOn track for Day 1 leaderboard · Strong signals: mac, agents, agent · Missing: macos, model, apple
98%98% predicted probability of success on Product Hunt, based on ML models trained on real launch data.
best fitHighest predicted score across all platforms for this description.
Indie HackersFits the IH revenue-focused audience · Strong signals: para, including · Missing: supports, reddit linkedin, podcasting
78%78% predicted probability of success on Indie Hackers, based on ML models trained on real launch data.
Hacker NewsMay not resonate with HN audience · Strong signals: lua, open source, ide · Missing: https docs, excited, just released
48%48% predicted probability of success on Hacker News, based on ML models trained on real launch data.
nativeThis product was originally launched on this platform.
TrustMRRLess likely to generate early MRR · Strong signals: video, para · Missing: mobile apps, ios, personal
32%32% predicted probability of success on TrustMRR, based on ML models trained on real launch data.
AppSumoMay struggle as an AppSumo deal · Missing: plus, platform, intuitive
26%26% predicted probability of success on AppSumo, based on ML models trained on real launch data.
Acquire.comPre-revenue stage for this audience · Missing: arr, mrr, revenue
23%23% predicted probability of success on Acquire.com, based on ML models trained on real launch data.
BetaListMay not resonate with beta-testers · Missing: web3, chat, crypto
0%0% predicted probability of success on BetaList, based on ML models trained on real launch data.

Correct prediction on native model

Similar products

Ru
Run LLM-generated code in sandboxed envs52%Launch Intel prediction score: how likely this product is to succeed on its source platform, based on its name, tagline, and description.

Run LLM-generated code in sandboxed envs

Hacker News2
No
No Code CI50%Launch Intel prediction score: how likely this product is to succeed on its source platform, based on its name, tagline, and description.

No Code CI

Hacker News1
No
No code CI for Raku modules52%Launch Intel prediction score: how likely this product is to succeed on its source platform, based on its name, tagline, and description.

No code CI for Raku modules

Hacker News3
A
A go package to capture stdout and stderr generated by your code43%Launch Intel prediction score: how likely this product is to succeed on its source platform, based on its name, tagline, and description.

A go package to capture stdout and stderr generated by your code

Hacker News1
Mu
Mutation testing to secure Cursor generated code39%Launch Intel prediction score: how likely this product is to succeed on its source platform, based on its name, tagline, and description.

Mutation testing to secure Cursor generated code

Hacker News1
Cl
ClamBot – AI agent that runs all LLM-generated code in a WASM sandbox38%Launch Intel prediction score: how likely this product is to succeed on its source platform, based on its name, tagline, and description.

ClamBot – AI agent that runs all LLM-generated code in a WASM sandbox

Hacker News4
CI
CI-Wallboard55%Launch Intel prediction score: how likely this product is to succeed on its source platform, based on its name, tagline, and description.

CI-Wallboard

Hacker News4
In
Integrating Packagr with Gitlab CI32%Launch Intel prediction score: how likely this product is to succeed on its source platform, based on its name, tagline, and description.

Integrating Packagr with Gitlab CI

Hacker News2
Pr
Protecting Your Gitlab CI Piplines36%Launch Intel prediction score: how likely this product is to succeed on its source platform, based on its name, tagline, and description.

Protecting Your Gitlab CI Piplines

Hacker News1
Ar
ArchSentry – Deterministic architectural enforcement for CI32%Launch Intel prediction score: how likely this product is to succeed on its source platform, based on its name, tagline, and description.

ArchSentry – Deterministic architectural enforcement for CI

Hacker News3