Skip to content

Research agent-device CI hardening for the observed infra flakes #56

Description

@thiagobrez

Part of #53

Question

Beyond the systematic iOS 26 failure, the remaining device-job failures are infrastructure flakes, each with a different cause: iOS runner ... already owned by another agent-device daemon, RUNNER_BUSY: still finishing a previous command that exceeded its execution watchdog, artifact restored but runner did not connect: dyld: Library not loaded: /usr/lib/libcurl.4.dylib, plus adb reverse-listener and emulator-boot flakes on Android.

The repo pins agent-device 0.20.10 (ADR 0006). What do the agent-device docs/repo (github.com/callstack/agent-device) recommend for CI daemon lifecycle, runner ownership, and artifact-restore issues — and do releases newer than 0.20.10 fix any of the observed errors? Check changelogs/issues for the three iOS error signatures, recommended CI setup (daemon-per-job vs shared, cleanup hooks, watchdog config), and whether the current prepare-agent-device-ios-runner.mjs dance (alert seeding, exit-75 sentinel, simulator erase between attempts) matches upstream guidance or works around already-fixed bugs. Deliver: a findings doc mapping each observed flake signature to a recommended remedy (upgrade, config change, or genuine local workaround still needed).

Activity

Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Metadata

Metadata

Assignees

No one assigned

    Projects

    No projects

      Milestone

      No milestone

      Relationships

      None yet

      Development

      No branches or pull requests

      Issue actions