The workflow had been running for six hours when the first alert fired. The apply stage was exiting with a JSONDecodeError, the supervisor was restarting it, and every restart died on the same line of the same file. The queue showed zero unacked messages, which meant the work was already marked...
Source: [Dev.to](https://dev.to/robinzzz/the-agent-crash-looped-on-a-truncated-line-a-ledger-debugging-retrospective-3706)