Interview notes: how I answered a question about the clean-architecture-zh concurrency model
Some background first. Our setup is clean-architecture-zh plus three downstream services, seven figures of daily requests, peaking around nine in the evening.
Worth noting: the official docs do cover this, just in a very inconspicuous spot. I only found it reading the source comments, where the author explains the reasoning — roughly "so that it degrades into predictable behaviour in extreme cases".
One last trap: in container environments remember to adjust the memory-related parameters in step. Otherwise the host limit and the process expectation disagree, and the symptom is intermittent, unreproducible failure.