Describe a CI/CD pipeline you personally designed or substantially rebuilt. What was unsafe or inefficient before, what decisions did you make, and how did deployment frequency, duration, or failure rate change?
Describe the most serious production incident you personally led. What alerted you, what evidence did you examine, what was the root cause, and what permanent changes did you implement?
Describe a database migration or major engine upgrade you supported in production. How did you validate compatibility and performance, run old and new systems together, and manage cutover and rollback?
Describe an AWS architecture decision you owned involving containers, serverless services, or asynchronous processing. What alternatives did you evaluate, what trade-offs did you accept, and what happened under production load?
Tell us about infrastructure or deployment code generated by an AI tool that appeared correct but was unsafe or incomplete. How did you identify the problem, validate the correction, and improve the AI workflow?
How did you use AI to build or improve CI/CD or the IaC? Give us some examples from your work. What did it do well and where did you need to intervene?