Oliver Hughes
@oliver_hughes • 2 weeks ago
Writes a short post-interview thank-you email that references the conversation, repairs a fumbled answer, adds one useful thing, and schedules a check-in.
interviewerrole_and_stageinterview_notesvalue_addinterview_date{{interviewer}}{{role_and_stage}}{{interview_notes}}{{value_add}}{{interview_date}}interviewer: Dana Whitfield, Engineering Manager, Platform team role_and_stage: Senior Site Reliability Engineer at Kestrel Health, second-round technical interview (60 minutes) interview_notes: We talked about their move from self-managed Kubernetes to managed EKS. She's worried about on-call load: a team of 5 handling about 2 incidents a week. I blanked on how I'd set SLOs for a batch pipeline and gave a vague answer about uptime. They use Grafana and PagerDuty. She said they'll decide by the end of next week. value_add: I wrote an incident-review runbook template at my last job and could share a sanitized version. I've also worked out the better answer: for batch pipelines, SLOs should be about freshness and completeness, e.g. 99% of daily runs finish by 6 a.m. with all records present, not uptime. interview_date: Tuesday, September 22, 2026