← back to queue

Site Reliability Engineer, Cloud Cost Utilization

GitLab · Remote, US · greenhouse · fit score 33.0
Open / Apply ↗
⬇ Tailored résumé Queue →

Job description

<div class="content-intro"><p>GitLab is the intelligent orchestration platform for DevSecOps. GitLab enables organizations to increase developer productivity, improve operational efficiency, reduce security and compliance risk, and accelerate digital transformation. More than 50 million registered users and more than 50% of the Fortune 100* trust GitLab to ship better, more secure software faster.</p> <p>The same principles built into our products are reflected in how our team works: we embrace AI as a core productivity multiplier, with all team members expected to incorporate AI into their daily workflows to drive efficiency, innovation, and impact. GitLab is where careers accelerate, innovation flourishes, and every voice is valued. Our high-performance culture is driven by our values&nbsp;and continuous knowledge exchange, enabling our team members to reach their full potential while collaborating with industry leaders to solve complex problems. <a href="https://www.youtube.com/watch?v=OuZIb5zszQI">Co-create the future with us</a> as we build technology that transforms how the world develops software.</p> <p>*<em>Fortune 500® is a registered trademark of Fortune Media IP Limited, used under license. Claim based on GitLab data. Fortune 100 refers to the top 20% ranked companies in the 2025 Fortune 500 list, published in June 2025. Fortune and Fortune Media IP Limited are not affiliated with, and do not endorse products or services of GitLab.</em></p></div><p class="font-cla

✍️ Tailored application (cached — grounded in your resume)

61 / 100 honest fit

Daniel brings strong reliability engineering fundamentals — AWS (Lambda, S3, DynamoDB), Python automation, CI/CD pipelines, monitoring, and deep root-cause analysis — plus hands-on operation of self-hosted production infrastructure with watchdogs and automated data pipelines, all directly relevant to an SRE role. The clear gap is that he has no explicit SRE title or dedicated cloud-cost/FinOps experience, and his cloud depth is AWS rather than the multi-cloud/Kubernetes/Terraform stack many SRE teams use; his reliability-first mindset and infra ownership are a genuine on-ramp, but the cost-utilization specialization would be new.

Tailored resume highlights

"Why this company?" (draft — personalize before sending)

GitLab expects every engineer to treat AI as a core productivity multiplier, and that's already how I work — I use Claude Code daily to design, ship, and operate roughly a dozen production apps end-to-end on my own infrastructure. That same hands-on ownership, plus a career in safety-critical medical-device software where reliability is a hard requirement, is why an SRE role at GitLab fits me: I care about systems that stay up and behave correctly. Working on cloud cost utilization would let me apply my AWS, automation, and monitoring background to making that reliability efficient at scale.

Cover letter (review before sending)

Dear GitLab Hiring Team, I'm applying for the Site Reliability Engineer, Cloud Cost Utilization role. My background is reliability engineering in a domain where correctness is non-negotiable: at Medtronic I built backend services, AWS solutions (Lambda, S3, DynamoDB), and Python automation frameworks integrated into Jenkins CI/CD, and led root-cause analysis of complex cross-team issues in safety-critical medical-device software. Beyond my day job, I design, deploy, and operate my own production services on a self-hosted Linux server — gunicorn behind hardened endpoints, per-IP rate limiting, automated cron pipelines, and self-healing watchdogs — so I understand cloud resources and reliability from an owner's perspective, not just a consumer's. I also use LLM/agentic tooling (Claude Code) daily to ship end-to-end, which fits GitLab's expectation that AI be a core productivity multiplier. I'd be new to dedicated cloud-cost optimization, but the SRE fundamentals — monitoring, automation, fault tolerance, and disciplined debugging — are exactly how I already work. I'd welcome the chance to bring that to GitLab. Sincerely, Daniel Maynard

ATS keywords you have but should add

Site Reliability Engineering (SRE)cloud cost optimizationobservabilitymonitoringinfrastructure as codeproduction operationsuptimefault toleranceLinuxcloud resource utilization

Everything is grounded strictly in your resume — review before you submit. Nothing is auto-sent.