For Employers

PulsePoint

Site Reliability Engineer, K8s

Posted 5 months ago
2-5 years experience
Apply Now

Please mention DailyRemote when applying

?/100
Resume Match Score

Match your resume skills with our AI powered skill match!

Get professional review

Questions interviewers often ask for this role, with sample answers.

Create a cover letter for this job

Upload your resume and we draft a letter for this exact role, tailored to what it asks for.

  • Tailored to this role
  • Based on your resume
  • Fully editable
AI Summary

The Site Reliability Engineer will own service health, incident response, and infrastructure monitoring for GCP-based APIs and data services. They will also manage cloud cost visibility and collaborate with engineering teams to ensure operational excellence and security.

WebMD and its affiliates is an Equal Opportunity/Affirmative Action employer and does not discriminate on the basis of race, ancestry, color, religion, sex, gender, age, marital status, sexual orientation, gender identity, national origin, medical condition, disability, veterans status, or any other basis protected by law.

Position Overview
Our BI team runs a set of GCP-based APIs and data services that a lot of internal products depend on. As we've grown, keeping things running has increasingly been a side responsibility for engineers who are primarily building features — and that's not sustainable. We're looking for an SRE to own that space: service health, incident response, infrastructure monitoring, and making sure we're not blindly burning cloud budget.
The Site Reliability Engineer will ensure the availability, performance, and security of the Business Intelligence team's GCP-hosted APIs and data infrastructure. This role is responsible for proactive monitoring, incident response, and continuous improvement of platform reliability across a cloud-native stack. The engineer will work closely with backend and data engineers to maintain service health and drive operational excellence. This position also carries responsibility for GCP cost visibility, helping the team track and optimize cloud spend through structured monitoring and alerting.
Responsibilities
  • Monitor and maintain uptime of GCP-hosted APIs and services, keeping performance within agreed targets
  • Lead incident response for BI platform services — triage, resolve, and follow up with post-mortems that actually prevent recurrence
  • Build and manage observability infrastructure: dashboards, alerts, and logging across GCP services
  • Track GCP cloud spend and set up cost alerting to flag anomalies before they become problems
  • Review and fix security gaps — IAP configs, service account permissions, API access controls
  • Work with data and backend engineers to shore up reliability of data pipelines and BigQuery workflows
  • Contribute to infrastructure-as-code and help keep deployments documented and reproducible
Qualifications 
  • 2+ years in a Site Reliability, DevOps, or Cloud Infrastructure role in a production environment
  • Bachelor's degree in Computer Science, Engineering, or related field, or equivalent hands-on experience
  • Practical experience with GCP — Cloud Run, API Gateway, and BigQuery in particular
  • Experience with monitoring and observability tooling (Cloud Monitoring, Datadog, or similar)
  • Solid grasp of cloud security fundamentals — IAM, network controls, access management
  • Proficiency with Git and version control in a team setting
Please list the preferred skills here:
  • CI/CD pipelines and deployment automation (GitHub Actions, Cloud Build, or similar)
  • Terraform or other infrastructure-as-code tools
  • Python for scripting or automation
  • MySQL, Spanner, or BigQuery at any meaningful depth
  • GCP cost management and spend optimization
  • Experience with dbt or Looker
  • Comfortable working across CET/EST hours in a distributed team



Automatically Apply to the Best Remote Jobs

Stop the endless job search. Our AI finds and applies to the best jobs for you.

Try it Now
Keep looking

Similar Jobs

See all Remote Software Development jobs →

Senior Salesforce Developer (Remote)

Full Time India Software Development

Senior Software Application Developer

Full Time Canada 110K - 140K per year Software Development

Senior Spectrum Engineer

Full Time Canada, United States 151K - 176K per year Software Development

Senior Project Manager - Cloud & Compliance Advisory

Full Time United States $85000 - $141K per year Software Development

Application Developer - Senior

Full Time United States $120K - $175K per year Software Development

Quality Assurance Manager - Senior

Full Time United States $120K - $165K per year Software Development
Apply Now

Personalize your Remote Job Search in 3 Easy Steps!

Featuring 219,042+ Jobs in Site Reliability Engineer

Answer easy questions

Answer easy questions

219,042+ jobs across 15+ categories

Get your best job matches

Get your best job matches

Only hand-screened, legit jobs

Find a remote job faster

Find a remote job faster

No ads, scams, or junk

“I was the first applicant for a remote marketing position that got listed on the company website the same day I applied. Had an interview within 48 hours!”

Sarah J. — Sarah J. · Marketing Manager ★★★★★ Verified