For Employers

Exavalu Solutions India Pvt Ltd

Senior Observability Engineer / Platform Engineer

Posted 22 days ago
5-10 years experience
Apply Now

Please mention DailyRemote when applying

?/100
Resume Match Score

Match your resume skills with our AI powered skill match!

Get professional review

Create a cover letter for this job

Upload your resume and we draft a letter for this exact role, tailored to what it asks for.

  • Tailored to this role
  • Based on your resume
  • Fully editable
AI Summary

Design, implement, and maintain enterprise observability platforms covering metrics, logs, traces, and events. Collaborate with engineering and SRE teams to improve reliability, performance, and availability through proactive monitoring and incident management.

This is a remote position.

Senior Observability Engineer / Platform Engineer (Offshore)

Location Offshore (India) Experience 6-10 Years

Role Summary

We are looking for a hands-on Observability Engineer with strong experience in cloud-native platforms, monitoring, logging, alerting, and operational excellence. The ideal candidate will have experience building and managing enterprise observability solutions across Kubernetes and public cloud environments, enabling proactive monitoring, incident response, and reliability engineering practices.

Key Responsibilities

Design, implement, and maintain enterprise observability platforms covering metrics, logs, traces, and events. Build observability solutions using tools such as Prometheus, Grafana, OpenSearch, Splunk, Elastic, Datadog, Dynatrace, New Relic, or equivalent platforms. Develop dashboards, SLOs, SLIs, alerting rules, and service health monitoring frameworks. Integrate monitoring and observability capabilities within Kubernetes and cloud-native environments. Enable incident management, root cause analysis, and operational troubleshooting through observability best practices. Automate monitoring configuration and platform onboarding using Infrastructure-as-Code and CI/CD pipelines. Collaborate with engineering, platform, SRE, and operations teams to improve reliability, performance, and availability. Support observability maturity initiatives including distributed tracing, AIOps, and intelligent alerting.



Requirements

Required Skills

6+ years of experience in Platform Engineering, SRE, DevOps, Cloud Operations, or Observability Engineering.

Strong expertise in: Prometheus Grafana Splunk / OpenSearch / Elastic Distributed tracing solutions (Jaeger, Tempo, OpenTelemetry, etc.) Good understanding of Kubernetes, Docker, and containerized workloads.

Experience with AWS, Azure, or GCP environments. Knowledge of incident management, alert tuning, and troubleshooting production environments. Experience with scripting and automation using Python, Shell, or similar languages.

Familiarity with CI/CD and Infrastructure-as-Code tools such as Terraform, Jenkins, GitHub Actions, or ArgoCD. Preferred Skills Exposure to SRE practices, error budgets, SLO/SLI frameworks.

Experience with AIOps, automated remediation, or intelligent incident response. Knowledge of OpenTelemetry implementation.

Working experience in enterprise-scale production platforms. Exposure to security and governance considerations in cloud-native environments.



Benefits

Diversity Inclusion:

At Exavalu, we are committed to building a diverse and inclusive workforce. We welcome applications for employment from all qualified candidates, regardless of race, color, gender, national or ethnic origin, age, disability, religion, sexual orientation, gender identity or any other status protected by applicable law. We nurture a culture that embraces all individuals and promotes diverse perspectives, where you can make an impact and grow your career.

Exavalu also promotes flexibility  depending on the needs of employees, customers and the business. It might be part-time work, working outside normal 9-5 business hours or working remotely. We also have a welcome back program to help people get back to the mainstream after a long break due to health or family reasons.



Automatically Apply to the Best Remote Jobs

Stop the endless job search. Our AI finds and applies to the best jobs for you.

Try it Now
Keep looking

Similar Jobs

See all Remote Software Development jobs →

Data Engineer

Full Time, Contract United States $103K - $143K per year Software Development

2109 Devops Engineer - Cloud

Full Time Brazil Software Development

AI Product Architect

Full Time United States Software Development

AD/Sr. AD, Business Analytics - Oncology (Remote)

Full Time United States $170K - $269K per year Software Development

Staff Data Engineer

Full Time United States Software Development

IT Systems Engineer

Full Time United States Software Development
Apply Now

Personalize your Remote Job Search in 3 Easy Steps!

Featuring 220,124+ Jobs in Platform Engineer

Answer easy questions

Answer easy questions

220,124+ jobs across 15+ categories

Get your best job matches

Get your best job matches

Only hand-screened, legit jobs

Find a remote job faster

Find a remote job faster

No ads, scams, or junk

I was the first applicant for a remote marketing position that got listed on the company website the same day I applied. Had an interview within 48 hours!”

Sarah J. — Sarah J. · Marketing Manager ★★★★★ Verified