For Employers

Cogniify

Databricks Data Engineer - India

Posted 21 days ago
5-10 years experience
Apply Now

Please mention DailyRemote when applying

?/100
Resume Match Score

Match your resume skills with our AI powered skill match!

Get professional review

Questions interviewers often ask for this role, with sample answers.

Create a cover letter for this job

Upload your resume and we draft a letter for this exact role, tailored to what it asks for.

  • Tailored to this role
  • Based on your resume
  • Fully editable
AI Summary

The role involves supporting, maintaining, and optimizing existing Databricks-based data applications and production pipelines. You will collaborate with engineering teams to ensure system reliability, performance, and scalability while managing workspace configurations and data quality.

Location: Remote

Work Hours: EST (Eastern Standard Time) aligned

Experience: 6–9 years

Role Overview

We are looking for an experienced Databricks Data Engineer to support, maintain, and enhance existing Databricks-based data applications and pipelines. The role focuses on ensuring reliability, performance, and scalability of production Databricks workloads rather than building net-new platforms from scratch. You will work closely with data, analytics, and engineering teams to keep critical data applications stable, optimized, and aligned with business needs.

Key Responsibilities

  • Support and maintain existing Databricks applications, notebooks, jobs, and Delta Lake pipelines in production.

  • Monitor, troubleshoot, and resolve issues related to job failures, performance degradation, data quality, and cluster utilization.

  • Optimize existing Spark jobs, SQL queries, and Delta tables for cost, performance, and reliability.

  • Manage and improve Databricks workspace configurations, including clusters, job scheduling, access controls, and Unity Catalog (where applicable).

  • Implement and maintain data quality checks, logging, alerting, and basic observability for Databricks workloads.

  • Collaborate with stakeholders to understand requirements for enhancements or bug fixes on existing applications.

  • Perform incremental improvements, refactoring, and technical debt reduction on current Databricks solutions.

  • Ensure adherence to best practices around security, governance, and cost management within the Databricks environment.

  • Document existing pipelines, dependencies, and operational runbooks.

  • Participate in on-call or support rotations as needed to maintain production stability (within EST working hours).

Required Qualifications

  • 6–9 years of overall experience in data engineering, with strong hands-on experience in Databricks.

  • Solid proficiency in Apache Spark (PySpark and/or Scala) and SQL.

  • Proven experience supporting and optimizing production Databricks workloads (jobs, notebooks, Delta Lake, workflows).

  • Strong understanding of Delta Lake concepts (ACID transactions, time travel, optimization techniques such as Z-ordering, vacuum, optimize).

  • Experience with Databricks Job clusters, Interactive clusters, and performance tuning (partitioning, caching, shuffle optimization, autoscaling).

  • Familiarity with data modeling, ETL/ELT patterns, and production data pipeline support.

  • Experience working with cloud platforms (preferably Azure, AWS, or GCP) in the context of Databricks.

  • Ability to troubleshoot complex Spark and Databricks issues independently.

  • Strong communication skills and ability to work effectively in a remote, EST-aligned team.

Preferred Qualifications

  • Experience with Unity Catalog, Databricks SQL, or Lakehouse architecture.

  • Knowledge of CI/CD practices for Databricks (e.g., Databricks Asset Bundles, Git integration, Terraform/ARM templates).

  • Familiarity with orchestration tools (Airflow, Azure Data Factory, or Databricks Workflows).

  • Exposure to data quality frameworks, monitoring tools, or cost optimization initiatives on Databricks.

  • Experience supporting analytics or BI teams consuming Databricks data products.

Work Arrangement

Fully remote

Must be available and productive during EST business hours

Collaborative remote environment with regular syncs and support responsibilities

Automatically Apply to the Best Remote Jobs

Stop the endless job search. Our AI finds and applies to the best jobs for you.

Try it Now
Keep looking

Similar Jobs

See all Remote Software Development jobs →

GTM Engineer (m/f/d)

Full Time Germany Software Development

Analytics Engineer, Network Intelligence

Full Time Georgia Software Development

Security Engineer - Vulnerability Management

Full Time United States Software Development

Senior Kafka/RabbitMQ Platform Engineer

Freelance Bosnia and Herzegovina, Germany, Oman +4 more Software Development

Sr Analytics Engineer

Other, Full Time United States $110K - $160K per year Software Development

Architect - R01571180

Full Time United States $160 - $165 per year Software Development
Apply Now

Personalize your Remote Job Search in 3 Easy Steps!

Featuring 218,224+ Jobs in Data Engineer

Answer easy questions

Answer easy questions

218,224+ jobs across 15+ categories

Get your best job matches

Get your best job matches

Only hand-screened, legit jobs

Find a remote job faster

Find a remote job faster

No ads, scams, or junk

“I was the first applicant for a remote marketing position that got listed on the company website the same day I applied. Had an interview within 48 hours!”

Sarah J. — Sarah J. · Marketing Manager ★★★★★ Verified