For Employers

Smart Working Solutions

Lead Data Engineer - Founding Member (Contract, Full-Time) [HR208] (UK)

Posted an hour ago
5-10 years experience
Apply Now

Please mention DailyRemote when applying

?/100
Resume Match Score

Match your resume skills with our AI powered skill match!

Get professional review

Questions interviewers often ask for this role, with sample answers.

Create a cover letter for this job

Upload your resume and we draft a letter for this exact role, tailored to what it asks for.

  • Tailored to this role
  • Based on your resume
  • Fully editable
AI Summary

The Lead Data Engineer will architect and build scalable data infrastructure to support AI products and real-time data pipelines. They will also define engineering standards and collaborate with cross-functional teams to ensure data reliability and performance.

About Smart Working

At Smart Working, we believe your job should not only look right on paper but also feel right every day. This isn’t just another remote opportunity - it’s about finding where you truly belong, no matter where you are. From day one, you’re welcomed into a genuine community that values your growth and well-being. Our mission is simple: to break down geographic barriers and connect skilled professionals with outstanding global teams and products for full-time, long-term roles. We help you discover meaningful work with teams that invest in your success, where you’re empowered to grow personally and professionally.

Join one of the highest-rated workplaces on Glassdoor and experience what it means to thrive in a truly remote-first world.

About the Role

We are looking for a Lead Data Engineer to build and lead the data infrastructure powering an intelligent AI assistant  platform. This role will architect and scale the systems that power our AI products, from real-time data pipelines and analytics infrastructure to vector databases and machine learning data workflows.

You will work closely with AI engineers, backend engineers, and product teams to ensure our platform can process large volumes of operational data reliably and intelligently. You will define our data architecture, tooling, and engineering standards, and play a key role in building the foundations of the future data team.

\n


Responsibilities
  • Architect and build scalable data pipelines and infrastructure to support AI and product systems.
  • Design and maintain data ingestion, transformation and storage architectures for operational and AI workloads.
  • Develop and manage batch and real-time data pipelines.
  • Build and optimise systems for vector search, retrieval and machine learning data pipelines.
  • Ensure data reliability, security and governance across the platform.
  • Collaborate with AI and backend engineering teams to support training, inference and product features.
  • Implement monitoring, observability and data quality frameworks.
  • Optimise the performance of large-scale datasets and query systems.
  • Contribute to technical architecture decisions and long-term data strategy.
  • Act as the founding data hire, defining culture, standards and the hiring bar for the data function as it scales.
  • Partner directly with founders and product leadership to translate data capabilities into product decisions.


Requirements
  • 7+ years of professional experience, with the majority of that experience in dedicated data engineering roles.
  • Strong experience designing and building data pipelines and distributed data systems.
  • Experience working with relational databases, with PostgreSQL preferred, although MySQL or similar is acceptable.
  • Experience working with NoSQL databases.
  • Strong programming experience in Python.
  • Demonstrated ability to make and justify architectural decisions, rather than only implementing them.
  • Experience building scalable backend systems.
  • Experience designing data models and storage architectures.
  • Strong understanding of data processing performance and optimisation.
  • Experience with some of the following data frameworks and infrastructure technologies is highly desirable: Apache Spark, Apache Airflow, Kafka, and Elasticsearch or OpenSearch.
  • Experience with relevant database technologies is highly desirable, including PostgreSQL, MongoDB, and vector databases such as Qdrant, Milvus or pgvector.
  • Experience with Python data-processing libraries such as Pandas or Polars is highly desirable.


Nice to Have
  • Experience working on AI or machine learning platforms.
  • Familiarity with stream processing and event-driven architectures.
  • Experience with cloud infrastructure such as GCP, AWS or Azure.
  • Experience working in high-growth startups or early-stage companies.
  • Experience with vector databases used in modern AI systems.


\n

Automatically Apply to the Best Remote Jobs

Stop the endless job search. Our AI finds and applies to the best jobs for you.

Try it Now
Keep looking

Similar Jobs

See all Remote Software Development jobs →

Senior Software Developer

Full Time United States $76100 - $108K per year Software Development

Senior Software Engineer

Full Time United States $119K - $234K per year Software Development

Principal Software Engineer

Full Time United States $142K - $304K per year Software Development

Principal Software Engineer

Full Time Colombia Software Development

Web Developer - Java Backend Developer

Full Time India Software Development

Ingeniero/a Cloud GCP ID89421

Full Time Spain Software Development
Apply Now

Personalize your Remote Job Search in 3 Easy Steps!

Featuring 215,940+ Jobs in Data Engineer

Answer easy questions

Answer easy questions

215,940+ jobs across 15+ categories

Get your best job matches

Get your best job matches

Only hand-screened, legit jobs

Find a remote job faster

Find a remote job faster

No ads, scams, or junk

I was the first applicant for a remote marketing position that got listed on the company website the same day I applied. Had an interview within 48 hours!”

Sarah J. — Sarah J. · Marketing Manager ★★★★★ Verified