Analyze complex, large-scale datasets to identify actionable insights and trends using advanced SQL, Python, and PySpark. Collaborate with stakeholders to translate business requirements into analytical solutions while ensuring data quality and process scalability.
SMASH, Who we are?
We are agents for tech professionals in Costa Rica and Colombia that help them build careers in the United States.
We believe in long-lasting relationships with our talent. We invest time getting to know them and understanding what they seek as their professional next step.
We aim to find the perfect match. As agents, we pair our talent with our US clients, not only by their technical skills but as a cultural fit. Our core competency is to find the right talent fast.
We purposefully move away from the “contractor” or “outsourcing” type of relationship. Our clients don’t want contractors or “just a service.” Neither does our talent.
Our Benefits
- Wellness Coverage
- Remote Work
- Birthday day off
- Recognition and rewards system
- Referrals Program
- Business skill coaching
- English classes for Smashers and relatives
- Learning opportunities
This position is Remote to work with a US Company; you will require to have Citizenship or a work permit from Costa Rica to apply for this role.
Role summary
We are looking for an experienced Technical Data Analyst with a strong background in data analysis, large-scale data processing, and cloud-based data environments.
The ideal candidate will bring hands-on expertise with SQL, Python, Spark, Spark SQL, PySpark, and Microsoft Azure and will be comfortable working with complex datasets and technical data workflows. Experience with Microsoft Fabric and within the Pharmaceutical, Life Sciences, or Insurance industries will be highly valued.
Responsibilities
- Analyze complex and large-scale datasets to identify trends, patterns, anomalies, and actionable insights.
- Develop advanced SQL queries to extract, transform, validate, and analyze data.
- Use Python and PySpark for data processing, transformation, automation, and analysis.
- Work with Apache Spark and Spark SQL to process and analyze large datasets efficiently.
- Support data analysis initiatives within Microsoft Azure environments.
- Collaborate with technical and business stakeholders to understand data requirements and translate them into analytical solutions.
- Perform data profiling, validation, reconciliation, and quality analysis to ensure accuracy and consistency.
- Investigate data issues and perform root-cause analysis across datasets and data workflows.
- Support the development and improvement of scalable data processes and analytical solutions.
- Document data requirements, business rules, analytical findings, and technical processes.
- Communicate technical findings clearly to both technical and non-technical stakeholders.
Requirements – Must-haves
- 5–6+ years of experience working on technical Data Analyst or closely related data projects.
- Strong hands-on SQL experience.
- Strong hands-on Python experience for data analysis and processing.
- Experience with Apache Spark.
- Strong experience with Spark SQL.
- Hands-on experience with PySpark.
- Experience working within Microsoft Azure data environments.
- Experience analyzing and transforming large and complex datasets.
- Strong understanding of data structures, data quality, and data validation.
- Strong analytical and problem-solving skills.
- Ability to independently investigate technical data issues and drive them toward resolution.
- Strong communication skills and ability to collaborate across technical and business teams.
Nice-to-haves (optional)
- Hands-on experience with Microsoft Fabric – strongly preferred.
- Pharmaceutical industry experience.
- Life Sciences industry experience.
- Insurance industry experience.
- Experience working with enterprise-scale cloud data platforms.
- Exposure to modern data engineering, data lake, or data warehouse architectures.
Languages