For Employers

mpathic

Remote Job Openings at mpathic (15)

Mental Health Expert (Brazil)

Brazil 5-10 yrs exp Healthcare

The role involves designing and testing AI chat scenarios by roleplaying clinical interactions to ensure cultural authenticity and safety. Experts will also collaborate with researchers to develop behavioral taxonomies, rubrics, and mental health policies for large language models.

Program Manager, Human Data Operations (Utah)

United States $120K - $150K per year 5-10 yrs exp Product

The Program Manager will lead end-to-end delivery of complex AI safety and human data operations, managing cross-functional teams and customer engagements. They are responsible for establishing project plans, ensuring quality standards, and acting as the primary point of contact for stakeholders.

Technical Product Manager

United States $120K - $160K per year 5-10 yrs exp Product

The Technical Product Manager will own the end-to-end delivery of platform features, from initial requirements and scoping to launch and feedback. They will act as the connective tissue between engineering, research, sales, and delivery teams to ensure alignment on product vision and quality standards.

Technical Program Manager, Life Sciences Customer Delivery

United States $100K - $120K per year 5-10 yrs exp Product

The Technical Program Manager will lead the implementation and delivery of mpathic's solutions across clinical research programs and healthcare studies. They will serve as the primary operational point of contact, managing project timelines, site training, and cross-functional coordination between engineering, product, and clinical teams.

Mental Health Expert (Contract)

United States $30 - $200 per hour 5-10 yrs exp Healthcare

The role involves collaborating on AI safety initiatives by roleplaying clinical scenarios, identifying behavioral edge cases, and providing expert feedback on mental health policy. You will also design testing scenarios, conduct qualitative analysis, and develop behavioral taxonomies to ensure accurate AI responses to psychological distress.

Radiology Specialist (Contract)

United States 2-5 yrs exp Others

Design and evaluate clinically realistic tasks to benchmark multimodal AI systems in the radiology domain. Provide expert reference answers and detailed feedback to improve AI diagnostic reasoning and instruction-following.

Healthcare Systems Specialist (Contract)

United States 5-10 yrs exp Healthcare

Evaluate and improve AI models supporting clinical operations and administrative workflows to ensure accuracy and practical application. Review real-world healthcare scenarios and provide expert feedback on AI-generated responses regarding EHRs and regulatory standards.

Primary Care Experts (Contract)

United States 5-10 yrs exp Others

Design realistic clinical scenarios and evaluate AI-generated responses to ensure clinical accuracy and safety. Provide actionable feedback to improve the reliability and trustworthiness of healthcare AI models.

Pharmaceutical Expert (Contract)

United States 5-10 yrs exp Others

Evaluate AI-generated prior authorization recommendations and specialty medication requests for accuracy and alignment with payer requirements. Provide actionable feedback to improve AI performance on complex pharmacy authorization workflows.

Cybersecurity for AI Safety (Contract)

United States 5-10 yrs exp Software Development

Design expert-level cybersecurity prompts and evaluate AI-generated responses for technical accuracy and safety. Provide actionable feedback to improve the reliability and security of frontier AI systems.

Psychiatry Specialist (Contract)

United States 2-5 yrs exp Others

Design clinically realistic psychiatric scenarios and prompts to test and improve the reasoning of large language models. Evaluate AI-generated responses for accuracy and safety while providing structured feedback to researchers.

Financial Risk & Safety Specialist (AI Systems)

United States $30 - $200 per hour 5-10 yrs exp Software Development

Responsibilities focus on reviewing AI-generated financial content and conversations to identify risks such as misleading guidance, overconfidence, and inappropriate agreement, while also participating in adversarial probing exercises to surface failure modes.