For Employers

uRun

Remote Job Openings at uRun (6)

Don't see your role? Reach out anyway.

United States $200K - $400K per year 10+ yrs exp Others

Build and architect the Universal Runtime (uRun) layer to enable real-time, stateful inference for AI. Lead technical direction or execute high-velocity engineering to solve ambiguous infrastructure problems.

Founding Engineer - Site Reliability

United States $185K - $285K per year 5-10 yrs exp Software Development

As the founding SRE, you will define the reliability culture, observability stack, and incident response processes from scratch. You will partner with ML infrastructure engineers to ensure the stability and scalability of the interactive AI inference cloud.

Founding Engineer - ML Performance

United States $250K - $395K per year 5-10 yrs exp Software Development

Develop custom CUDA kernels and optimize model inference to achieve sub-50ms latency and 10-100x performance gains. Own the end-to-end inference pipeline, focusing on GPU utilization, memory bandwidth, and distributed memory optimizations.

Founding ML infrastructure Engineer

United States $200K - $350K per year 10+ yrs exp Software Development

Design and scale a GPU compute platform supporting 1,000+ clusters to enable real-time, stateful AI inference. Own the full infrastructure stack from bare metal to model serving, including resource orchestration and production reliability.

Founding Engineer - Platform

United States $250K - $350K per year 10+ yrs exp Software Development

Design and own the scalable, low-latency infrastructure powering the uRun real-time inference runtime. You will manage GPU-heavy workloads, streaming pipelines, and define platform standards for security and observability.

Founding Engineer - Software

United States $200K - $350K per year 5-10 yrs exp Software Development

Build and maintain the backend services, APIs, and core application systems that power the uRun real-time inference runtime. Design scalable systems for real-time interaction and session state while shaping the overall architecture for fault tolerance and performance.