Please mention DailyRemote when applying
See how much of this job your resume covers, and what’s missing.
Want a recruiter to go through it line by line?
Get professional reviewQuestions interviewers often ask for this role, with sample answers.
Upload your resume and we draft a letter for this exact role, tailored to what it asks for.
You will own the design, operation, and scaling of AWS infrastructure and production software to ensure platform reliability and performance. This involves managing cloud costs, incident response, and collaborating with content engineering to support hands-on security learning environments.
SOFTWARE ENGINEERING
Keep a platform serving 8 million+ learners fast, reliable and cost-efficient and help decide how it should be built.
Join TryHackMe as a Senior Cloud/Software Engineer. This isn't a traditional DevOps role. It's for a strong software engineer who has gone deep into cloud infrastructure, reliability and distributed systems. You'll work across AWS and Vercel on infrastructure shaped by unusually demanding cyber security learning environments.
Location/Time zone: this is a remote role, but strong overlap with UK business hours is required.
TryHackMe operates across AWS and Vercel, supporting live virtual machines, hands-on security exercises and a high-traffic web product. You'll work in a small, focused Cloud/SRE squad with real ownership over how the platform evolves.
We're looking for a senior software engineer with strong hands-on AWS and infrastructure experience who can own problems across application code, cloud infrastructure, reliability, security and cost. You should be as comfortable changing production TypeScript/Node.js as you are debugging a distributed system or evolving its infrastructure.
This is not a traditional DevOps role focused on pipelines or infrastructure in isolation. We need someone who owns the whole problem, including existing systems, and can make pragmatic decisions about how they should evolve.
You'll be working on things like:
Designing and operating software, tooling and AWS infrastructure across networking, compute, containers, databases, IAM, storage and security
Writing and maintaining production Node.js/TypeScript alongside infrastructure-as-code and automation
Cloud cost optimisation and finding more efficient ways to run the platform without compromising reliability
Reliability, observability, incident response, disaster recovery and failure testing
Operating and scaling production databases, including performance, capacity and availability concerns
Vercel and our evolving application hosting model
Developer tooling and automation that reduce operational overhead
Working with Content Engineering on the virtual-machine infrastructure powering our hands-on learning experience
Supporting security and assurance work, including technical controls and audit findings
Day to day, you'll own technical problems from discovery through production, debug across software and infrastructure, lead code and design reviews, mentor other engineers and help raise the technical bar across the team.
Strong software engineering fundamentals and a track record of shipping and owning production software
Strong hands-on AWS experience, beyond basic service familiarity or certification-level knowledge
Solid infrastructure depth across networking, compute, containers, serverless, IAM, storage, monitoring and security
Strong production debugging skills and a good understanding of distributed systems
Experience operating and scaling production databases
Strong infrastructure-as-code and automation experience
Good judgement around reliability, performance, security, cost and long-term maintainability
Confidence with AI development tools such as Cursor, Codex or Claude, paired with the judgement to challenge their output
Clear communication and the ability to work effectively in a small, high-ownership team
Able to work with strong overlap with UK business hours
Bonus: NoSQL/MongoDB experience, Vercel, Next.js infrastructure, Kubernetes/ECS, observability, high-scale SaaS, disaster recovery or chaos engineering.
Our infrastructure problems aren't generic. They're genuinely unusual, and you'll have real influence over how we solve them, not just a rulebook to follow.
π£ 100% Remote - work from anywhere you want!
π» Tools - a dedicated work laptop + any accessories you need to do your best work.
π Swag Pack - start your TryHackMe journey with a branded swag bundle!
πͺ Personal Development - Β£2,500 training budget to acquire certifications, and more.
β±οΈ Company Retreat - an annual company retreat, fully paid for us!
π Lunch on us - covered during our recurring company virtual lunches.
π§‘ Health Insurance - if you're in a country that doesn't have public health care.
πΌ Enhanced Maternity & Paternity - an enhanced package on top of statutory requirements.
πΈ 401k / Pension - TryHackMe makes it easy to save money for your retirement.
Stage 1: Introductory Chat with Talent Partner (~15 mins)
Stage 2: Technical & Experience Interview with the Head of Engineering
Stage 3: Take-Home AI-Assisted Technical Assessment
Stage 4: Founders Interview (Culture & Operating Alignment) with Co-Founder
Stop the endless job search. Our AI finds and applies to the best jobs for you.
Featuring 215,178+ Jobs in Software Engineer
Answer easy questions
215,178+ jobs across 15+ categories
Get your best job matches
Only hand-screened, legit jobs
Find a remote job faster
No ads, scams, or junk
“I was the first applicant for a remote marketing position that got listed on the company website the same day I applied. Had an interview within 48 hours!”