Please mention DailyRemote when applying
Match your resume skills with our AI powered skill match!
Questions interviewers often ask for this role, with sample answers.
Upload your resume and we draft a letter for this exact role, tailored to what it asks for.
You will own the production reliability, observability, and performance of real-time trading systems, including trading engines and execution gateways. Responsibilities include debugging production incidents, hardening market data ingestion, and building data-integrity tooling to ensure accurate PnL and reconciliation.
About the company
Hi, we're Ondo Finance. Our mission is to provide institutional-grade, blockchain-enabled investment products and services. We have both a technology arm that develops decentralized finance technology, and an asset management arm that creates and manages tokenized funds. We are the global leader in tokenized treasuries, tokenized stocks and ETFs, and are building the future of institutional-grade financial services onchain.
Founded by folks from Goldman Sachs Digital Assets Team, we’re backed by some of the best investors in the world including Founders Fund, Coinbase Ventures, Pantera Capital, Tiger Global, and more. We are currently the leaders in the space in terms of AUM and are well capitalized to continue growing the firm. We're fully remote, with team members across the U.S.
About the role
Ondo operates real-time trading systems that run around the clock across traditional and crypto venues. The platform spans low-latency Rust engines, a fleet of Go services for trading, execution, and PnL accounting, and a multi-region Kubernetes footprint on AWS.
We are looking for an SRE with strong systems programming skills to own the reliability, observability, and performance of this platform. This is a hands-on role: you will read and modify Go and Rust code, debug latency regressions down to the feed handler, run incident response during market hours, and build the automation that keeps a 24/7 trading system healthy with a small team.
Target outcomes
Responsibilities
Requirements
Nice to haves
Tech stack
Go, Rust, Python | Kubernetes (EKS), Flux, SOPS | AWS (multi-region), S3 parquet lake | Prometheus, Grafana, Datadog | Postgres, CockroachDB, BigQuery | Databento, venue WebSocket/REST feeds
What we offer
Stop the endless job search. Our AI finds and applies to the best jobs for you.
Featuring 212,512+ Jobs in Site Reliability Engineer
Answer easy questions
212,512+ jobs across 15+ categories
Get your best job matches
Only hand-screened, legit jobs
Find a remote job faster
No ads, scams, or junk
“I was the first applicant for a remote marketing position that got listed on the company website the same day I applied. Had an interview within 48 hours!”