Build and maintain the out-of-band control plane for server fleets, focusing on BMC interactions and DPU provisioning. Develop observable Rust services and collaborate with hardware vendors to debug firmware and production behavior.
Vultr
50 Remote Job Openings at Vultr
The role involves owning baseline security standards for Linux, containers, and Kubernetes while bridging the gap between security policy and engineering practice. You will act as a consultant to engineering teams to assess security posture and deliver actionable remediation plans.
Design and maintain automated golden image pipelines and configuration management to enforce security baselines across Linux workloads. Integrate security scanning and policy-as-code controls into CI/CD pipelines to prevent non-compliant workloads from reaching production.
Design, build, and maintain OS images for Cloud Compute and Bare Metal platforms while migrating legacy scripts to cloud-init. Develop automation tooling for image creation and manage the full lifecycle of marketplace applications.
Build and maintain production-grade automation and core services for Vultr's global cloud platform. Focus on hardening infrastructure, improving CI/CD pipelines, and enhancing the reliability of cloud-native services.
The role involves automating production infrastructure using Ansible and Terraform while maintaining cloud-native services. The engineer will develop tooling in Go, Python, and Bash to support platform operations and participate in on-call rotations.
Design, develop, and maintain high-performance PHP applications and RESTful APIs for a global cloud platform. Optimize MySQL database schemas and implement asynchronous workflows to handle millions of daily requests.
Design, develop, and scale high-performance backend PHP applications and MySQL database schemas for a global cloud platform. Collaborate with engineering teams to implement RESTful APIs and asynchronous workflows while mentoring junior engineers.
Lead customers in designing, implementing, and optimizing AI/ML workloads on Vultr's GPU platform. Collaborate with internal teams to translate customer needs into product features and create reference architectures.
Lead the FedRAMP authorization process and develop the product roadmap for Vultr's public sector and sovereign cloud offerings. Bridge the gap between federal security controls and engineering execution to ensure compliance with strict regulatory baselines.
The IAM Engineer will own and expand Okta integrations and implement RBAC standardizations across the organization. They will also design automation strategies using Okta Workflows and BetterCloud to reduce manual provisioning overhead.
Support the health and stability of global network infrastructure by troubleshooting performance issues and onboarding BGP customers. Research network events to identify technical debt and drive end-to-end improvements across platforms.
The GTM Engineer will build and optimize the systems powering go-to-market motions, focusing on automated workflows and audience architecture. They will own end-to-end campaign execution and create the copywriting frameworks for AI-generated outbound messaging.
Design and execute performance benchmarks for AI training and inference workloads to identify bottlenecks. Systematically tune GPU parameters and develop automation tools to maximize throughput and validate new hardware platforms.
Validate, troubleshoot, and optimize high-speed networking fabrics for GPU clusters supporting AI training and inference. This includes tuning performance parameters for distributed workloads and managing fabric health using NVIDIA UFM.
Design and implement the canonical data model for the control plane ERP transformation, focusing on product catalogs and financial transactions. Build and maintain internal APIs and event streams while refactoring legacy code into clean domain boundaries.
Senior Software Engineer, ERP Platforms (Integration & Reporting)
Vultr
·
Full Time
·
23 days ago
Vultr
Build and maintain reporting infrastructure and ETL pipelines to ensure reliable financial data flow from the control plane to the data warehouse. Design API surfaces and event contracts to enable downstream systems to consume data reliably for revenue and margin analysis.
Lead a cross-functional squad to evolve the internal control plane into a structured, ERP-grade platform. Own the technical roadmap for billing, invoicing, and revenue recognition while partnering with Finance and Data teams.
Design, implement, and optimize advanced optical transport networks using DWDM, ROADM, and OTN technologies. Provide Tier III technical support and lead network maintenance, upgrades, and operational excellence initiatives.
The role involves administering and optimizing SIEM and EDR platforms while building detection content and automated SOAR workflows. Additionally, the engineer will conduct threat hunts and support incident response activities to improve security visibility.
Architect and scale a global network telemetry and data intelligence platform to optimize infrastructure and security. Lead DDoS detection frameworks, threat mitigation, and telemetry-driven billing ecosystems while managing a high-performing engineering team.
Design and build a next-generation bare metal deployment pipeline to achieve sub-60-second provisioning times. Develop automation toolkits for firmware management and integrate VPC networking for bare metal instances.
Drive revenue growth for AI infrastructure by managing strategic customer relationships and guiding them through the full sales cycle. Collaborate cross-functionally to align product capabilities with customer technical requirements and business goals.
Oversee multiple technical support teams to resolve complex platform-level issues and act as an escalation point for high-priority incidents. Coordinate with engineering and system administration teams to improve platform reliability and maintain internal knowledge base documentation.
The BI Architect will design, build, and maintain internal datasets and automated reporting dashboards using Power BI and BigQuery. They will manage the semantic layer, optimize data models, and collaborate cross-functionally to provide business insights for executive leadership.
Define and deliver AI infrastructure capabilities for large-scale GPU workloads by translating customer requirements into technical specifications. Partner with engineering and operations teams to build scalable offerings for training, inference, and cluster orchestration.
Own the end-to-end roadmap for the Observability Platform, focusing on telemetry ingestion, visualization, and alerting for large-scale GPU clusters. Partner with engineering teams to ensure all new infrastructure is observable by design and translate low-level signals into actionable health views.
Modernize the customer-facing control panel by migrating legacy templates to React and Tailwind CSS. Collaborate with cross-functional teams to build accessible and intuitive UIs for cloud infrastructure products.
Lead the Platform, Lifecycle & Troubleshooting team to resolve complex incidents and manage server repurposing and migrations. Drive operational maturity by developing automation, runbooks, and mentoring senior engineers to improve uptime and performance.
Design and build the observability pipeline for global datacenter infrastructure, including bare metal servers and provisioning workflows. Establish standards for telemetry collection and create actionable dashboards and alerting for operational stakeholder teams.
Lead and develop a distributed team of engineers focusing on identity management, billing, and internal administration platforms. Drive engineering execution and modernization of a PHP monolith while aligning technical strategy with product roadmaps.
Serve as a senior security authority supporting complex technical sales engagements and strategic customer interactions. Articulate the security architecture of the Vultr platform and lead security due diligence during RFP responses.
Lead the engineering and operations team responsible for Vultr's global cloud storage portfolio across block, object, and file services. Oversee the design, deployment, and optimization of storage platforms to ensure high availability and performance for AI and HPC workloads.
Own the GPU Orchestration product line, managing the roadmap for Kubernetes, Slurm, and Run:ai integrations for AI and HPC workloads. Drive the end-to-end cluster lifecycle and establish resource management capabilities for GPU workloads.
Lead the identification and evaluation of new data center sites to support global infrastructure expansion. Coordinate with utilities and government agencies to secure power, land, and permits for AI compute deployments.
Lead strategic sourcing and supplier management for critical cloud infrastructure, including GPUs, CPUs, and data center services. Manage the end-to-end procurement cycle, negotiate complex vendor agreements, and develop scalable procurement policies.
The Procurement Analyst manages vendor data, analyzes spend, and drives process efficiency to support infrastructure scaling. Key duties include maintaining procurement databases, supporting RFP/RFQ processes, and collaborating with Finance and technical teams.
The Senior Talent Acquisition Specialist will strategically source top talent for various corporate roles and manage the full-cycle recruitment process. This includes conducting compensation analysis, partnering with hiring managers to build teams, and maintaining strong candidate relationships.
Lead the planning and execution of GPU data center deployment projects while coordinating across engineering, facilities, and supply chain teams. Manage the deployment lifecycle of high-density compute environments, including infrastructure design and risk mitigation.
The Technical Account Manager leads the post-sales technical success for customers deploying large-scale AI and GPU workloads on the Vultr platform. This role involves advising on cluster architecture, optimizing performance, and managing long-term technical strategy for high-growth AI accounts.
The Staff AI/ML Infrastructure Engineer will design, maintain, and optimize high-performance GPU and bare metal infrastructure. They will lead technical direction, mentor engineers, and collaborate with hardware vendors to ensure reliable, scalable AI/ML platform performance.
The Technical Support Specialist will respond to customer inquiries and platform alerts, troubleshoot technical issues, and coordinate resolutions across the organization. This role is collaborative and customer-facing, focusing on maintaining platform reliability and meeting service-level objectives.
Sr. Technical Program Manager, Data Center & Network Delivery
Vultr
·
Full Time
·
4 months ago
Vultr
The Senior Technical Program Manager will oversee the end-to-end delivery of complex data center builds and facilitate critical conversations among stakeholders. This role requires maintaining program visibility and ensuring alignment across multiple workstreams.
The specialist will be responsible for driving and improving the Return Merchandise Authorization (RMA) process for emerging technologies, including GPU/CPU fault validation and interaction with hardware vendors. Key duties involve scheduling on-site support, validating system software/firmware, and creating trouble reports in external vendor portals.
The administrator will design, configure, and implement Linux system upgrades and enhancements to ensure scalability and reliability for global customer infrastructure. This role involves administering servers, serving as a technical escalation point for complex issues, and collaborating with engineering teams on data center deployments.
The specialist will partner with Hiring Managers to define roles and develop recruitment strategies, posting openings across various channels. They will own the full-cycle recruiting process, conducting structured interviews and assessing candidates for technical skills and cultural fit.
The CPU-focused Technical Account Manager owns the post-sales technical relationship for a portfolio of customers, driving successful onboarding, managing technical health, and ensuring long-term value realization from CPU-based IaaS offerings. This involves advising on architecture, optimizing performance and spend, navigating roadmaps, and resolving complex issues in partnership with internal teams.
The Account Executive will be responsible for generating Cloud Compute revenue, focusing on customers within target accounts, and partnering with the Sales Development team to assess customer needs. This role requires ownership of results, forecasting, leveraging sales tech stack tools, and collaborating with engineering and product teams.
The Strategic Customer Success Manager will serve as the primary post-sales contact for top-tier strategic accounts, driving customer adoption and retention. They will build executive relationships and coordinate service delivery to ensure contractual obligations are met.
The Director of Service Delivery will lead the end-to-end delivery of customer infrastructure deployments and ensure projects are delivered on time and within scope. This role involves managing teams and processes to meet customer expectations and service level agreements.