Source Job

US Unlimited PTO

  • Act as operational backbone, ensuring stability, security, and high availability of legacy production VMs running Java/Tomcat and MySQL 5.7.
  • Manage incident response, performance monitoring, and root-cause analysis using modern observability platforms.
  • Collaborate on cloud architecture design and automation to modernize infrastructure and support new products.

Linux MySQL Google Cloud Platform Terraform

20 jobs similar to Senior Infrastructure Engineer (Legacy Operations & Cloud Modernization)

Jobs ranked by similarity.

$145,000–$145,000/yr
US

  • Design, deploy, and maintain GCP infrastructure, including Compute Engine, GKE, Cloud Storage, IAM, and Cloud Interconnect, following well-architected principles.
  • Automate provisioning using Terraform, implement IAM best practices, and partner on security audits and hardening.
  • Monitor performance, availability, and cost, and support cross-cloud and on-premises networking as needed.

Dragos is the global leader in xOT cybersecurity, protecting critical infrastructure systems that deliver water, power, and healthcare. The remote-first team spans North America, Europe, the Middle East, and APAC, built on authenticity, transparency, and trust.

US Unlimited PTO

  • Own and drive key infrastructure modernization initiatives toward container-orchestrated infrastructure.
  • Design and maintain infrastructure as code across multiple cloud providers.
  • Provide technical leadership and mentorship across the Systems Engineering team.

Intellum is the leader in corporate education technology, powering large learning programs for brands like Google, Meta, and Amazon. We are a remote-first company with a culture that values curiosity, creativity, perseverance, and kindness, and we invest in our people through personal development budgets and annual retreats.

$125,000–$135,000/yr
US

  • Own and evolve monitoring and alerting for MySQL and PostgreSQL infrastructure across multiple datacenters.
  • Establish a quarterly backup verification and disaster recovery testing program for all database systems.
  • Serve as first-tier on-call responder for database incidents and maintain replication health.

Vultr provides high-performance cloud infrastructure for enterprises and AI innovators globally. It is the world's largest privately-held cloud infrastructure company, with 33 data centers, hundreds of thousands of customers, and a $3.5 billion valuation.

US

  • Build and operate monitoring, tracing, alerting, and observability infrastructure for system reliability.
  • Drive platform security initiatives with preventative controls and resilient architecture.
  • Lead incident response and recovery, including root-cause analysis and preventative measures.

This role is with a partner company managing AI-powered products. They are a growing technology organization with a fully distributed US-based team and a collaborative culture focused on large-scale infrastructure and AI technology.

Germany 6w PTO

  • Architect and automate platforms to enable fast, secure code delivery with CI/CD integration.
  • Own the GKE/Kubernetes infrastructure on Google Cloud, ensuring scalability and reliability.
  • Proactively engineer security, eliminating bottlenecks and driving developer experience.

Smartclip is an ad tech company that builds a platform for developers to deploy code quickly and securely. They foster a culture of ownership, fast iterations, and minimal bureaucracy, with a remote-first approach and on-site meetings in Berlin.

Global

  • Support management and maintenance of Google Cloud Platform infrastructure.
  • Assist in ensuring reliability, scalability, and performance of cloud services.
  • Contribute to GitLab CI/CD pipelines, work with Kubernetes, and write automation scripts.

Miratech is a global IT services and consulting company that helps visionaries change the world through digital transformation. With nearly 1,000 professionals across 25 countries, the company maintains a culture of Relentless Performance with a 99% project success rate and over 30% year-over-year growth.

$220,000–$300,000/yr
US 4w PTO

  • Improve infrastructure performance, reliability, and cost optimization.
  • Automate operational workflows and maintain CI/CD pipelines for security and compliance.
  • Collaborate with global engineering teams to standardize environments and deployment patterns.

Kalepa builds AI that performs professional labor in insurance underwriting, claims, pricing, and operations. Backed by IA Ventures and Inspired Capital, the team brings experience from Facebook, Palantir, Google, and other top tech companies.

US

  • Own, troubleshoot, and solve customer technical issues using best practices.
  • Identify cases that require escalation to product or engineering.
  • Design and implement automation solutions to scale support offerings.

Eon transforms cloud backups into useful assets with their Cloud Backup Posture Management platform. They are an ambitious, collaborative startup backed by prominent investors.

UK Ireland Estonia Netherlands Sweden Israel Eastern Europe Portugal Unlimited PTO

  • Develop and evolve foundational software and services enabling product and development teams.\n- Architect, design, and implement Infrastructure as Code using Terraform.\n- Deploy, manage, and optimize Kubernetes clusters on GCP (GKE) and AWS (EKS).

DoiT is a global technology company that helps cloud-driven organizations leverage cloud for business growth and innovation through data, technology, and human expertise. They work with over 4,000 customers worldwide and foster a remote-first, entrepreneurial culture.

UK

  • Architect and build a robust, scalable, and highly available distributed infrastructure.
  • Build a cutting-edge cloud-native platform on top of the public cloud and automate cloud resource management.
  • Work closely with core database development and security teams to produce the SaaS offering.

ClickHouse is a real-time analytics and data warehousing company recognized on the Forbes Cloud 100 list. With over 4,000 customers and rapid growth, the company is a leader in its space.

  • Architect and automate scalable cloud environments across AWS and Azure using Terraform, Ansible, Helm, and CDK.
  • Serve as Linux subject matter expert, managing system builds, core services, and performance from kernel up.
  • Lead CI/CD pipelines, observability, security, and incident response to ensure platform reliability.

Fueled is a leading digital strategy, design, and engineering agency. The 300+ person team has designed and built hundreds of digital products for major brands like Google, Apple, and The New York Times, and thrives in a culture that values flexibility, creativity, and cutting-edge technology.

$230,000–$270,000/yr
US 16w maternity 16w paternity

  • Lead the design, development, and maintenance of highly scalable infrastructure systems.
  • Drive the technical vision and roadmap for infrastructure teams.
  • Mentor and guide engineers, fostering a culture of continuous learning and improvement.

Maven Clinic is the world's largest virtual clinic for women and families, providing end-to-end women's and family health programs. Founded in 2014 with over $425 million in funding, the company has been recognized as a top workplace and leader in innovation, fostering an award-winning culture focused on making healthcare work for all.

$150,000–$185,000/yr
US Unlimited PTO

  • You will lead the reliability and operational evolution of our platform, building and improving system resiliency and establishing SLIs and SLOs.
  • You will partner with product engineering teams to own and operate their services, evolving observability platforms and strengthening incident practices.
  • You will contribute to day-to-day cloud infrastructure work alongside reliability specialty, including on-call rotation.

Rocket Money is a financial technology company that empowers people to live their best financial lives by providing insights and services to save time and money. The company runs hundreds of services in production, processing billions of transactions, and has a culture of reliability and innovation.

Mexico

  • Own and evolve enterprise Linux platforms across on-premises and cloud environments, driving platform strategy, architecture, automation, security, and reliability.
  • Lead complex initiatives end to end, acting as a senior escalation point during incidents and driving root-cause analysis and permanent corrective actions.
  • Embed security-by-design principles, partner with security teams on hardening and compliance, and mentor engineers to advance platform engineering maturity.

The company is a technology organization focused on enterprise IT infrastructure and platform engineering. It values ownership, strategic thinking, and mentorship, with a collaborative environment spanning multiple technical teams.

France

  • Evolve an Internal Developer Platform enabling development teams to deploy and operate applications securely with high self-service.
  • Play a crucial role in infrastructure architecture in a multi-cloud environment with a focus on GCP.
  • Build and maintain the platform used by over 50 internal clients, then support their concrete use by teams.

Lifen believes medical data can transform healthcare by reducing administrative burden, improving care coordination, and accelerating scientific discovery. Since 2015, the company has connected 800 hospitals and 150,000 healthcare professionals, with over 150 employees working remotely and from offices to unlock the potential of health data.

Global

  • Own and scale cloud infrastructure including compute, networking, storage, and data systems.
  • Lead BYOC and private cloud deployments with infrastructure-as-code and GitOps foundations.
  • Establish reliability through service-level objectives, observability, and incident response processes.

A technology company builds a developer-focused platform with scalable cloud infrastructure. This is a remote-first opportunity with a small, autonomous engineering team operating in North America, LATAM, and Europe, offering high autonomy and ownership.

UK

  • Lead the design and development of shared backend services and platform capabilities, driving architectural decisions across multiple teams.
  • Build scalable, highly available cloud-native solutions using Java and AWS technologies, and mentor engineers to promote engineering excellence.
  • Collaborate with product, architecture, and engineering teams to solve complex technical challenges, championing best practices in quality, performance, and security.

Turnitin is a recognized innovator in the global education space, partnering with institutions to promote honesty and fairness in assessments for over 25 years. With a remote-first culture and team members in over 35 countries, we empower you to work with purpose and accountability in a diverse, collaborative community.

$5,000–$6,000/mo
Latin America

  • Architect resilient cloud-native infrastructure for Java backend services.
  • Lead client discussions, translating requirements into actionable roadmaps.
  • Standardize containerization with Docker and Kubernetes, and build CI/CD pipelines.

GoFasti is a Talent-as-a-Service company that connects world-class developers and designers from Latin America with global companies. We are a growing team focused on making remote work remarkable for both talent and clients.

Canada

  • Design, develop, test, and maintain software that powers product provisioning, licensing, and purchasing capabilities.
  • Build scalable, reliable, and maintainable backend services using modern engineering practices.
  • Collaborate closely with engineers, product managers, and cross-functional teams to deliver impactful features.

Sonatype provides end-to-end software supply chain security, combining proactive protection against malicious open source, enterprise SBOM management, and leading dependency management. Trusted by over 2,000 organizations including 70% of the Fortune 100 and 15 million developers, the company values diversity, inclusivity, and flexible work.

Brazil

  • Drive the performance, stability, security, and reliability of production environments with a focus on automation and proactive improvements.
  • Design and maintain infrastructure using Infrastructure as Code tools like Terraform, and manage Kubernetes and cloud environments.
  • Lead vulnerability management, incident response, and secure CI/CD practices to ensure resilience and operational excellence.

Jobgether is a platform that uses AI-powered matching to connect candidates with hiring companies. It processes applications and shares shortlists with employers, offering a remote-first and inclusive work environment.