Source Job

Europe

  • Lead end-to-end technical engagements: Partner directly with engineering teams to diagnose, unblock, and resolve complex infrastructure challenges.
  • Execute critical migrations: Develop reference implementations, tooling, and guidance to transition teams off deprecated systems seamlessly.
  • Accelerate platform adoption: Act as primary technical contact for new teams onboarding to Planet's core infrastructure.

Kubernetes Google Cloud Platform Terraform Go Python

20 jobs similar to Senior Infrastructure Solutions Engineer

Jobs ranked by similarity.

Europe 6w PTO

  • Design, build, and operate infrastructure for real-time systems handling millions of concurrent connections and billions of monthly API requests.
  • Drive Kubernetes end to end: cluster architecture, workload design, and migration of existing services from AWS to GCP.
  • Own cloud cost and efficiency optimization, measuring impact against real spend and utilization data.

Stream powers real-time chat, video, activity feeds, and AI moderation for billions of end-users across thousands of apps. We are a Series B company with around 145 employees from over 35 countries, offering a fast-paced startup culture with real ownership.

Romania

  • Design and advance core infrastructure for multi-cloud Kubernetes clusters and developer toolchains.
  • Automate operations and engineering tasks to improve productivity and reliability.
  • Build machine learning infrastructure to enable AI teams to train and deploy large-scale models.

Cresta provides an AI platform that transforms customer conversations into competitive advantages by combining conversational AI, real-time agent augmentation, and conversation intelligence. The company has raised over $270 million from top investors like a16z, Greylock, and Sequoia, and is led by AI industry veterans.

US

  • Drive complex infrastructure migrations and build platform tooling and automation across multiple production environments.
  • Support development teams by consulting on infrastructure needs and improving observability and incident response.
  • Provide operational support and maintain platform reliability through structured debugging and on-call rotations.

PENN Entertainment is North America's leading provider of integrated entertainment, sports content, and casino gaming experiences. We operate across numerous locations in North America and foster a culture that cares about career growth and skill expansion.

$126,000–$174,000/yr
US

  • Define DevOps strategy and lead infrastructure architecture across multi-environment, multi-region cloud systems.
  • Architect and own scalable Kubernetes platforms, infrastructure as code, and DevSecOps implementation.
  • Drive platform reliability, performance SLAs, cost optimization, and lead complex migrations and AI/ML platform infrastructure.

Robots & Pencils is an applied AI engineering firm that designs and ships AI co-workers for enterprise operations. Founded in 2009, the company has delivery centers across Canada, the US, Eastern Europe, and Latin America, with teams averaging over 15 years of experience.

France

  • Evolve an Internal Developer Platform enabling development teams to deploy and operate applications securely with high self-service.
  • Play a crucial role in infrastructure architecture in a multi-cloud environment with a focus on GCP.
  • Build and maintain the platform used by over 50 internal clients, then support their concrete use by teams.

Lifen believes medical data can transform healthcare by reducing administrative burden, improving care coordination, and accelerating scientific discovery. Since 2015, the company has connected 800 hospitals and 150,000 healthcare professionals, with over 150 employees working remotely and from offices to unlock the potential of health data.

UK

  • Lead Cloud Platform and SRE teams to scale securely and efficiently.
  • Drive infrastructure strategy, including Kubernetes (GKE) clusters and developer platform.
  • Champion SRE culture with SLOs, error budgets, and observability.

Prolific builds human data infrastructure for AI development. The company is a fast-growing, mission-driven organization with cross-functional teams and a strong ownership culture.

$59,400–$65,880/yr
Europe

  • Lead the design, implementation, and ongoing improvement of reliable, scalable, and secure production platforms and services.
  • Work closely with cross-functional teams to build and maintain resilient infrastructure and deployment patterns.
  • Provide technical leadership and mentorship, promoting strong engineering standards and operational best practices.

Cision is a global leader in PR, marketing and social media management technology and intelligence, helping brands connect with customers and stakeholders. They have offices in 24 countries, a network of over 1.1 billion influencers, and a culture that champions diversity, equity, and inclusion.

UK Ireland Estonia Netherlands Sweden Israel Eastern Europe Portugal Unlimited PTO

  • Develop and evolve foundational software and services enabling product and development teams.\n- Architect, design, and implement Infrastructure as Code using Terraform.\n- Deploy, manage, and optimize Kubernetes clusters on GCP (GKE) and AWS (EKS).

DoiT is a global technology company that helps cloud-driven organizations leverage cloud for business growth and innovation through data, technology, and human expertise. They work with over 4,000 customers worldwide and foster a remote-first, entrepreneurial culture.

$220,000–$292,000/yr
US Unlimited PTO

  • Own the platform including GCP, Kubernetes, Temporal, GPU fleet, and deploy/rollback machinery.
  • Contribute to AI enablement substrate: GPU capacity, training/inference pipelines, and cost optimization.
  • Strengthen team practices through tooling, standards, tests, observability, and release processes.

Descript is building a simple, intuitive, fully-powered editing tool for video and audio — an editing tool built for the age of AI. They are a team of 150 backed by top investors like OpenAI and Andreessen Horowitz, with a culture that values collaboration and serendipitous discovery.

Romania

  • Architect, deploy, and manage highly available, fault-tolerant cloud infrastructure across Google Cloud Platform (GCP) and Google Kubernetes Engine (GKE).
  • Maintain and scale declarative infrastructure using Terraform across a multi-hundred-file estate, enforcing GitOps workflows with Atlantis.
  • Build, maintain, and optimize robust automated pipelines for continuous integration and delivery using GitHub Actions, Jenkins, and ArgoCD.

Point Wild helps customers monitor, manage, and protect against the risks associated with their identities and personal information in a digital world. Backed by WndrCo, Warburg Pincus and General Catalyst, Point Wild is a scrappy, nimble organization dedicated to creating the world’s most comprehensive portfolio of industry-leading cybersecurity solutions.

US

  • Own core platform infrastructure including Terraform migration, CI/CD pipelines, and environment provisioning as part of a growing team.
  • Build and maintain tooling and abstractions that let product engineers ship and run code without solving infrastructure problems from scratch.
  • Advance observability foundation, standardize monitoring and alerting, and support GCP infrastructure scaling.

Astra builds mission-critical infrastructure for moving money at scale, processing billions in annual transaction volume with 99.9%+ uptime. We are a remote-first company hiring within the U.S., with a small team focused on thoughtful collaboration and clarity.

$12,500–$20,800/mo
Turkey

  • Own the technical architecture and evolution of core infrastructure.
  • Engineer for scale and performance through capacity modeling and bottleneck diagnosis.
  • Participate in on-call rotation and drive technical recovery during incidents.

Sezzle is a fintech company that revolutionizes shopping through interest-free installment plans, blending cutting-edge technology with financial empowerment. It has a dynamic and innovative team culture focused on shaping the future of fintech and retail.

Global

  • Own and scale cloud infrastructure including compute, networking, storage, and data systems.
  • Lead BYOC and private cloud deployments with infrastructure-as-code and GitOps foundations.
  • Establish reliability through service-level objectives, observability, and incident response processes.

A technology company builds a developer-focused platform with scalable cloud infrastructure. This is a remote-first opportunity with a small, autonomous engineering team operating in North America, LATAM, and Europe, offering high autonomy and ownership.

Brazil 4w PTO

  • Own critical infrastructure across compute, networking, CI/CD, Kubernetes, and observability.
  • Manage Kubernetes environments and infrastructure-as-code with Terraform, improving developer experience and reducing operational friction.
  • Lead production incident response, influence architecture, and integrate AI-powered tools to boost engineering efficiency.

Jobgether is an AI-powered recruitment platform that connects candidates with global hiring companies. This role is with a partner company, a globally distributed technology organization offering a collaborative, informal culture and long-term opportunities.

$150,000–$165,000/yr
US Unlimited PTO

  • Design and manage high-availability platforms using Kubernetes, Terraform, and Ansible with native-AI capabilities.
  • Develop and operate the observability stack: Grafana, Mimir, Loki, Tempo, and Prometheus on Kubernetes via GitLab CI/CD.
  • Build automation scripts in Python, maintain GitOps pipelines, and mentor mid-level engineers.

Flexential builds and operates critical IT platforms including observability, DevOps, and ITSM technologies. The company fosters a collaborative engineering culture and values diversity.

Canada USA Unlimited PTO

  • Own the observability, logging and alerting for Kubernetes clusters and critical workloads.
  • Build and maintain automation for lifecycle management of Kubernetes clusters.
  • Identify and root-fix reliability bottlenecks before they become incidents.

Wrapbook is an AI platform for production finance, built for feature films and TV, trusted by Netflix and Paramount. Backed by top investors, our team of over 350 employees uses AI to transform how finance teams work.

Germany 6w PTO

  • Architect and automate platforms to enable fast, secure code delivery with CI/CD integration.
  • Own the GKE/Kubernetes infrastructure on Google Cloud, ensuring scalability and reliability.
  • Proactively engineer security, eliminating bottlenecks and driving developer experience.

Smartclip is an ad tech company that builds a platform for developers to deploy code quickly and securely. They foster a culture of ownership, fast iterations, and minimal bureaucracy, with a remote-first approach and on-site meetings in Berlin.

India

  • Lead the architecture and implementation of complex cloud solutions across AWS and GCP.
  • Drive cloud automation and optimization initiatives to improve scalability and reliability.
  • Provide technical leadership and mentorship to engineers while collaborating with global teams.

The company focuses on cloud infrastructure and platform engineering. They operate with global teams and emphasize automation, security, and reliability.

  • Own the reliability, performance, and scalability of Runlayer's infrastructure across AWS and GCP.
  • Manage Kubernetes clusters, database reliability, and CI/CD pipelines for rapid deployments.
  • Lead incident response and partner with product engineers to design resilient systems for enterprise customers.

Runlayer builds a unified platform for MCPs, Skills, and AI Agents, providing enterprises with security, governance, and observability to deploy AI safely and at scale. Founded by engineers who built AI Actions for OpenAI and Zapier Agents, the team has raised $42M from Felicis and Khosla Ventures, serving companies like Gusto, Instacart, and Opendoor.

$180,000–$220,000/yr
US

  • Design, implement, and maintain reliable, scalable, and secure infrastructure to support applications and automation systems.
  • Automate infrastructure provisioning, configuration management, and deployment pipelines using tools like Terraform and ArgoCD.
  • Implement observability solutions and enforce security best practices to ensure uptime and system performance.

Bright Machines is a next-generation, AI-enabled manufacturer focused on data center infrastructure production, using proprietary AI-based robotics and software to assemble hardware products for hyperscalers and OEMs. The company is headquartered in San Francisco, California, with an integration center in Guadalajara, Mexico, and has been recognized by Forbes' AI 50 and other leading organizations.