Source Job

US

  • Implement and manage the infrastructure stack to enable the engineering team to ship quickly and effectively.
  • Proactively identify and eliminate bottlenecks in the devops process to ensure optimal developer velocity.
  • Maintain Tempo chain reliability, validator infrastructure, and explorer reliability.

Rust Kubernetes Terraform Prometheus Linux

20 jobs similar to Platform Engineer

Jobs ranked by similarity.

Global

  • Maintain mission-critical bare-metal infrastructure and ensure SLI/SLO/SLA compliance.
  • Improve and create automation for deploying web2/web3 infrastructure using Ansible and high-level languages.
  • Conduct R&D for new blockchain projects to optimize node performance and participate in on-call rotation.

P2P.org is the largest institutional staking provider with a TVL of over $10B and a market share exceeding 20% in restaking. The company unites talented individuals globally, sharing a passion for decentralized finance and a culture of ownership and continuous learning.

US Unlimited PTO

  • Own the infrastructure layer for AI workloads including inference serving, Kubernetes, and agent-sandboxing platforms.
  • Manage the serving tier for open-weight models, Kubernetes operators, and stateful data planes.
  • Oversee the sandbox runtime, control-plane services, and observability tooling.

AZX accelerates positive impact in critical industries through AI transformation, specializing in physics-informed ML and enterprise AI solutions for climate and sustainability. Founded in 2024, the company is a profitable public benefit corporation with a growing team working with category leaders in real estate, energy, logistics, and utilities.

Europe

  • Manage and troubleshoot complex distributed large-scale software systems
  • Build scalable, secure and reliable container-based infrastructure
  • Automate software delivery processes with CI/CD pipelines

Coinspaid Dev is the engineering brand behind the technology, infrastructure, and R&D expertise built within Coinspaid, focusing on advancing blockchain infrastructure engineering. With over 120 engineers and more than 11 years of industry experience, they bring together teams building distributed systems and blockchain infrastructure across 20+ blockchain networks.

$250,000–$285,000/yr
US Unlimited PTO

  • Define architecture and best practices for the platform and infrastructure layer the product is built on.
  • Own the deploy pipeline and lead the move to a GitOps model (Argo) for fast, safe releases.
  • Design and harden multi-tenant isolation and blast-radius protection for top-tier customers, including dedicated deployments.

We are the Engineering Operations Platform - mission control for the AI software factory, providing visibility, governance, and golden paths. We are a group of 80 passionate individuals, backed by $60M Series C from Sequoia, IVP, and others, with a fully remote culture.

Global Unlimited PTO

  • Operate the Monad node fleet, including health, sync, upgrades, and incident response for validators, full nodes, and archive nodes.
  • Own infrastructure-as-code with Ansible, Terraform, and Kubernetes, and build observability with Prometheus, Grafana, and Loki.
  • Design and build AI agent tooling for automated operations, including runbooks-as-code and deterministic guardrails.

Category Labs designs and builds decentralized technology, including the Monad blockchain, a high-performance EVM-compatible Layer 1. The team raised $225M in series A funding and is a lean, collaborative group of engineers and researchers with a culture of low ego and high-quality output.

US

  • You'll contribute to infrastructure scaling to infinitely many apps, improving performance and reliability across backend services.
  • You'll support observability efforts, help implement SLOs, and build foundational services for next-generation cloud infrastructure.
  • You'll participate in triage and on-call processes to diagnose issues and implement changes to prevent recurrence.

Bubble is an AI visual development platform that empowers anyone to create software without code, from first-time entrepreneurs to enterprise teams. With over 6 million users in more than 100 countries and a mission to break down barriers to entrepreneurship, the company fosters a collaborative and inclusive culture focused on empowering builders worldwide.

US

  • Build and maintain SRE microservices under the guidance of senior engineers.
  • Deliver infrastructure and application updates using GitOps and automated CI/CD pipelines.
  • Participate in a mentored on-call rotation and write clear runbooks and incident post-mortems.

Bitdeer is a world-leading technology company providing AI and Bitcoin mining infrastructure solutions. Headquartered in Singapore, the company operates data centers across multiple countries and focuses on building computational infrastructure.

US

  • Design, build, and maintain robust, scalable, and secure infrastructure systems supporting Laurel's AI-driven platform.
  • Manage and optimize cloud infrastructure (AWS and Azure), Kubernetes orchestration, and CI/CD pipelines to increase deployment frequency and reliability.
  • Implement comprehensive observability, monitoring, and alerting to maintain system health and partner with engineering teams to optimize performance and cost-efficiency.

Laurel is an AI Time platform for professional services firms, automating work time capture and connecting time data to business outcomes for clients like EY and Crowell & Moring. The company comprises top AI, product, and engineering talent, is VC-backed by Google Ventures and IVP, and fosters an inclusive, ambitious culture.

US

  • Own the infrastructure end-to-end for ScaleOps' self-hosted and SaaS platforms.
  • Manage cloud infrastructure across AWS, GCP, and Azure, including networking, security, and compute.
  • Collaborate with customers and internal teams to ensure rapid feature delivery without compromising reliability.

ScaleOps is redefining autonomous cloud and AI infrastructure, freeing DevOps from manual resource management. Backed by $210M+ in funding, they are trusted by leading enterprises and Fortune 100 companies, with a fast-paced, innovative culture.

Poland

  • Design and develop highly performant backend services for real-time data processing and web APIs.
  • Define and own reliability objectives and error budgets for core API services.
  • Build observability through metrics, logging, tracing, and dashboards, and participate in on-call rotation.

The company provides technology that protects businesses and users from online fraud. It is a fully remote, globally distributed organization.

Global

  • Own and scale cloud infrastructure including compute, networking, storage, and data systems.
  • Lead BYOC and private cloud deployments with infrastructure-as-code and GitOps foundations.
  • Establish reliability through service-level objectives, observability, and incident response processes.

A technology company builds a developer-focused platform with scalable cloud infrastructure. This is a remote-first opportunity with a small, autonomous engineering team operating in North America, LATAM, and Europe, offering high autonomy and ownership.

US

  • Design and maintain CI/CD and MLOps pipelines for software and machine learning models.
  • Build and scale cloud-native infrastructure using Kubernetes, Docker, and GPU clusters.
  • Champion Infrastructure as Code and observability to ensure high availability and governance.

Bitdeer is a world-leading technology company for AI and Bitcoin mining infrastructure, providing comprehensive solutions and cloud capabilities. Headquartered in Singapore, the company has deployed data centers across multiple countries and fosters a culture of innovation.

$140,000–$220,000/yr
North America LATAM Europe

  • Own and scale the cloud infrastructure behind our open-source platform: compute, networking, and the data layer.
  • Lead BYOC: turn customer-cloud deployments into a real product, with provisioning, upgrades, and observability that scale past bespoke work per deal.
  • Make reliability a product feature: meaningful SLOs, and an incident process people trust.

Nango is a developer infrastructure company that provides API access for agents and apps, enabling AI applications to connect to the real world through integrations. With over 400 paying customers and a team of 14 from top tech companies like AWS, GitHub, and Okta, they are a YC-backed, multi-million ARR company that values ownership and autonomy.

$82,550–$101,600/yr
UK

  • Build and maintain cloud infrastructure and tooling using Kubernetes, Terraform, and AWS/GCP to support Monzo's microservices.
  • Develop backend services and APIs that abstract infrastructure complexity, enabling self-service for engineering teams.
  • Work across platform and software engineering, tackling high-scale traffic management, infrastructure migrations, and security tooling.

Monzo is a digital bank on a mission to make money work for everyone, offering personal and business bank accounts, savings, investments, and pensions. We’re a growing company with 10 million customers, focused on solving problems and changing lives through innovative financial products and a diverse, inclusive culture.

Global

  • Be on an on-call rotation responding to production incidents and support service engineers.
  • Run infrastructure with Ansible, Puppet, Terraform, and Kubernetes, making monitoring alert on symptoms.
  • Design and maintain core infrastructure scaling to hundreds of thousands of concurrent users.

Our client's Cloud Operations team is expanding its SRE function, keeping user-facing services and production systems running smoothly. The team specializes in systems like networking, Linux kernel, and distributed systems, blending pragmatic operations with software engineering.

UK Unlimited PTO 18w maternity 12w paternity

  • Own the technical strategy for multi-ecosystem scaling, defining architecture for onboarding new language ecosystems.
  • Drive end-to-end remediation automation, leading redesign of CVE workflows to close the loop from detection to verified release.
  • Set platform-wide technical direction spanning package index, build pipelines, and orchestration tooling to serve customers and ecosystem teams.

Chainguard is the trusted source for open source, delivering hardened, secure, and production-ready builds of open source software. They serve Fortune 500 enterprises and global industry leaders, and are venture-backed by leading investors, fostering a culture of customer obsession and intentional action.

Global

  • Build and operate Go backend services for key management and signing at scale.
  • Design and implement cryptographic recovery and trust models for non-custodial wallets.
  • Work with external auditors, Kubernetes, and distributed systems in a security-critical environment.

Offchain Labs builds the Arbitrum stack, the leading Ethereum scaling solution. They are a well-funded company with $124 million in backing, a small but high-ownership team, and a culture of security-first engineering and collaborative review.

US

  • Drive complex infrastructure migrations and build platform tooling and automation across multiple production environments.
  • Support development teams by consulting on infrastructure needs and improving observability and incident response.
  • Provide operational support and maintain platform reliability through structured debugging and on-call rotations.

PENN Entertainment is North America's leading provider of integrated entertainment, sports content, and casino gaming experiences. We operate across numerous locations in North America and foster a culture that cares about career growth and skill expansion.

Global Unlimited PTO

  • Design, build, and maintain scalable distributed backend systems powering platform and product capabilities.
  • Own backend services end-to-end, from architecture to monitoring, while collaborating with product and engineering peers.
  • Build and evolve Web3 platform backend components, including protocol integrations and blockchain network support.

Zerion builds backend systems for Web3 infrastructure, powering API and wallet products trusted by 50+ top crypto brands. They have a fully remote team with strong ownership and high code quality, processing over 3.1B requests and supporting 1.7M+ active funded wallets.

$120,000–$155,000/yr
Global

  • Own infrastructure as code across development, staging, and production environments
  • Build, maintain, and improve CI/CD pipelines for reliable and efficient deployments
  • Manage cloud infrastructure, establish scalable engineering practices, and lead incident response

CelebriOS is a software company building B2B SaaS products that help businesses make better decisions and streamline operations. The company has a remote-first working environment and a benefits package designed to support their team.