Source Job

$165,000–$230,000/yr
Global

  • Design and build agent runtime infrastructure with Firecracker, Rust, and Go
  • Define and enforce security boundaries for running untrusted AI agents
  • Architect global scale distributed systems for scheduling and orchestration

Go Rust Distributed Systems Security Kubernetes

20 jobs similar to Software Engineer - Platform (Senior/Staff)

Jobs ranked by similarity.

US

  • Design and develop foundational components and frameworks for our Agentic AI platform.
  • Collaborate with cross-functional teams to create platform solutions that empower developers.
  • Provide production support and ensure platform stability, working closely with ML engineers.

Legion builds secure, reliable AI systems that integrate with complex platforms, optimizing workflows and enhancing human capability. They work with partners like Palantir, Nvidia, HPE, and Oracle, and are looking for bold thinkers to shape the future of grounded AI.

India

  • Design and build production AI agent systems that diagnose, investigate, and remediate infrastructure issues across one of the world’s largest GPU fleets.
  • Build the distributed services, orchestration framework, knowledge graph, and retrieval systems that power infrastructure agents.
  • Own services end to end, including architecture, implementation, testing, deployment, observability, and production operations.

Together AI is a research-driven artificial intelligence company that builds open and transparent AI systems. The company has contributed to leading open-source research like FlashAttention and RedPajama, and aims to lower the cost of modern AI through co-designed software, hardware, algorithms, and models.

India

  • Design and build production AI agent systems for diagnosing and remediating infrastructure issues in large-scale GPU environments.
  • Develop distributed services, orchestration frameworks, knowledge graphs, and retrieval systems to power AI agents.
  • Own services end-to-end from architecture through production, collaborating with infrastructure and engineering teams.

The company builds AI agents that operate and automate large-scale GPU infrastructure. The engineering team is highly collaborative and remote, fostering ownership and autonomy.

US

  • Design, build, and maintain highly reliable backend services and distributed systems for RapidFort's security platform.
  • Build systems for processing and analyzing large volumes of security, vulnerability, container, and runtime data.
  • Solve complex problems involving concurrency, performance, scalability, and distributed processing across Linux, Kubernetes, and cloud environments.

RapidFort is a cybersecurity company focused on securing and optimizing modern software supply chains and cloud-native environments. The company works with enterprise and U.S. public-sector customers in security-sensitive environments, fostering a culture of technical ownership and collaboration.

US Unlimited PTO

  • Own the infrastructure layer for AI workloads including inference serving, Kubernetes, and agent-sandboxing platforms.
  • Manage the serving tier for open-weight models, Kubernetes operators, and stateful data planes.
  • Oversee the sandbox runtime, control-plane services, and observability tooling.

AZX accelerates positive impact in critical industries through AI transformation, specializing in physics-informed ML and enterprise AI solutions for climate and sustainability. Founded in 2024, the company is a profitable public benefit corporation with a growing team working with category leaders in real estate, energy, logistics, and utilities.

Canada Unlimited PTO

  • Design and operate secure, multi-tenant infrastructure for agentic AI workloads at the Linux and network boundary.
  • Build process isolation and sandboxing mechanisms using namespaces, cgroups, gVisor, Firecracker, and workload identity frameworks.
  • Collaborate with Platform, Security, and Backend teams to own infrastructure from architecture through production operations.

Jobgether is an AI-powered job matching platform that connects candidates with hiring companies. It has a remote-first culture and uses technology to review applications fairly.

Global

  • Evolving Supabase Edge Runtime, an open-source Rust-based host that runs Deno isolate and enforces per-request memory and CPU limits.
  • Implementing monitoring, alerting, and OpenTelemetry tracing to drive latency and reliability improvements.
  • Expanding functions for more use cases like AI inference, MCP servers, and improving developer experience with Supabase CLI.

Supabase is the Postgres development platform, providing a complete backend solution including Database, Auth, Storage, Edge Functions, Realtime, and Vector Search. With a globally distributed team of ~400 members across 60+ countries, they are open-source-first and move fast, building in public.

$140,000–$225,000/yr
US Canada Unlimited PTO

  • Build and maintain backend services for our LLM gateway, including routing, rate limiting, and observability.
  • Contribute to sandboxing and isolation infrastructure for safe agent-generated code execution.
  • Write high-performance backend code in Go, Rust, or async Python, supporting Kubernetes-based platform services.

AZX accelerates positive impact in critical industries through AI transformation. Founded in 2024 and profitable from the start, we work with category leaders in real estate, energy, logistics, and utilities.

$150,000–$300,000/yr
US

  • You own the platform runtime, voice infrastructure, enterprise security, and evaluation layer for all clients.
  • Work directly with the founder on architecture, from design through production and incident response.
  • Build self-serve infrastructure, lead incident response, and keep the platform fast and reliable on Kubernetes.

Rifa AI builds an AI agents platform for contact centers in regulated industries, helping enterprises deploy trustworthy voicebots. They are a small, passionate engineering team with paying enterprise clients and growing revenue, backed by Seaborne Capital.

APAC

  • Design and deliver significant components and core subsystems of our Kubernetes platform, such as secrets management, workload identity, storage, or cluster networking, from design through production operation.
  • Contribute to the architecture of distributed workloads, working with dependent teams to get runtime and isolation models right, while spending most time hands-on in code.
  • Own operability of built systems including SLOs, failure modes, upgrades, migrations, and on-call, and mentor earlier-career engineers.

ServiceNow is the AI control tower for business reinvention, bringing together any AI, any data, and any workflow to help 85% of the Fortune 500 work smarter, faster, and better. The company fosters an AI-native culture where technology and talent are unstoppable together, with a focus on freeing people from busywork.

UK Unlimited PTO 18w maternity 12w paternity

  • Own the technical strategy for multi-ecosystem scaling, defining architecture for onboarding new language ecosystems.
  • Drive end-to-end remediation automation, leading redesign of CVE workflows to close the loop from detection to verified release.
  • Set platform-wide technical direction spanning package index, build pipelines, and orchestration tooling to serve customers and ecosystem teams.

Chainguard is the trusted source for open source, delivering hardened, secure, and production-ready builds of open source software. They serve Fortune 500 enterprises and global industry leaders, and are venture-backed by leading investors, fostering a culture of customer obsession and intentional action.

$175,000–$250,000/yr
US Unlimited PTO 2w maternity 2w paternity

  • Design, build, and maintain internal software and tooling used across Tiberius.
  • Develop services, automation, and operational tools primarily in Go and Rust.
  • Build, operate, and continuously improve infrastructure running on Kubernetes.

Tiberius Aerospace builds next-generation weapons systems for the United States, United Kingdom, and their allies through faster iteration and tighter software-hardware integration. They foster a high-trust, mission-driven culture with significant autonomy for engineers.

$180,000–$220,000/yr
US Unlimited PTO

  • Design, build, and maintain agent infrastructure and platforms, including the TRACE Graph, embeddings, and semantic search.
  • Collaborate with detection engineers and threat hunters to encode domain expertise into agents.
  • Build automated evaluations and benchmarks for non-deterministic agentic systems.

Nebulock is an agentic threat hunting platform that autonomously surfaces behaviors, not just IOCs, from various data sources. As a startup, we emphasize collaboration, low ego, and a relentless focus on delivering customer value.

Global

  • Build and operate Go backend services for key management and signing at scale.
  • Design and implement cryptographic recovery and trust models for non-custodial wallets.
  • Work with external auditors, Kubernetes, and distributed systems in a security-critical environment.

Offchain Labs builds the Arbitrum stack, the leading Ethereum scaling solution. They are a well-funded company with $124 million in backing, a small but high-ownership team, and a culture of security-first engineering and collaborative review.

Canada

  • Build and operate backend platform services for APIs, Kafka-based event routing, workflow execution, and secrets management.
  • Design event-driven workflows and durable executions using technologies such as Kafka and Temporal.
  • Own production reliability, observability, and security across multi-tenant, multi-cloud environments.

This partner company builds core platform services for workflow execution, identity, event routing, secrets management, and audit infrastructure. It is a high-growth, engineering-led organization focused on innovation, accountability, and continuous improvement.

LATAM

  • Design and implement high-performance components in Rust, C++, or Go for the DEX engine and protocol runtime.
  • Build low-latency pipelines for order execution, event propagation, and state updates, optimizing concurrency and memory layout.
  • Investigate and resolve performance bottlenecks using profiling and benchmarking, driving system design decisions for senior candidates.

Nexus is the engine for verifiable finance, building a blockchain designed to embed the global financial system into one compounding system. Headquartered in San Francisco with a growing presence in Buenos Aires, the cross-continental team is backed by leading investors and works with more than 100 partners.

Global

  • Design and improve the core orchestration engine for managing Supabase Branches lifecycle.
  • Provision ephemeral sandboxed execution environments on EKS/ECS for untrusted build workloads.
  • Implement job scheduling, pipeline optimization, and observability for thousands of concurrent builds.

Supabase is the Postgres development platform, built by developers for developers. They are a globally distributed team of ~400 members across 60+ countries, operating fully remote with an open-source-first culture.

US

  • Build and run monitoring, tracing, and alerting infrastructure to ensure platform reliability and security.
  • Lead incident response and recovery, including root cause analysis, and improve deployment processes for fast, safe code changes.
  • Collaborate with engineering teams to deliver a stable, scalable platform and handle load for resource-intensive applications.

WellSaid Labs is the leading AI voiceover studio for enterprise and professional use, providing ultra-realistic voices that the world’s biggest brands trust. We are a fully distributed team across the U.S. with a focus on responsible AI and an inclusive culture.

$200,000–$350,000/yr
US

  • Design, build, and operate production software systems.
  • Own significant technical projects from architecture through deployment.
  • Develop scalable services, APIs, infrastructure, and internal tooling.

The company builds sophisticated infrastructure and software systems. It is a rapidly scaling technology company that values ownership and technical excellence.

Latin America

  • Build and expand our healthcare data platform and member management solution.
  • Write high quality code through various channels and in collaboration with the Product team.
  • Have a high degree of autonomy and influence on the culture and direction of the technical organization.

Ubiminds connects Latin American tech professionals with software companies in the US and Canada. With 9 years in the market and GPTW-certified, they support hundreds of professionals in data, design, product, and engineering.