Source Job

$185,000–$260,000/yr
US Unlimited PTO

  • Lead architecture and implementation of Greyspace, a simulated internet for SimSpace cyber ranges.
  • Build distributed services and realistic network behavior including DNS, TLS, web, and mail.
  • Drive technical standards, mentor engineers, and partner cross-functionally to deliver production-grade systems.

Distributed Systems Networking Kubernetes API Design CI/CD

20 jobs similar to Staff Software Engineer, Greyspace

Jobs ranked by similarity.

$163,815–$227,130/yr
Canada 20w maternity 16w paternity

  • Design and document networking features and connectivity solutions.
  • Enhance product capabilities focusing on Funnel and DERP relay infrastructures.
  • Investigate and resolve complex escalated network difficulties.

Tailscale is making safe connection effortless by delivering software that securely interconnects people and their devices, no matter where they are. Founded in 2019 and fully distributed, they are backed by Accel, CRV, Insight, Heavybit, and Uncork Capital.

India

  • Own architecture health of large-scale distributed systems, including failure modes, capacity constraints, and consistency guarantees.
  • Identify and remediate systemic risks such as single points of failure, unbounded queues, and data-loss scenarios.
  • Work hands-on with Node.js/Go and GCP technologies to prototype solutions and resolve complex failures.

This company operates large-scale distributed systems processing billions of events and messages. Its engineering culture values technical rigor, proactive problem solving, and clear cross-team communication.

Global

  • Own technical design and execution for major products, systems, and engineering initiatives.
  • Design scalable, reliable systems across backend, frontend, APIs, databases, and infrastructure.
  • Lead technically critical initiatives and mentor engineers across teams while staying hands-on.

BJAK builds technology products used by millions of users across Southeast Asia. It is a fast-moving, distributed international team working across multiple countries and time zones.

$3,000–$5,000/mo
Costa Rica

  • Architect bespoke AI-driven simulations and core services for realistic cyber range training.
  • Lead projects and mentor engineers on best practices for distributed, event-driven systems.
  • Partner across teams to deliver fault-tolerant, containerized solutions in air-gapped environments.

SimSpace provides an AI-powered cyber training and simulation platform for governments, militaries, and enterprises to test and validate defenses. Trusted by experts from U.S. Cyber Command, it fosters a human-centered culture of continuous learning and collaboration.

$145,500–$235,400/yr
US

  • Design, implement, test, and operate production services and APIs.
  • Lead projects or meaningful components of projects from problem definition through deployment and iteration.
  • Improve the observability and operability of the systems you own, including metrics, logs, traces, alerting, and incident learnings.

LaunchDarkly provides a platform that helps engineering teams release software and AI with speed, safety, and control using feature flags and observability. The company is growing and emphasizes teamwork, humility, openness, curiosity, and inclusive collaboration.

Mexico

  • Architect and evolve core control and context planes, including service registries, SLO enforcement, and automated canary releases.
  • Own the service chassis and golden path, maintaining multi-language Java/Python libraries, Helm charts, and deployment pipelines.
  • Drive reliability engineering practices, mentor engineers, and lead architectural strategy for distributed systems at production scale.

The company builds large-scale web data products and distributed engineering infrastructure for AI-driven workflows. It operates a remote-first, globally distributed engineering culture focused on reliability, autonomy, and technical excellence.

Europe

  • Lead and mentor a high-performing team of senior engineers building distributed systems for decentralized compute and storage.
  • Drive technical direction, stay hands-on with architecture, design, and implementation of P2P networking, compute orchestration, and core platform services.
  • Champion reliability, observability, and engineering practices including automated testing, CI/CD, incident response, and production readiness.

Pragmatike builds decentralized compute and storage platforms powered by distributed systems. They foster a remote-first, async writing-first engineering culture with a focus on pragmatic decision-making and continuous improvement.

$170,000–$235,000/yr
US

  • Design and implement backend services for licensing, entitlements, feature access, and usage limits across NodeZero's product and APIs.
  • Build and evolve provisioning, admin experience, MSP/MSSP capabilities, and audit logging for a multi-tenant SaaS platform.
  • Operate production services with monitoring, incident response, and a high bar for design quality and test coverage.

Horizon3 is a fast-growing, remote cybersecurity company that helps organizations proactively find, fix, and verify exploitable attack vectors through its NodeZero autonomous pentesting platform. The team is a fusion of former special operations cyber operators and startup engineers, fostering a culture of respect, collaboration, ownership, and results.

India

  • Own the architecture health of a billion-scale distributed system, including failure modes, capacity limits, and cross-deployment interactions.
  • Approve critical-path designs and hunt gaps like single points of failure, unbounded queues, and missing idempotency proactively.
  • Build the parts nobody else can, prototype risky architectural bets, and ship remediations after serious incidents.

HighLevel is an AI-powered business operating system that gives agencies and SMBs the infrastructure to build, automate and scale. With over 2,000 team members across 10+ countries, HighLevel operates as a global, remote-first organization built for speed and ownership.

$170,350–$275,550/yr
US 16w maternity 16w paternity

  • Design and build core systems for Docker's enterprise agentic platform using Go and Kubernetes.
  • Own production reliability from on-call to SLOs and incident response.
  • Set technical direction and mentor engineers while driving security, identity, and network boundaries.

Docker builds developer tools trusted by over 20 million monthly users and billions of container pulls. It is a globally distributed, remote-first team with a culture of trust, flexibility, and engineering excellence.

Global

  • Design and build a global edge router platform to dynamically steer customer traffic.
  • Implement advanced observability, optimize routing performance, and collaborate cross-functionally.
  • Participate in on-call rotations and incident response for edge-related issues.

Supabase is the Postgres development platform, built by developers for developers, providing a complete backend solution including Database, Auth, Storage, Edge Functions, Realtime, and Vector Search. With ~400 team members across 60+ countries and over $1B raised, we are a born-remote, open-source-first company with a strong culture of collaboration and innovation.

$192,000–$192,000/yr
Global

  • Lead large-scale Sensor Platform initiatives in collaboration with product, infrastructure, and engineering peers.
  • Own root cause analysis and postmortems for complex production issues, ensuring durable fixes.
  • Establish standards for documentation, test harnesses, and observability across the codebase.

Dragos is the global leader in OT cybersecurity, combining technology, threat intelligence, and expert services to protect critical infrastructure. The team is remote-first, mission-driven, and built on authenticity, transparency, and trust.

$265,000–$285,000/yr
US

  • Design and develop back-end systems for cloud-based cybersecurity solutions.
  • Create and implement distributed systems architecture for endpoint management platforms.
  • Collaborate with cross-functional teams and participate in code reviews.

CrowdStrike is a global leader in cybersecurity, protecting organizations from breaches with its advanced AI-native platform. Since 2011, the company has grown to process trillions of events daily and fosters a culture of innovation, flexibility, and responsible AI adoption.

Global Unlimited PTO

  • Architect and build robust, scalable, and highly available distributed infrastructure.
  • Build a cutting-edge cloud-native platform on public cloud and automate resource management.
  • Improve reliability, security, and cost efficiency of cloud services.

ClickHouse builds a real-time analytics database platform and manages ClickHouse Cloud data plane end-to-end with compute, networking, and security. As a rapidly scaling global startup, the company operates across 25+ countries and fosters a flexible, remote-friendly culture with equity and healthcare benefits.

US

  • Lead and develop a globally distributed Core Infrastructure team to evolve compute, networking, storage, and cloud systems.
  • Own technical strategy and improve reliability, capacity management, and cloud efficiency across the platform.
  • Build automation, self-service infrastructure, and leverage AI tools to accelerate engineering execution and reduce toil.

Samsara is the pioneer of the Connected Operations Cloud, helping organizations harness IoT data to improve safety, efficiency, and sustainability. As a public company processing over 25 trillion data points annually, it fosters a growth-minded, customer-focused culture with a globally distributed team.

Ireland

  • Design and deliver major components of Twilio's carrier test and observability platform.
  • Build high-throughput data and alerting systems that turn telemetry into actionable signals.
  • Extend and harden production LLM systems for automated carrier troubleshooting with safe guardrails.

Twilio is a customer engagement platform that delivers innovative communications solutions to hundreds of thousands of businesses and empowers millions of developers worldwide. The company is remote-first, values connection and global inclusion, and has a vibrant culture where diverse employees make a global impact.

$217,000–$303,000/yr
United States

  • Partner with cross-functional teams to design and deliver scalable backend systems for major product initiatives.
  • Own the full software lifecycle from technical design to rollout, using A/B experiments and data analysis to drive decisions.
  • Build and maintain high-performance APIs and distributed services using modern languages and tools.

Reddit is a community of communities, built on shared interests, passion, and trust. It is home to the most open and authentic conversations on the internet, with 100,000+ active communities and approximately 130 million daily active unique visitors.

$219,000–$352,000/yr
Global

  • Define and drive the technical vision for scalable, secure, and high-performance platform architecture.
  • Provide technical leadership and mentorship, setting best practices and standards across the engineering organization.
  • Lead research and prototyping of emerging technologies, and ensure reliability through automation and security best practices.

Docker is a developer tooling platform trusted by 20M+ users, with products like Docker Desktop, Docker Hub, and Docker Scout. It is a globally distributed, remote-first team building AI-ready infrastructure for software delivery.

$77,000–$106,700/yr
France

  • Design and develop scalable search and indexing systems for an AI search engine.
  • Ensure operational excellence by participating in on-call rotation and maintaining system quality.
  • Collaborate with a global remote team to solve distributed system challenges.

Algolia is a pioneer and market leader in AI Search, empowering over 18,000 businesses to deliver blazing-fast search experiences. With $150 million in Series D funding and a valuation of $2.25 billion, the company fosters a high-trust, flexible culture and values diversity and collaboration.

$140,000–$215,000/yr
US

  • Lead and scale a global network engineering team focused on IP and optical backbone infrastructure.
  • Drive cloud connectivity strategy, automation, and operational excellence across hybrid clouds.
  • Collaborate with vendors and internal partners to evolve network architecture for reliability and scale.

CrowdStrike is a global cybersecurity leader that stops breaches with an AI-native platform processing nearly 3 trillion events daily. The company fosters a mission-driven culture of flexibility, autonomy, and responsible AI adoption, with employees known as CrowdStrikers.