Own architecture health of large-scale distributed systems, including failure modes, capacity constraints, and consistency guarantees.
Identify and remediate systemic risks such as single points of failure, unbounded queues, and data-loss scenarios.
Work hands-on with Node.js/Go and GCP technologies to prototype solutions and resolve complex failures.
This company operates large-scale distributed systems processing billions of events and messages. Its engineering culture values technical rigor, proactive problem solving, and clear cross-team communication.
Build and operate core platform components like queues, storage layers, caching, schedulers, and rate limiters.
Own architecture, implementation, monitoring, and production reliability for high-throughput distributed systems.
Work with Node.js/TypeScript, Go, GCP, Kubernetes, Redis, MongoDB, and other large-scale data technologies.
The hiring company builds and operates platform infrastructure for large-scale CRM, automation, and communication systems, processing billions of events and messages monthly. Its engineering culture emphasizes ownership, collaboration, and high standards for reliability and production quality.
Design and build platform components like queueing pipelines, storage layers, caching, and rate limiters for systems handling billions of events.
Own components end-to-end, from architecture and design to production health, on Node.js microservices and GCP infrastructure.
Debug cross-layer incidents, write clear design docs, and drive fixes for scalability issues like cache stampedes and hot shards.
HighLevel is an AI-powered business operating system for agencies, entrepreneurs, and SMBs to build, automate, and scale. With over 2,000 team members across 10+ countries, it operates as a global, remote-first organization focused on speed, ownership, and community-driven growth.
Design and optimize data models, queries, and indexing across MongoDB, Firestore, and ElasticSearch for high-scale systems.
Own ElasticSearch reliability, including ingestion, indexing, shard strategy, and query performance for billions of documents.
Define data flow and integration boundaries between storage, cache, and APIs with clear contracts and fault isolation.
HighLevel is an AI-powered business operating system that provides agencies, entrepreneurs, and SMBs with infrastructure to build, automate, and scale. With over 2,000 team members across 10+ countries, HighLevel operates as a global, remote-first organization built for speed and ownership.
Design and develop scalable search and indexing systems for an AI search engine.
Ensure operational excellence by participating in on-call rotation and maintaining system quality.
Collaborate with a global remote team to solve distributed system challenges.
Algolia is a pioneer and market leader in AI Search, empowering over 18,000 businesses to deliver blazing-fast search experiences. With $150 million in Series D funding and a valuation of $2.25 billion, the company fosters a high-trust, flexible culture and values diversity and collaboration.
Architect and ship scalable, mission-critical brokerage systems and APIs.
Provide technical leadership, mentorship, and engineering standards across teams.
Drive operational excellence through monitoring, alerting, and incident response.
Alpaca provides agent-first brokerage infrastructure for stocks, ETFs, options, crypto, fixed income, and 24/5 trading. The company has 400+ globally distributed members, is backed by $400M, and values curiosity, empathy, and accountability.
Own and drive impactful distributed systems problems end-to-end, from inception through production launch.
Collaborate with a strong engineering team to prioritize and solve the most important problems for the company.
Raise the quality bar while keeping systems reliable and operationally lean, and mentor fellow engineers.
StarTree is a cloud-based software company that enables businesses to derive advanced insights from real-time and historical data using Apache Pinot. The company was founded by the core engineering team behind Apache Pinot, has secured Series B funding, and was named one of The Information's 50 Most Promising Startups.
Design and build distributed data systems handling large-scale ingestion and processing.
Drive architectural decisions and take end-to-end ownership of critical components.
Collaborate with product teams to translate ambiguous requirements into robust technical solutions.
Our partner builds a large-scale, multi-chain data platform that ingests, models, and delivers blockchain data to users and developers. They are a remote-first, distributed team with a strong engineering culture focused on ownership and collaboration.
Lead architectural evolution by driving cross-team initiatives toward simplicity and scalability.
Coach and mentor engineers through pairing, design reviews, and hands-on coaching in XP practices.
Champion sustainable quality by promoting TDD, pair programming, and continuous delivery.
Prerender helps marketers create dynamic sites that are easily indexable by search engines. More than 100,000 websites globally trust Prerender to improve organic search ranking.
Lead and grow a high-performing engineering team focused on the Payments Platform, including hiring, mentoring, and performance management.
Drive end-to-end execution of platform initiatives, ensuring predictable delivery across multiple concurrent workstreams.
Champion technical excellence through architecture evolution, reliability practices, and cross-functional collaboration.
HighLevel is an AI-powered business operating system that provides infrastructure for agencies, entrepreneurs, and SMBs to build, automate, and scale their operations. With over 2,000 team members across 10+ countries, HighLevel is a global, remote-first organization that emphasizes speed, ownership, and people-first culture.
Own large slices of the system end to end, from approach to operation.
Turn Beads into a platform and take Gas City to the cloud.
Define SLOs, observability, backups, and security baseline for enterprise readiness.
Gas City builds the open-source stack teams use to run coding agents at scale, including the Beads work graph and Gas City agent orchestration. It's a small, flat organization moving toward revenue with a focus on reliability and agent-driven development.
Lead the development of Reddit's Ingestion Platform, designing and delivering reliable software for distributed data movement across streaming and batch workloads.
Own the architecture of the platform's control and data planes, including pipeline APIs, connectors, and sink integrations, expanding beyond Kafka-to-BigQuery to S3/GCS and Apache Iceberg.
Mentor engineers, drive migrations from legacy systems, and establish robust reliability, security, and operational practices for pipelines running on Kubernetes.
Reddit is a community of communities built on shared interests, passion, and trust, hosting authentic conversations across 100,000+ active communities and approximately 130 million daily active unique visitors. The company fosters an open, collaborative culture with a focus on reliability, performance, and efficiency.
Partner with cross-functional teams to design and deliver scalable backend systems for major product initiatives.
Own the full software lifecycle from technical design to rollout, using A/B experiments and data analysis to drive decisions.
Build and maintain high-performance APIs and distributed services using modern languages and tools.
Reddit is a community of communities, built on shared interests, passion, and trust. It is home to the most open and authentic conversations on the internet, with 100,000+ active communities and approximately 130 million daily active unique visitors.
Design and build highly scalable distributed systems to support global expansion across multiple countries and regulatory environments.
Partner closely with product, operations, payments, logistics, and platform teams to own critical operational and transactional systems.
Work on cross-border transaction infrastructure, tax and regulatory compliance, international post-purchase experiences, and marketplace operational tooling.
Whatnot is the largest live shopping platform in North America and Europe, enabling sellers to build businesses across hundreds of categories. They are a remote co-located team with hubs across the US, UK, Ireland, Poland, Germany, and Australia, and were recently named the #1 Best Startup Employer in America.
Design and build agent runtime infrastructure with Firecracker, Rust, and Go
Define and enforce security boundaries for running untrusted AI agents
Architect global scale distributed systems for scheduling and orchestration
We build agent sandboxes—runtime infrastructure that is fast, durable, and secure by default for AI systems. We are a small, globally distributed team based in San Francisco backed by forward-thinking investors.
Define and evolve the long-term database and data infrastructure strategy across the company.
Lead complex, cross-team initiatives such as database migrations, sharding, and multiregion architectures.
Mentor engineers and raise the technical bar for database and distributed-systems engineering.
HighLevel is an AI-powered business operating system for SMBs, providing infrastructure to build, automate, and scale operations. With over 2,000 team members across 10+ countries, the company operates as a global, remote-first organization focused on speed and ownership.
Work on the heart of broker-dealer operations: clearing and settlement for US equities and options, ensuring systems are reliable and scalable.
Design and build new systems to enable new product offerings and revenue streams, from ideation to deployment.
Participate in an on-call rotation to maintain system health while prioritizing elimination of issues to protect team free time.
Alpaca is a US-headquartered global leader in agent-first brokerage infrastructure for stocks, ETFs, options, crypto, and more, serving hundreds of financial institutions across 40 countries. The team is a dynamic group of 400+ globally distributed members who thrive working from their favorite places around the world, with a culture valuing curiosity, empathy, and accountability.
Design, build, and operate GitLab Orbit backend services, primarily in Rust, within a distributed, cloud-native environment.
Improve deployment, monitoring, and operations using Kubernetes, Helm, Terraform, and cloud services from AWS or GCP.
Automate operational work, strengthen observability, and manage production issues to reduce single points of failure.
GitLab is the intelligent orchestration platform for DevSecOps, helping organizations increase developer productivity, improve operational efficiency, and accelerate digital transformation. Trusted by more than 50 million registered users and over 50% of the Fortune 100, GitLab fosters a high-performance culture driven by shared values and continuous knowledge exchange.
Architect, build, and scale secure software systems including microservices and REST/gRPC APIs.
Lead design of distributed systems using Messaging, Search stores, and cloud storage solutions.
Mentor junior engineers, monitor system health, and optimize services with a DevOps mindset.
Zscaler accelerates digital transformation so customers can be more agile, efficient, resilient, and secure. The Zscaler Zero Trust Exchange platform protects thousands of customers from cyberattacks and data loss, with thousands of employees and a culture focused on ownership, collaboration, and challenge.