Design and implement real-time services with high throughput and low latency, working across the full software development lifecycle.
Collaborate with stakeholders to understand customer needs and deliver simple, robust, and scalable solutions.
Embrace challenges of scaling a complex distributed platform with global points of presence, ensuring high availability and reliability.
Twilio is shaping the future of communications by delivering innovative solutions to hundreds of thousands of businesses and empowering millions of developers worldwide. The company is a remote-first organization with a strong culture of connection and global inclusion.
Build and operate agents across the testing lifecycle — test design, generation, maintenance, execution and failure triage.
Own the infrastructure that makes agent-driven testing reliable: environments, test data, tool interfaces, guardrails, runtime and cost control.
Define how we measure agent output quality, so teams can trust results instead of hand-reviewing them.
Camunda is the enterprise platform for agentic orchestration, enabling coordination of AI agents, people, and systems across business processes. Trusted by over 700 organizations, we are a fast-growing, fully remote company with a culture of ownership and innovation.
Own and drive impactful distributed systems problems end-to-end, from inception through production launch.
Collaborate with a strong engineering team to prioritize and solve the most important problems for the company.
Raise the quality bar while keeping systems reliable and operationally lean, and mentor fellow engineers.
StarTree is a cloud-based software company that enables businesses to derive advanced insights from real-time and historical data using Apache Pinot. The company was founded by the core engineering team behind Apache Pinot, has secured Series B funding, and was named one of The Information's 50 Most Promising Startups.
Own the reliability posture of production services, including availability, latency, capacity, and performance.
Define and operate against SLIs and SLOs, using error budgets to drive engineering priorities.
Lead incident response, write post-mortems, and build automation to measurably improve service reliability.
Twilio is a cloud communications platform that delivers innovative solutions to hundreds of thousands of businesses and empowers millions of developers worldwide to create personalized customer experiences. The company is remote-first with a strong culture of connection, global inclusion, and a focus on solving problems and taking initiative.
Design, build, and operate core billing platform services that run at massive scale on AWS.
Leverage technologies like Apache Kafka, REST APIs, and Kubernetes to create high-quality, fully performing software.
Troubleshoot operational issues, improve procedures, and execute the full software development life cycle.
Twilio delivers innovative communication solutions to hundreds of thousands of businesses, empowering millions of developers worldwide. They are a remote-first company with a strong culture of connection and global inclusion.
Lead the development of Reddit's Ingestion Platform, designing and delivering reliable software for distributed data movement across streaming and batch workloads.
Own the architecture of the platform's control and data planes, including pipeline APIs, connectors, and sink integrations, expanding beyond Kafka-to-BigQuery to S3/GCS and Apache Iceberg.
Mentor engineers, drive migrations from legacy systems, and establish robust reliability, security, and operational practices for pipelines running on Kubernetes.
Reddit is a community of communities built on shared interests, passion, and trust, hosting authentic conversations across 100,000+ active communities and approximately 130 million daily active unique visitors. The company fosters an open, collaborative culture with a focus on reliability, performance, and efficiency.
Design and build production-grade systems end-to-end, from problem definition through deployment and operations.
Work across application services, distributed systems, infrastructure, data pipelines, and ML systems, debugging complex issues across multiple layers.
Frame problems correctly, applying ML when needed, and ensure reliability, performance, and cost efficiency.
Moniepoint Inc. is Africa's all-in-one financial platform, helping 20 million businesses and individuals access payments, banking, credit, cross-border, and business management tools. As Nigeria's largest merchant acquirer processing over $250 billion annually, we prioritize our people's well-being and foster a culture of innovation and teamwork.
Lead architecture and technical direction for workflow automation and agentic AI-driven execution.
Build prototypes and reference implementations to prove architectural ideas with real experiments.
Drive the evolution of workflow capabilities toward AI-native patterns including agentic execution and autonomous flows.
ServiceNow is the AI control tower for business reinvention, helping 85% of the Fortune 500 work smarter, faster, and better. They build an AI-native culture where technology and talent are unstoppable together.
Design and develop scalable search and indexing systems for an AI search engine.
Ensure operational excellence by participating in on-call rotation and maintaining system quality.
Collaborate with a global remote team to solve distributed system challenges.
Algolia is a pioneer and market leader in AI Search, empowering over 18,000 businesses to deliver blazing-fast search experiences. With $150 million in Series D funding and a valuation of $2.25 billion, the company fosters a high-trust, flexible culture and values diversity and collaboration.
Architect and evolve core control and context planes, including service registries, SLO enforcement, and automated canary releases.
Own the service chassis and golden path, maintaining multi-language Java/Python libraries, Helm charts, and deployment pipelines.
Drive reliability engineering practices, mentor engineers, and lead architectural strategy for distributed systems at production scale.
The company builds large-scale web data products and distributed engineering infrastructure for AI-driven workflows. It operates a remote-first, globally distributed engineering culture focused on reliability, autonomy, and technical excellence.
Own the architecture health of a billion-scale distributed system, including failure modes, capacity limits, and cross-deployment interactions.
Approve critical-path designs and hunt gaps like single points of failure, unbounded queues, and missing idempotency proactively.
Build the parts nobody else can, prototype risky architectural bets, and ship remediations after serious incidents.
HighLevel is an AI-powered business operating system that gives agencies and SMBs the infrastructure to build, automate and scale. With over 2,000 team members across 10+ countries, HighLevel operates as a global, remote-first organization built for speed and ownership.
Drive technical vision and set the roadmap for the squad, partnering with engineering leadership.
Architect the end-to-end automation platform for package creation, test generation, and image building.
Build AI-powered tooling with LLM integration for manifest generation and quality gates.
Chainguard delivers hardened, secure builds of open source software to secure the software supply chain. They serve Fortune 500 enterprises and are backed by top investors, with a remote-first culture focused on trust, intentional action, and customer obsession.
Set technical direction and own architecture for major services and integrations within Internal Platforms.
Design, build, and maintain Java-based backend services that are scalable, resilient, and well-tested.
Mentor engineers, partner with cross-functional teams, and drive predictable delivery for complex initiatives.
Fanatics builds a leading global digital sports platform for commerce, collectibles, and betting & gaming. With over 22,000 employees, it is committed to enhancing the fan experience and delighting sports fans globally.
Build and maintain core infrastructure for Quora's ML platform, ensuring high availability, scalability, and performance.
Build and improve distributed systems serving ML models in production, from Large Recommendation Models to Large Language Models.
Work on GPU model serving, optimizing latency, throughput, and cost to support larger and more capable models.
Quora's mission is to grow the world's collective intelligence through two platforms: Quora for global knowledge sharing and Poe for AI agent collaboration. We are a remote-first company with passionate, collaborative, and high-performing global teams, rooted in transparency and experimentation.
Design, build, and evolve the event-sourced engine and its C8 REST/gRPC/MCP API, ensuring safe public-API evolution.
Take on hard, ambiguous problems end-to-end — define the problem, write solution designs, and drive delivery across the team.
Ship high-impact engine features fast while maintaining an extremely high quality bar and simplifying the engine.
Camunda is an enterprise platform for agentic orchestration, coordinating AI agents, people, and systems across complex business processes. With over 700 organizations worldwide, including 9 of top 10 US banks, they are a fully remote, global team transforming into an AI-first organization and are Great Place to Work certified.
Partner with cross-functional teams to design and deliver scalable backend systems for major product initiatives.
Own the full software lifecycle from technical design to rollout, using A/B experiments and data analysis to drive decisions.
Build and maintain high-performance APIs and distributed services using modern languages and tools.
Reddit is a community of communities, built on shared interests, passion, and trust. It is home to the most open and authentic conversations on the internet, with 100,000+ active communities and approximately 130 million daily active unique visitors.
Provide senior technical leadership across multiple engineering teams and business domains.
Drive architecture, design, and implementation of scalable, reliable, and maintainable software systems.
Mentor staff and senior engineers and influence technical strategy across the organization.
Oportun is a mission-driven financial services company that provides responsible credit, savings, and budgeting tools to help members build a better financial future. It has provided over $22.7 billion in credit and values speed, high standards, and using AI to work smarter.
Own hands-on technical contribution, team leadership, and delivery accountability for the engineering organization.
Manage technical execution, architecture, engineering quality, and production reliability with a focus on security.
Partner cross-functionally with product, customer experience, and leadership to drive platform vision.
Authorium is a high-growth GovTech SaaS company that helps government agencies modernize administrative operations by replacing legacy systems with a unified platform. The team is small but growing, moves fast, and collaborates across on-shore and off-shore squads to deliver mission-critical software.
Own architecture health of large-scale distributed systems, including failure modes, capacity constraints, and consistency guarantees.
Identify and remediate systemic risks such as single points of failure, unbounded queues, and data-loss scenarios.
Work hands-on with Node.js/Go and GCP technologies to prototype solutions and resolve complex failures.
This company operates large-scale distributed systems processing billions of events and messages. Its engineering culture values technical rigor, proactive problem solving, and clear cross-team communication.
Design and build distributed data systems handling large-scale ingestion and processing.
Drive architectural decisions and take end-to-end ownership of critical components.
Collaborate with product teams to translate ambiguous requirements into robust technical solutions.
Our partner builds a large-scale, multi-chain data platform that ingests, models, and delivers blockchain data to users and developers. They are a remote-first, distributed team with a strong engineering culture focused on ownership and collaboration.