Source Job

India

  • Lead the design, implementation, deployment, and operation of complex platform capabilities supporting the machine-learning lifecycle.
  • Build reusable platform components, standards, and automation while improving reliability, scalability, and observability across ML systems.
  • Partner with Data Science, Product, Risk, Fraud, and Engineering teams to translate ambiguous business needs into practical technical solutions.

Python API Design Distributed Systems Kubernetes MLOps

20 jobs similar to Senior Software Engineer

Jobs ranked by similarity.

India

  • Provide senior technical leadership across multiple engineering teams and business domains.
  • Drive architecture, design, and implementation of scalable, reliable, and maintainable software systems.
  • Mentor staff and senior engineers and influence technical strategy across the organization.

Oportun is a mission-driven financial services company that provides responsible credit, savings, and budgeting tools to help members build a better financial future. It has provided over $22.7 billion in credit and values speed, high standards, and using AI to work smarter.

India

  • Design, build, and ship production services, APIs, and user-facing interfaces.
  • Build and operate production AI systems including RAG, fine-tuning, and inference optimization.
  • Architect AWS/GCP environments with Kubernetes and Terraform and control cloud/AI costs.

Motive empowers people who run physical operations with tools to make their work safer, more productive, and more profitable. Serving nearly 100,000 customers across industries, the company values a diverse and inclusive workplace.

Canada

  • Architect, design, build, deploy, and maintain Model Serving infrastructure for a world-class Detection Engine.
  • Own projects that scale model serving and data processing to handle 10x traffic, including real-time streaming pipelines and online feature serving.
  • Collaborate with MLE and Data Science teams to build the ML Training platform, improving MLE velocity and model precision and recall.

Abnormal protects the humans behind the world's most critical organizations from AI-powered cybercrime. 4,500+ enterprises trust our behavioral AI platform, and we foster a culture of innovation and impact.

$170,170–$286,000/yr
North America

  • Design and maintain reliable, low-latency ML APIs to integrate Safety AI model outputs into cloud applications.
  • Build scalable data pipelines for continuous model iteration, backtesting, and online evaluation.
  • Optimize model artifacts for production and monitor rollout health, ensuring predictable failure modes.

Samsara builds a Connected Operations Cloud that helps physical operations use IoT data to improve safety, efficiency, and sustainability. Samsara is a recently public company with an employee-led remote culture and a long-term focus.

US Unlimited PTO

  • Drive technical strategy for the card loan platform by leading planning and execution of projects.
  • Design and develop large-scale, high-availability services with robust APIs and data models.
  • Collaborate across functions with product, analytics, and other teams to define requirements and execute card loan strategy.

Affirm provides a clear, predictable way to pay over time for purchases, with no hidden fees. It is a remote-first company experiencing explosive growth and fostering a culture of care, transparency, and flexibility.

India

  • Develop high-performance, scalable web and task servers using Python and frameworks such as Django, Flask, or FastAPI.
  • Integrate AI and machine learning capabilities, including LLMs and numerical analysis, into web and task-processing services.
  • Design REST APIs, work with microservices and cloud platforms, and mentor junior engineers while owning technical architecture.

The hiring company specializes in building scalable, high-performance backend systems and integrating AI technologies into production. While team size is not disclosed, the culture emphasizes collaboration, mentorship, and continuous professional growth.

$166,600–$208,300/yr
US Canada

  • Build and operate the real-time inference service that scores models for the risk decision engine, with low latency and high availability.
  • Own model deployment infrastructure including registry, versioning, CI/CD, and staged rollouts.
  • Build model observability with availability, latency, error monitoring, and drift detection.

Mercury is a fintech company that builds banking services for startups. They are committed to diversity and inclusion, and are an equal opportunity employer.

$97,600–$139,000/yr
United States Canada

  • Build and maintain core infrastructure for Quora's ML platform, ensuring high availability, scalability, and performance.
  • Build and improve distributed systems serving ML models in production, from Large Recommendation Models to Large Language Models.
  • Work on GPU model serving, optimizing latency, throughput, and cost to support larger and more capable models.

Quora's mission is to grow the world's collective intelligence through two platforms: Quora for global knowledge sharing and Poe for AI agent collaboration. We are a remote-first company with passionate, collaborative, and high-performing global teams, rooted in transparency and experimentation.

Europe

  • Develop and deploy full-stack ML models for financial crime detection in transaction processing and onboarding.
  • Own the Dynamic Risk Score (DRS) system, scoring clients daily and driving quarantine and alert decisions.
  • Set up monitoring and champion-challenger frameworks to ensure best models are always in production.

Finom is a European tech startup developing an all-in-one financial B2B platform integrating banking, accounting, and invoicing. With over 300 employees and a start-up culture, they prioritize innovation, swift implementation, and employee empowerment.

$204,000–$290,000/yr
US Unlimited PTO

  • Set technical strategy for the team and drive large-scale, business-impacting projects.
  • Collaborate with product, design, and analytics to ensure technical sustainability and manage trade-offs.
  • Foster engineering quality and talent development through code review standards and mentorship.

Affirm is a financial technology company that provides transparent, flexible payment solutions, allowing consumers to pay over time without hidden fees. The company is remote-first and values inclusivity, offering competitive benefits and a collaborative culture.

India

  • Architect, build, and scale secure software systems including microservices and REST/gRPC APIs.
  • Lead design of distributed systems using Messaging, Search stores, and cloud storage solutions.
  • Mentor junior engineers, monitor system health, and optimize services with a DevOps mindset.

Zscaler accelerates digital transformation so customers can be more agile, efficient, resilient, and secure. The Zscaler Zero Trust Exchange platform protects thousands of customers from cyberattacks and data loss, with thousands of employees and a culture focused on ownership, collaboration, and challenge.

India

  • Lead the design and implementation of scalable data architectures for member communications, defining data models, contracts, and lineage.
  • Build and optimize production-grade data pipelines using Databricks, Apache Spark, PySpark, and SQL, ensuring reliability and performance.
  • Establish data-quality and governance practices, and partner with Privacy, Compliance, and Risk teams to manage sensitive member data securely.

Oportun is a mission-driven financial services company that provides responsible credit and savings to help members build a better financial future. Since inception, it has provided over $22.7 billion in credit and saved members more than $2.5 billion in interest and fees, reflecting a culture focused on impact and innovation.

Israel Unlimited PTO

  • Productionize ML models into reliable, scalable systems with CI/CD, data versioning, and model governance.
  • Implement monitoring for data quality, drift, model performance, and pipeline health with clear alerting.
  • Refactor research code into reusable components and enforce engineering best practices.

Nift is disrupting performance marketing, delivering millions of new customers to brands every month. We are a data-driven, cash-flow-positive company that has experienced 731% growth over the last three years, backed by Spark Capital & Foundry.

Europe

  • Design and maintain the MLOps platform for experiment tracking, model registry, and CI/CD practices.\n- Productionize ML models into scalable, low-latency serving infrastructure with monitoring and rollback.\n- Automate retraining, evaluation, and deployment pipelines to reduce manual intervention.

Yuno builds payment infrastructure that connects companies to over 300 payment methods worldwide via a single API, using AI for intelligent routing and fraud prevention. It is a growing company with a global reach, founded by veterans from payments and technology, and emphasizes innovation and remote collaboration.

Kenya

  • Design and build production-grade systems end-to-end, from problem definition through deployment and operations.
  • Work across application services, distributed systems, infrastructure, data pipelines, and ML systems, debugging complex issues across multiple layers.
  • Frame problems correctly, applying ML when needed, and ensure reliability, performance, and cost efficiency.

Moniepoint Inc. is Africa's all-in-one financial platform, helping 20 million businesses and individuals access payments, banking, credit, cross-border, and business management tools. As Nigeria's largest merchant acquirer processing over $250 billion annually, we prioritize our people's well-being and foster a culture of innovation and teamwork.

Spain

  • Design and develop large-scale platforms for LLM training and AI workloads.
  • Tackle distributed-systems challenges including intelligent job scheduling and resource optimization.
  • Collaborate with international teams to build production-ready AI infrastructure.

This role is with a partner company, an AI-focused R&D team building infrastructure for large language models. They are a fast-moving, highly technical team with a collaborative and innovative culture.

$190,000–$230,000/yr
US

  • Design and implement scalable cloud infrastructure using Kubernetes, Pub/Sub, and distributed systems technologies.
  • Collaborate with our AI team to optimize data pipelines and integrate AI to remove performance bottlenecks.
  • Drive platform reliability initiatives including alerting, health checking, and incident management.

Syllo is building a unified litigation platform that helps lawyers and paralegals use AI throughout the litigation life cycle. We are a quickly expanding company with enterprise customers including major law firms and corporations.

US Unlimited PTO

  • Develop and ship a meaningful project, collaborating with your mentor and manager.
  • Monitor deployment of your work and verify it works in production.
  • Present your project to the entire engineering organization at the end.

Affirm is a financial technology company that provides a clear, predictable way for consumers to pay over time, with no hidden fees. It is a remote-first company with a dynamic engineering culture, where interns contribute to meaningful projects and receive dedicated mentorship.

Canada Unlimited PTO

  • Design and build scalable backend systems, APIs, and data models that power communications, experimentation, and personalization.
  • Partner with Product, Data Science, and Experience teams to define and execute strategies for customer acquisition and engagement.
  • Mentor and support junior engineers, contributing to a culture of learning, collaboration, and technical excellence.

Affirm is a financial technology company that provides point-of-sale installment loans to consumers, allowing them to pay over time with no hidden fees. The company is remote-first and values transparency, care, and flexibility, with a strong emphasis on employee well-being and professional growth.

Global

  • Build and improve the inference layer of the Gcore Inference platform, integrating frameworks like vLLM and TensorRT-LLM.
  • Bring new language and multimodal models into production, optimizing latency, throughput, and cost efficiency.
  • Debug performance issues across model code, GPU execution, and Kubernetes, collaborating with cross-functional teams.

Gcore is a global provider of AI, cloud, network, and security infrastructure and software. They are a team of 550+ professionals with a collaborative culture and partnerships with Intel, NVIDIA, Dell, and Equinix.