Similar Jobs

See all

Responsibilities:

  • Help build and maintain the core infrastructure that powers Quora's ML platform, ensuring high availability, scalability, and performance.
  • Build and improve the distributed systems that serve our ML models in production, from Large Recommendation Models (LRMs) to Large Language Models (LLMs).
  • Work on GPU model serving, optimizing latency, throughput, and cost to support larger and more capable models.

Minimum Requirements:

  • Availability for meetings and impromptu communication during Quora's coordination hours (Mon-Fri: 9am-3pm Pacific Time).
  • A 2025 or 2026 graduate with or pursuing a B.S., M.S., or Ph.D. in Computer Science, Engineering, or a related technical field.
  • Genuine interest in large-scale distributed systems, infrastructure, and machine learning.

Preferred Requirements:

  • Previous software engineering experience via an internship, work experience, open-source contribution, or coding competition.
  • Coursework or hands-on experience with ML frameworks such as PyTorch or TensorFlow.
  • Exposure to Kubernetes, Docker, or cloud technologies like AWS.

Quora

Quora's mission is to grow the world's collective intelligence through two platforms: Quora for global knowledge sharing and Poe for AI agent collaboration. We are a remote-first company with passionate, collaborative, and high-performing global teams, rooted in transparency and experimentation.

Apply for This Position