Develop and maintain scalable, resilient data pipelines using PySpark and distributed processing frameworks.
Modernize legacy processes and design data solutions on AWS, building data products and complex transformations.
Support data quality and optimization initiatives, collaborating with stakeholders to translate requirements into technical solutions.
Jobgether is an AI-powered job matching platform that connects candidates with hiring companies. It focuses on efficiency and fairness in recruitment, promoting diversity and inclusion in the workplace.
Write, optimize, and maintain complex SQL across multiple database platforms and Databricks.
Design and build data ingestion pipelines using API integrations, CDC patterns, and modern data engineering practices.
Develop and maintain Databricks Lakehouse architectures following the Medallion model.
Coforge is a global IT services company specializing in digital transformation and technology solutions. With over 20,000 employees worldwide, the company fosters a collaborative and innovative culture focused on delivering value to clients.
Design and implement a petabyte-scale AWS data platform using advanced engineering and cloud-native practices.
Build and optimize highly scalable data pipelines using JVM languages, Spark, and Python for production-ready data processing.
Act as a principal-level technical leader, mentoring engineers and driving architectural decisions to guide platform evolution.
Experian is a global data and technology company that powers opportunities for people and businesses around the world. With 25,200 employees across 32 countries, they have a people-centric, inclusive culture recognized as a World's Best Workplace.
Design, develop, and maintain scalable relational database solutions for cloud-native and enterprise applications.
Tune database performance across SQL, indexing, schema design, and configuration to meet demanding workloads.
Plan and execute database migrations and platform modernization efforts with strong validation and rollback practices.
Redapt Inc. is a pioneering data center infrastructure integrator and cloud services provider. They focus on delivering innovative solutions that power customers' demanding applications and enable data-driven insights.
Build, train, and deploy machine learning models and data pipelines using Python, pandas, NumPy, and scikit-learn, ensuring scalability and reliability in production.
Design, implement, and continuously improve LLM-powered applications and Retrieval-Augmented Generation (RAG) pipelines, including prompt engineering, embedding strategies, and vector store integration.
Conduct exploratory data analysis and design data visualizations to uncover trends and communicate key metrics to both technical and non-technical stakeholders.
Great Gray Group is a leading independent provider of trustee and administrative services for Collective Investment Trusts, with over $370 billion in assets under management. They are a high-growth organization backed by Madison Dearborn Partners, fostering a culture of collaboration, innovation, and continuous improvement.
Define strategic Databricks adoption roadmaps and platform guardrails for enterprise clients.
Embed with client teams to implement high-priority use cases, then guide them toward self-service.
Act as a data engineering SME in pre-sales and coach client teams on Databricks best practices.
Caylent is an AI-first cloud services company that helps organizations turn ambitious ideas into meaningful business impact. They are a fully remote global company with employees across Canada, the United States, and Latin America, fostering a culture of technological curiosity and inclusion.
Architect enterprise-scale Databricks data platforms with guardrails, automation, and standards to enable safe self-service by business teams.
Design high-volume batch and real-time streaming pipelines using PySpark and Structured Streaming, integrating Unity Catalog with third-party metadata.
Communicate architectural decisions to technical and business stakeholders, produce documentation, and partner with client teams to translate use-case needs into platform capabilities.
Caylent is an AI-first cloud services company that helps organizations turn ambitious ideas into meaningful business impact using AWS, AI, and Anthropic's Claude platform. They are a fully remote global company with employees across Canada, the US, and Latin America, fostering a culture of technological curiosity and inclusion.
Define and build the Bioinformatics function and roadmap for scalable genomics-based algorithm development.
Architect and oversee scalable AWS infrastructure and implement MLOps for clinical production.
Lead, mentor, and recruit a high-performing team of bioinformaticians and data scientists.
The company develops advanced genomic products and clinical algorithms. It fosters an inclusive, collaborative environment with experienced professionals across data science, genetics, medicine, and engineering.
Develop and implement data analysis techniques and ML models using customer data.
Create scalable data pipelines for feature extraction on cloud services.
Collaborate with engineering teams to productionize work and automate processes.
RRD is a leading global provider of marketing, packaging, print, and supply chain solutions. With 32,000 employees across 28 countries, the company values innovation, authenticity, and integrity.
Design, build, and maintain scalable event-streaming pipelines that ingest data from client systems and third-party sources.
Develop and operate analytical databases and data models optimized for high-volume event workloads and low-latency access.
Build production-grade services in Elixir and/or Python for event processing, transformation, and integration.
This company is building data infrastructure for high-volume client and internal event data. They are establishing a new data engineering team and offer a remote work environment with significant technical ownership.