Define and lead QA strategy for Kubernetes and cloud-native environments, aligning with business goals, release deadlines, and best practices.
Architect test frameworks for microservices and infrastructure components, designing complex scenarios for multi-cluster and hybrid-cloud setups.
Spearhead scalable test automation using CI/CD tools, Kubernetes APIs, and Helm, while mentoring engineers in advanced testing techniques.
Mirantis is the Kubernetes-native AI infrastructure company, helping organizations build scalable, secure infrastructure for AI and data-intensive applications. The company is a leader in container management and employs passionate, talented individuals working with Fortune 500 and Global 2000 customers.
Audit 30% of production output during the pilot phase to ensure rating quality and consistency.
Identify and flag miscalibration or inconsistent scoring across raters.
Provide clear, actionable feedback and corrections to raters.
Welo Data provides AI services, including data annotation and rating for machine learning projects. It is a community of freelancers working on various AI-related tasks.
Audit 30% of production output to ensure rating quality and consistency during the initial phase.
Identify miscalibration or inconsistent scoring across raters and flag issues early.
Provide clear, actionable feedback and corrections to raters to maintain alignment.
Welo Data provides AI services and data solutions for machine learning and artificial intelligence projects. They are a growing community of freelancers and contractors working on various AI-related tasks.
Lead, mentor, and develop a team of QA engineers across manual and automation specialties.
Establish and enforce QA standards, drive automation maturity, and promote quality advocacy.
Collaborate with cross-functional teams to ensure high-quality, secure software delivery.
Impiricus is the first and only AI-powered HCP Engagement Engine connecting healthcare professionals to pharma resources ethically. Named the #1 fastest growing company in North America by Deloitte in 2025, they foster a culture of ethical innovation, collaboration, and clinical impact.
Evaluate AI-generated responses for relevance, accuracy, and personalization using personalized prompts and data from connected Google applications.
Identify subtle issues such as incorrect assumptions, irrelevant recommendations, inconsistencies, and inappropriate personalization.
Provide clear, detailed, and structured feedback to support improvements to AI models and personalization systems.
Our partner company is seeking an AI Response Quality Evaluator to improve AI-generated responses. This is a project-based contract role with a remote, independent working environment and a duration of up to 16 weeks.
Plan testing scope and timelines, maintain test cases and documentation.
Test and stabilize patches in fast release cycles with clear defect reporting.
Improve testing processes by standardizing repetitive checks and leveraging AI tools.
Social Discovery Group (SDG) solves loneliness and disconnection through social discovery products that connect people online across cultures and regions. Our international team of digital nomads works remotely worldwide and we are a two-time Great Place to Work winner (USA & Japan, 2024-2025) and a top company for work-from-anywhere jobs (FlexJobs, 2025).
Lead and support projects providing independent quality assurance, IV&V, project management, and business process improvement for state government clients.
Report on project status, risks, and issues to client executive leadership and facilitate meetings with confidence.
Provide Workday Financials implementation expertise and review client project artifacts, while contributing to a positive team culture.
BerryDunn is a professional services firm that helps businesses, nonprofits, and government agencies solve challenges. The firm has been recognized for its inclusive culture and focus on learning, development, and well-being.
Design and implement scalable automated testing frameworks for data pipelines and ML models.
Validate large-scale distributed data systems for accuracy, reliability, and performance.
Mentor junior SDETs and contribute to testing standards and best practices.
Quanata is an insurance technology innovation company that engineers advanced risk prediction and prevention solutions and builds a full-stack, flexible, digital insurance platform. It is wholly owned by State Farm and prioritizes an inclusive and positive culture.
Design and execute comprehensive test plans for web and API automation, ensuring functionality aligns with specifications.
Collaborate with development teams to resolve issues, track results, and communicate status clearly.
Perform non-functional testing, including security, accessibility, and AI-feature validation.
Pulse iD is a fintech start-up providing an end-to-end intelligent platform for offer sourcing, merchant-funded rewards, and geolocation services. They collaborate with financial institutions and enterprises across Asia, striving to enhance client experiences and drive business growth.