Quick Overview
Seniority
Mid Senior
Work mode
Hybrid
Location
Lincolnshire, IL, United States
Posted
Yesterday
Machine LearningSeleniumComputer VisionCypressJiraPlaywrightPythonRESTiOSAndroid
Job Description
ApTask Global Workforce (AGW) is seeking a AI Quality Analyst with an information technology company. This is a 5+ month contract opportunity with our client located in Lincolnshire, IL. This is a hybrid position.
Summary:
The AI Quality Analyst is a critical role responsible for ensuring the performance, safety, and reliability of our cutting-edge AI/ML models. You will be at the forefront of our development lifecycle, designing and executing comprehensive evaluation strategies to identify model weaknesses, potential biases, and critical edge cases. This role requires a blend of analytical rigor, technical aptitude, and a deep curiosity for how AI models behave in real-world scenarios. You will not just find bugs, but provide the actionable insights that drive model improvement and guide our research and development efforts.
Responsibilities:
If you have the described qualifications and are interested in this exciting opportunity, please apply!
ApTask Global Workforce (AGW)
ApTask Global Workforce (AGW) is a certified Minority and Veteran workforce solutions company. AGW delivers operational, clinical, lab and professional talent with a strong focus on healthcare and life sciences. The company supports clients with reliable staffing, program expertise and a commitment to quality, speed and consistent delivery.
Benefits of working with ApTask Global Workforce include:
Summary:
The AI Quality Analyst is a critical role responsible for ensuring the performance, safety, and reliability of our cutting-edge AI/ML models. You will be at the forefront of our development lifecycle, designing and executing comprehensive evaluation strategies to identify model weaknesses, potential biases, and critical edge cases. This role requires a blend of analytical rigor, technical aptitude, and a deep curiosity for how AI models behave in real-world scenarios. You will not just find bugs, but provide the actionable insights that drive model improvement and guide our research and development efforts.
Responsibilities:
- Evaluation Strategy & Benchmark Development: Design, develop, and maintain a comprehensive suite of test cases and evaluation benchmarks. Proactively identify potential model failure points, including edge cases, adversarial inputs, and sources of bias.
- Error Analysis & Failure Triage: Conduct systematic error analysis to categorize model failures and identify underlying patterns. Triage defects, prioritize them based on severity and impact, and work with the development team to ensure resolution.
- Data Sourcing & Curation: Source, curate, and manage high-quality datasets for model evaluation and testing. This includes performing data annotation and validation to ensure the integrity of our ground-truth data.
- Exploratory & Adversarial Testing (Red Teaming):Perform unscripted, exploratory testing to discover unexpected model behaviors. Participate in red teaming exercises to intentionally challenge our models and identify potential safety and security vulnerabilities.
- Test Environment Management: Set up, maintain, and troubleshoot testing and demonstration environments to ensure a stable and reliable evaluation pipeline.
- Reporting & Insights: Analyze and synthesize test results into clear, actionable reports for both technical and non-technical stakeholders. Translate complex findings into concrete recommendations for model improvement.
- Process Improvement: Actively participate in post-hoc evaluation reviews and contribute to the continuous improvement of our testing methodologies, tools, and overall quality assurance processes.
- Proven experience in a quality assurance, testing, or data analysis role, preferably within the AI/ML domain.
- A deep understanding of the machine learning lifecycle and the common failure modes of AI models.
- Hands-on experience with data annotation, data validation, and managing large datasets.
- Meticulous attention to detail and a methodical approach to problem-solving.
- Strong analytical skills with the ability to identify patterns in data and draw meaningful conclusions.
- Expertise with industry-standard test automation tools and libraries (e.g., Selenium, Playwright, Cypress, REST-assured).
- Experience in testing across different platforms (e.g. comprehensive testing of mobile Android/iOS and web applications)
- Experience with bug tracking systems (e.g., Jira) and test case management tools.
- Scripting skills (e.g., Python) for test automation and data manipulation.
- (Preferred) experience in testing AI systems, including evaluating agentic responses, model performance metrics, and data integrity.
- Excellent communication skills, with the ability to clearly document bugs and articulate complex technical issues.
- Familiarity with computer vision or other specific AI domains relevant to our work.
If you have the described qualifications and are interested in this exciting opportunity, please apply!
ApTask Global Workforce (AGW)
ApTask Global Workforce (AGW) is a certified Minority and Veteran workforce solutions company. AGW delivers operational, clinical, lab and professional talent with a strong focus on healthcare and life sciences. The company supports clients with reliable staffing, program expertise and a commitment to quality, speed and consistent delivery.
Benefits of working with ApTask Global Workforce include:
- Medical
- Dental
- Vision
- Sick Pay (for applicable states/municipalities)
- Our team stays close to the process and is here to guide you every step of the way. To learn more, please visit our website
ApTask Global Workforce is an Equal Opportunity Employer. Candidates will receive consideration without regard to their race, color, religion, sex, sexual orientation, gender identity, national origin, disability or status as a protected veteran.
Similar jobs
- LA
Quality Engineer & Manager
NewLaborup
Oak Ridge, Tennessee🇺🇸On-site13 hours agoAuditingCNCCompliance+2Engineering - RH
Quality Assurance Tester
NewRobert Half
Scottsdale, AZ🇺🇸HybridYesterdayAWSScrumAgile+3Technology - CR
Mid-Level Tester
NewCredence
McLean, Virginia🇺🇸Hybrid11 hours agoSAFe401kScrum+6 - TH
Senior Quality Assurance Engineer
NewTherapyNotes.com
Philadelphia, Pennsylvania🇺🇸RemoteYesterdaySeleniumAgileC#+17Technology - SS
Product Quality and Test Engineer Group Lead
NewScientific Systems Company, Inc.
Burlington, Massachusetts🇺🇸Hybrid9 hours ago401kRoboticsCompliance+1Technology - DS
(196948) Lead Software Development Engineer in Test (Lead SDET)
NewDiversified Services Network, Inc.
Chicago, Illinois🇺🇸Hybrid15 hours ago401kSQLAWS+11Technology