AI Agentic Tester
Denver, MO (Remote)
Must-Have Skills
- Large Language Models (LLMs)
- Retrieval-Augmented Generation (RAG)
- Prompt Engineering Validation
- AI Model Validation & Evaluation
- API Testing (Postman, REST APIs, Swagger)
- Test Automation (Selenium / Playwright / PyTest)
- AI Evaluation Metrics (Accuracy, Hallucination, Relevance, Consistency)
- End-to-End Workflow Testing
- Azure OpenAI / AWS Bedrock / Gemini
Core Responsibilities
- Perform functional testing of AI agent workflows and end-to-end AI applications.
- Validate LLM responses for accuracy, relevance, consistency, and hallucination rates.
- Test RAG pipelines, prompt engineering, and AI agent orchestration.
- Execute API testing using Postman, Swagger, and REST APIs.
- Develop and maintain automation scripts using Python, Selenium, Playwright, or PyTest.
- Validate multi-agent interactions, tool calling, and workflow execution.
- Apply AI testing methodologies and evaluation metrics to ensure response quality.
- Work with Azure OpenAI, AWS Bedrock, Gemini, or similar AI platforms to validate AI solutions.
Originally posted on Himalayas