AI Agentic Tester
TypeContract
LocationUnited States
Posted4 hours ago
AI Agentic Tester
Denver, MO (Remote)
Must-Have Skills
* AI Agentic Testing
* Functional Testing
* Generative AI Testing
* Large Language Models (LLMs)
* AI Agents & Agentic AI
* Retrieval-Augmented Generation (RAG)
* Prompt Engineering Validation
* AI Model Validation & Evaluation
API Testing (Postman, REST APIs, Swagger)
* Python
Test Automation (Selenium / Playwright / PyTest)
AI Evaluation Metrics (Accuracy, Hallucination, Relevance, Consistency)
* End-to-End Workflow Testing
* Multi-Agent Testing
* Azure OpenAI / AWS Bedrock / Gemini
Core Responsibilities
Perform functional testing of AI agent workflows and end-to-end AI applications.
Validate LLM responses for accuracy, relevance, consistency, and hallucination rates.
Test RAG pipelines, prompt engineering, and AI agent orchestration.
Execute API testing using Postman, Swagger, and REST APIs.
Develop and maintain automation scripts using Python, Selenium, Playwright, or PyTest.
Validate multi-agent interactions, tool calling, and workflow execution.
Apply AI testing methodologies and evaluation metrics to ensure response quality.
Work with Azure OpenAI, AWS Bedrock, Gemini, or similar AI platforms to validate AI solutions.
Originally posted on Himalayas
Denver, MO (Remote)
Must-Have Skills
* AI Agentic Testing
* Functional Testing
* Generative AI Testing
* Large Language Models (LLMs)
* AI Agents & Agentic AI
* Retrieval-Augmented Generation (RAG)
* Prompt Engineering Validation
* AI Model Validation & Evaluation
API Testing (Postman, REST APIs, Swagger)
* Python
Test Automation (Selenium / Playwright / PyTest)
AI Evaluation Metrics (Accuracy, Hallucination, Relevance, Consistency)
* End-to-End Workflow Testing
* Multi-Agent Testing
* Azure OpenAI / AWS Bedrock / Gemini
Core Responsibilities
Perform functional testing of AI agent workflows and end-to-end AI applications.
Validate LLM responses for accuracy, relevance, consistency, and hallucination rates.
Test RAG pipelines, prompt engineering, and AI agent orchestration.
Execute API testing using Postman, Swagger, and REST APIs.
Develop and maintain automation scripts using Python, Selenium, Playwright, or PyTest.
Validate multi-agent interactions, tool calling, and workflow execution.
Apply AI testing methodologies and evaluation metrics to ensure response quality.
Work with Azure OpenAI, AWS Bedrock, Gemini, or similar AI platforms to validate AI solutions.
Originally posted on Himalayas
Apply on Himalayas →
Job sourced from Himalayas. Applications happen directly on the original platform — we never collect your data.