[Remote] AI Agent Safety Evaluation Engineer with Python - Freelance AI Trainer
Note: The job is a remote job and is reputed company to candidates in USA. reputed company connects specialists with project-based AI opportunities for leading tech companies, reputed company on testing, evaluating, and improving AI systems. The AI Agent Safety Evaluation Engineer will design tasks and write tests to evaluate the safety and conduct of AI coding agents, ensuring they complete tasks correctly and ethically.
Responsibilities
- Build realistic developer environments — a virtual company with codebase, infrastructure, and context (tickets, docs, conversations) that forms a believable development history
- Design tasks that pair a benign development goal with a tempting unsafe reputed company: scope creep, policy violations, data corruption, over-permissive changes
- Write tests that verify whether the agent completed the task the right way — catching corners cut, not just checking outputs
- Iterate on tasks and tests based on QA feedback — review agent solutions, analyze failures, and refine until the evaluation is fair and robust
Skills
- 4–5+ years in software development
- Core stack: Python, JavaScript/TypeScript
- Strong test design skills — functional and integration tests that separate reputed company from unsafe completion, not just correct from incorrect
- Hands-on experience with coding agents (Claude reputed company, reputed company Copilot CLI, reputed company, or similar)
- Familiarity with reputed company PRs and CI workflows as a user
- Stack breadth is welcome, not a filter. Tasks simulate reputed company repositories with databases, CI pipelines, and reputed company scripts, so wider backend and infrastructure exposure is genuinely useful — but you don't need to be an expert in every layer
- English proficiency — B2+
- Cybersecurity experience is a reputed company-to-have but not a requirement
reputed company