What We're Researching
We are hiring senior full-stack engineers to help build and refine next-generation coding environments that evaluate AI agents. This work directly supports the development of robust grading systems that accurately measure how effectively an AI agent solves complex coding tasks. The goal is to create secure evaluation suites that cannot be bypassed or easily cheated by the models.
How It Works
Throughout this long-term engagement, you will work remotely with real website clones to identify bugs and write technical specifications. You will build comprehensive automated test suites designed to grade AI-generated code. This involves analyzing how an agent interacts with a given task and ensuring the grading logic is watertight. The process requires deep technical scrutiny and creative problem-solving to account for unpredictable AI behavior.
Who This Is For
We welcome senior full-stack software engineers with extensive hands-on experience in React. You should be highly comfortable writing technical specs and building automated testing frameworks for complex web environments. Candidates with a strong background in web security, QA automation, or AI evaluation are highly encouraged to apply.
What You'll Do
- Interact with real website clones to identify bugs and workflow gaps
- Write detailed technical specifications for complex AI coding tasks
- Build robust automated test suites to securely grade AI agent performance
- Ensure grading logic is watertight and resistant to AI bypassing or shortcuts
Who Should Apply
- Senior-level experience in full-stack software engineering
- Deep professional expertise building web applications with React
- Strong background in automated testing and technical specification writing
- Available for an ongoing, high-commitment technical engagement
Contract & Payment Terms
- You will be engaged as an independent contractor.
- This is a fully remote opportunity that can be completed on your own schedule.
- Opportunities can be extended, shortened, or concluded early depending on needs and performance.
- Your participation will not involve access to confidential or proprietary information from any employer, client, or institution.
- Payments are processed weekly based on services rendered.
- We are unable to support H1-B or STEM OPT candidates at this time.
About Terac
Terac is an AI-powered qualitative research platform that automates the entire research process — from recruiting experts to conducting AI-driven interviews in 50+ languages, analyzing results, and managing compliance. Trusted by leading companies, Terac enables teams to run hundreds of interviews simultaneously and generate comprehensive insights in hours, not months.