- Salary
- $60 – $100/hr
- Workplace
- Remote
- Type
- Contract
- Department
- IT
- Closing date
- Today
- Source
- CareersPage
Description
About the work
We are building high-fidelity simulated work environments used to evaluate and improve AI agents on real marketing work. Each environment reproduces a marketing org's actual tool surface — email, storage, CRM, project management, social media management, web analytics, AEO/SEO, ads, CMS, product analytics and support — populated with realistic documents, dashboards, personas and deliberately planted problems.
You will help design and pressure-test the Product Marketing environments: the briefs, the artifacts, the judgment calls a strong practitioner would make, and the errors a weaker one would miss.
What you will do
- Design task briefs drawn from work you have actually done in product marketing, including what the task is really testing — the planted issue and the decision a strong practitioner should reach
- Specify the documents, dashboards, personas and tool states a realistic version of that task requires
- Write or review the reference answer and the criteria that separate a strong response from a plausible-but-wrong one
- Review AI agent attempts and judge them against your own standard
Tasks span the capabilities we measure: diagnosing what happened from messy or conflicting data, prioritizing and making tradeoffs, planning and executing, QA and reconciliation, triage and escalation, research and evaluation, reporting, and orchestrating multi-step work across several tools.
Product Marketing scope
Win/loss analysis, positioning, launch readiness and launch execution, adoption analysis, and competitive response.
Who we are looking for
- Hands-on practitioner experience in product marketing — you have owned launches, positioning and enablement, not only managed people who do it
- Comfortable working across a CRM such as Salesforce, product analytics such as Amplitude, and a project tracker such as Asana or Linear
- Prior experience building AI training environments, RL environments or simulated case studies is strongly preferred
- Rubric Academy or Rubric Bootcamp certification is strongly preferred
- Clear written reasoning — much of the value is in explaining why a decision is right
Screening
Applicants complete a short multiple-choice knowledge screener specific to this sub-domain before review.