Logo for Linda Werner & Associates, Inc

Content Specialist III | AI Evaluation & Prompting

Role overview

Qualifications

  • 5+ years of experience in writing, editing, journalism, production, linguistics, STEM, coding, policy, or another relevant subject matter field
  • 1+ year of hands on AI experience preferred, including prompting, annotation, evaluation, or red teaming
  • Bachelor’s degree or equivalent experience
  • Deep expertise in at least one subject area with the ability to evaluate content as a subject matter expert

Responsibilities

  • Test new AI model versions across a variety of topics, use cases, and conversation types
  • Evaluate model responses against established rubrics, guidelines, and quality standards
  • Identify and document specific examples of successful and unsuccessful model behavior
  • Write and refine system prompts that help shape model personality, tone, and behavior

Key facts

Hard skills

Other skills

  • Writing
  • Analytical Thinking
  • Editing
  • Analytical Skills
  • Detail Oriented
  • Adaptability

About the company

Linda Werner & Associates, Inc logo

Linda Werner & Associates, Inc

Business Consulting & Services

Linda Werner & Associates is a technology, consulting and technical staffing firm, providing acumen and solutions that include content management, learning, knowledge management, and Web solutions. Our delivery models range from consulting, through managed teams and staffing services. Linda Werner & Associates serves various industries, including software, retail, financial, telecommunications, and others.

Company details

IndustryBusiness Consulting & Services
Company size51-200

Your match analysis

See how your profile stacks up against this role.

We compared the job requirements to your profile to show where you're strong and where you fall short.

Job description

We are seeking an experienced Content Specialist III to help evaluate, refine, and improve advanced AI models and products.

In this role, you will work directly with evolving AI systems, testing how they respond across a wide range of topics and conversation types, identifying where they succeed or fall short, and helping shape the prompts, quality standards, and evaluation frameworks that influence how they communicate and behave.

This is a highly hands on opportunity for someone who brings together strong writing and content expertise, curiosity about AI, thoughtful judgment, and a sharp eye for quality. Your work will directly contribute to improving real world AI product experiences and how these systems perform for users.

What You Will Do

• Test new AI model versions across a variety of topics, use cases, and conversation types

• Evaluate model responses against established rubrics, guidelines, and quality standards

• Identify and document specific examples of successful and unsuccessful model behavior

• Write and refine system prompts that help shape model personality, tone, and behavior

• Assess whether evaluation rubrics and quality standards effectively measure model performance

• Perform quality reviews of both human and agent based evaluations to ensure accuracy and compliance

• Conduct hands on experiments with AI models and products

• Investigate model failures and document findings clearly

• Apply detailed instructions and evaluation criteria consistently across a high volume of work

• Adapt quickly as product priorities, model behavior, and evaluation needs evolve

Top Skills

• AI Model Evaluation

• Prompt Writing and Refinement

• Writing and Editing

• Rubric Based Quality Assessment

• Fact Checking and Analytical Judgment

Qualifications

• 5+ years of experience in writing, editing, journalism, production, linguistics, STEM, coding, policy, or another relevant subject matter field

• 1+ year of hands on AI experience preferred, including prompting, annotation, evaluation, or red teaming

• Bachelor’s degree or equivalent experience

• Deep expertise in at least one subject area with the ability to evaluate content as a subject matter expert

• Strong judgment and attention to detail with the ability to consistently apply detailed rubrics, verify information, and identify quality issues

Preferred Experience

• Experience evaluating generative AI or large language model outputs

• Strong fact checking and research skills

• Ability to verify claims against original sources

• Experience conducting detailed failure investigations

• Clear written communication and documentation skills

• Ability to flag questions and blockers early

• Comfortable shifting priorities as product needs change

This team operates in a fast moving environment where priorities can change from day to day. Successful candidates will be comfortable balancing quality, speed, volume, and independent judgment while following detailed instructions.

You should enjoy getting into the details, identifying subtle differences in AI responses, determining why something works or fails, and translating those observations into clear and actionable feedback.

Location: United States (Remote)

Role type: Contract 6 Month Position

Expected hours: 40 per week

Benefits:

  • Dental insurance
  • Health insurance
  • Health savings account
  • Life insurance
  • Paid time off
  • Retirement plan
  • Vision insurance

Schedule:

  • 8 hour shift
  • Monday to Friday

Application Question(s):

  • Do you or will you in the future require any sponsorship to work in the US?

Language:

  • English  (Required)

Apply once. Then go straight to the hiring manager.

After you apply, unlock the direct contact details of the people who actually make the call. A quick follow-up makes you 5x more likely to land an interview.

MR

Marcus Rivera

Chief Revenue Officer

m.rivera@company.com
linkedin.com/in/marcusrivera
Unlocked after you apply
·

Content Specialist Related jobs

Other jobs at Linda Werner & Associates, Inc

Premium

Reach out to the hiring manager directly.

Gain access to the contact details of the hiring managers who actually decide, and reach out to network with them directly. That, plus more when you upgrade:

  • Full match report with fit score and gaps
  • Career diagnostics on how recruiters read you
  • Curated company matches and warm intros
  • 48h early access to new roles

Cancel anytime.