MailerMen logo

LLM Red Teaming and Safety Prompts Intern

MailerMen

Posted yesterdayBe an early applicant

What this role pays, in context

This listing does not state its pay. AI & Prompt Engineering roles in India currently advertise a median of ₹19.9K a month.

25th pct ₹17.9K75th pct ₹21.6K

Median of 6 other live AI & Prompt Engineering listings in India on this board that state pay. Roles that do not disclose a salary are excluded rather than counted as zero.

Every assistant we run can be talked into something it should not do, and the only real question is whether we find the route first or a user does.

Eight weeks of breaking things

This short internship puts you on the adversarial side. You will systematically attack our candidate assistant and our recruiter agent: prompt injection hidden in uploaded resumes, instructions buried in job descriptions, roleplay framings that coax out the system prompt, and requests that nudge the model toward discriminatory screening advice. Finding a break is the easy half. You then write the guardrail prompt or input filter that closes it, add the case to a permanent regression suite, and confirm the fix has not made the assistant uselessly cautious. That tradeoff between safety and usefulness is the actual work.

Terms, stated plainly

The position is unpaid and runs for two months, with two days a week in the Bengaluru office. You will work directly with the engineer responsible for model safety, your test cases become part of what ships, and you will finish with a body of documented findings that demonstrates a real and rare skill.

Responsibilities

  • Systematically probe our assistants for prompt injection, jailbreaks and system prompt leakage
  • Test resume and job description uploads as injection vectors
  • Write guardrail prompts and input filters that close the gaps you find
  • Add every confirmed break to a permanent regression suite
  • Verify that new guardrails do not cause refusals on legitimate requests
  • Document each finding with reproduction steps and a severity note

Requirements

  • Final year student or recent graduate with an interest in security or AI safety
  • Basic Python and comfort working with an API client
  • Creative and persistent, the sort of person who finds the case nobody considered
  • Able to write clear, reproducible reports
  • Able to attend the Bengaluru office two days a week for the eight week term

Skills

Benefits

  • Certificate + Letter of Recommendation
  • Direct work with the engineer who owns model safety at MailerMen
  • Your test cases become part of the permanent safety suite
  • Flexible hours outside the two required office days