inhousefyi
← Back to listings

Red Teaming Fellowship

10a LabsRemote · Posted 2 months ago
Full-timeRemoteEst. 52,000 USD
Apply now

Description

About 10a Labs: 10a Labs is the safety and threat-intelligence layer trusted by frontier AI labs, AI unicorns, Fortune 10 companies, and leading global technology platforms. Our adversarial red teaming, model evaluations, and intelligence collection enable engineering, safety, and security teams to stay ahead of evolving threats and deploy AI systems safely.

About the role: As a Red Teaming Fellow, you will help evaluate the safety, security, and reliability of advanced AI systems. Fellows work alongside researchers, engineers, and subject-matter experts to identify vulnerabilities, uncover failure modes, and generate insights that help organizations deploy AI systems more safely.

This is a hands-on fellowship at the intersection of AI, security, and adversarial testing. Fellows will contribute to real-world evaluations of frontier models and AI-enabled systems, helping design and execute tests that probe how models behave under challenging or unexpected conditions.

In this role, you will:

  • Conduct adversarial testing of AI systems across safety, security, and misuse scenarios
  • Design and execute red teaming exercises against models, agents, and AI-enabled applications
  • Develop prompts, attack strategies, and test cases to identify vulnerabilities and failure modes
  • Analyze model behavior and document findings in clear, actionable reports
  • Support the development of evaluation frameworks, taxonomies, and testing methodologies
  • Research emerging AI threats, attack techniques, and risk trends
  • Collaborate with engineers, researchers, and subject-matter experts to investigate emerging risks and threat vectors
  • Develop front-end dashboards and other visualizations

Qualifications:

  • Pursuing or recently completed a degree in Computer Science, Cybersecurity, Engineering, Data Science, Mathematics, Political Science, International Relations, or a related field
  • Strong interest in AI safety, cybersecurity, red teaming, or adversarial testing
  • Familiarity with AI abuse including jailbreaking, system prompt extraction, indirect prompt injection, data exfiltration, etc.
  • Strong analytical, research, problem-solving, and communication skills
  • Ability to think creatively from an adversarial perspective
  • Ability to work independently and collaboratively in a remote environment
  • Experience with Python, Bash, large language models, AI systems, cybersecurity, or research methodologies preferred
  • Proficiency in a language other than English preferred

Benefits:

  • Flexible start / end dates
  • Remote work (based in the continental U.S.)
  • Flexible schedule, up to 20 hours per week (negotiable)
  • Hourly pay commensurate with experience and qualifications
  • $25 per hour for undergraduate students
  • $32.50 per hour for graduate students

Similar jobs

10a LabsRemote

Est. 110,000 USD

About 10a Labs: 10a Labs is the safety and threat-intelligence layer trusted by frontier AI labs, AI unicorns, Fortune 10 companies, and leading global technology platforms. Our adversarial red teaming, model evaluations…

Full-timeRemote
10a LabsRemote

Est. 60,000 USD

About 10a Labs: 10a Labs is the safety and threat-intelligence layer trusted by frontier AI labs, AI unicorns, Fortune 10 companies, and leading global technology platforms. Our adversarial red teaming, model evaluations…

Full-timeRemote
Research FellowJul 7, 2025
10a LabsRemote

Est. 52,000 USD

About 10a Labs: 10a Labs is the safety and threat-intelligence layer trusted by frontier AI labs, AI unicorns, Fortune 10 companies, and leading global technology platforms. Our adversarial red teaming, model evaluations…

Full-timeRemote
10a LabsRemote

Est. 147,500 USD

About 10a Labs: 10a Labs is the safety and threat-intelligence layer trusted by frontier AI labs, AI unicorns, Fortune 10 companies, and leading global technology platforms. Our adversarial red teaming, model evaluations…

Full-timeRemote
10a LabsRemote

Est. 135,000 USD

About 10a Labs: 10a Labs is the safety and threat-intelligence layer trusted by frontier AI labs, AI unicorns, Fortune 10 companies, and leading global technology platforms. Our adversarial red teaming, model evaluations…

Full-timeRemote
10a LabsRemote

Est. 127,000 USD

About 10a Labs: 10a Labs is the safety and threat-intelligence layer trusted by frontier AI labs, AI unicorns, Fortune 10 companies, and leading global technology platforms. Our adversarial red teaming, model evaluations…

Full-timeRemote
10a LabsRemote

Est. 92,500 USD

About 10a Labs: 10a Labs is the safety and threat-intelligence layer trusted by frontier AI labs, AI unicorns, Fortune 10 companies, and leading global technology platforms. Our adversarial red teaming, model evaluations…

Full-timeRemote

Est. 120,000 USD

Orion Innovation is a premier, award-winning, global business and technology services firm. Orion delivers game-changing business transformation and product development rooted in digital strategy, experience design, and…

Full-timeRemote
10a LabsRemote

Est. 120,000 USD

About 10a Labs: 10a Labs is the safety and threat-intelligence layer trusted by frontier AI labs, AI unicorns, Fortune 10 companies, and leading global technology platforms. Our adversarial red teaming, model evaluations…

Full-timeRemote
10a LabsRemote

Est. 165,000 USD

About 10a Labs: 10a Labs is the safety and threat-intelligence layer trusted by frontier AI labs, AI unicorns, Fortune 10 companies, and leading global technology platforms. Our adversarial red teaming, model evaluations…

Full-timeRemote
AnthropicRemote

Est. 362,500 USD

About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of co…

Full-timeRemote
AnthropicRemote

Est. 205,600 GBP

About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of co…

Full-timeRemote
10a LabsRemote

Est. 60,000 GBP

About 10a Labs: 10a Labs is the safety and threat-intelligence layer trusted by frontier AI labs, AI unicorns, Fortune 10 companies, and leading global technology platforms. Our adversarial red teaming, model evaluations…

Full-timeRemote
AnthropicRemote

Est. 80,600 GBP

About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of co…

Full-timeRemote
MeridialRemote

Est. 134,400 USD

Are you a cybersecurity expert eager to shape the future of AI? Large‑scale language models are evolving from clever chatbots into powerful engines of digital defense. With high‑quality training data, tomorrow’s AI can d…

Full-timeRemote
Hyphen Connect LimitedBoston, Massachusetts, United States

Est. 165,000 USD

We are searching for an AI Safety Specialist who will play a crucial role in enhancing the security and robustness of language models. You will ensure the safe deployment of AI systems by conducting adversarial testing,…

Full-time
AnthropicSan Francisco, California, United States

Est. 620,000 USD

About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of co…

Full-time
10a LabsRemote

Est. 80,000 GBP

About 10a Labs: 10a Labs is the safety and threat-intelligence layer trusted by frontier AI labs, AI unicorns, Fortune 10 companies, and leading global technology platforms. Our adversarial red teaming, model evaluations…

Full-timeRemote
Hyphen Connect LimitedSan Francisco, California, United States

Est. 190,000 USD

We are searching for an AI Safety Specialist who will play a crucial role in enhancing the security and robustness of language models. You will ensure the safe deployment of AI systems by conducting adversarial testing,…

Full-time
AnthropicSan Francisco, California, United States

Est. 445,000 USD

About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of co…

Full-time
AnthropicRemote

Est. 80,600 GBP

About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of co…

Full-timeRemote
10a LabsSydney, New South Wales, Australia

About 10a Labs: 10a Labs is the safety and threat-intelligence layer trusted by frontier AI labs, AI unicorns, Fortune 10 companies, and leading global technology platforms. Our adversarial red teaming, model evaluations…

Full-time
RedditRemote

Est. 228,950 USD

Reddit is a community of communities. It’s built on shared interests, passion, and trust, and is home to the most open and authentic conversations on the internet. Every day, Reddit users submit, vote, and comment on the…

Full-timeRemote
AnthropicSan Francisco, California, United States

Est. 352,500 USD

About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of co…

Full-time
MeridialRemote

Are you an AI QA expert eager to shape the future of AI? Large-scale language models are evolving from clever chatbots into enterprise-grade platforms. With rigorous evaluation data, tomorrow’s AI can democratize world-c…

Full-timeRemote
Hyphen Connect LimitedSeattle, Washington, United States

Est. 165,000 USD

We are searching for an AI Safety Specialist who will play a crucial role in enhancing the security and robustness of language models. You will ensure the safe deployment of AI systems by conducting adversarial testing,…

Full-time
AnthropicSan Francisco, California, United States

Est. 402,500 USD

About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of co…

Full-time
AnthropicRemote

Est. 362,500 USD

About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of co…

Full-timeRemote
AnthropicRemote

Est. 80,600 GBP

About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of co…

Full-timeRemote
AnthropicSan Francisco, California, United States

Est. 425,000 USD

About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of co…

Full-time