Description
About 10a Labs: 10a Labs is the safety and threat-intelligence layer trusted by frontier AI labs, AI unicorns, Fortune 10 companies, and leading global technology platforms. Our adversarial red teaming, model evaluations, and intelligence collection enable engineering, safety, and security teams to stay ahead of evolving threats and deploy AI systems safely.
About the role: As a Red Teaming Fellow, you will help evaluate the safety, security, and reliability of advanced AI systems. Fellows work alongside researchers, engineers, and subject-matter experts to identify vulnerabilities, uncover failure modes, and generate insights that help organizations deploy AI systems more safely.
This is a hands-on fellowship at the intersection of AI, security, and adversarial testing. Fellows will contribute to real-world evaluations of frontier models and AI-enabled systems, helping design and execute tests that probe how models behave under challenging or unexpected conditions.
In this role, you will:
- Conduct adversarial testing of AI systems across safety, security, and misuse scenarios
- Design and execute red teaming exercises against models, agents, and AI-enabled applications
- Develop prompts, attack strategies, and test cases to identify vulnerabilities and failure modes
- Analyze model behavior and document findings in clear, actionable reports
- Support the development of evaluation frameworks, taxonomies, and testing methodologies
- Research emerging AI threats, attack techniques, and risk trends
- Collaborate with engineers, researchers, and subject-matter experts to investigate emerging risks and threat vectors
- Develop front-end dashboards and other visualizations
Qualifications:
- Pursuing or recently completed a degree in Computer Science, Cybersecurity, Engineering, Data Science, Mathematics, Political Science, International Relations, or a related field
- Strong interest in AI safety, cybersecurity, red teaming, or adversarial testing
- Familiarity with AI abuse including jailbreaking, system prompt extraction, indirect prompt injection, data exfiltration, etc.
- Strong analytical, research, problem-solving, and communication skills
- Ability to think creatively from an adversarial perspective
- Ability to work independently and collaboratively in a remote environment
- Experience with Python, Bash, large language models, AI systems, cybersecurity, or research methodologies preferred
- Proficiency in a language other than English preferred
Benefits:
- Flexible start / end dates
- Remote work (based in the continental U.S.)
- Flexible schedule, up to 20 hours per week (negotiable)
- Hourly pay commensurate with experience and qualifications
- $25 per hour for undergraduate students
- $32.50 per hour for graduate students
Similar jobs
Est. 110,000 USD
About 10a Labs: 10a Labs is the safety and threat-intelligence layer trusted by frontier AI labs, AI unicorns, Fortune 10 companies, and leading global technology platforms. Our adversarial red teaming, model evaluations…
Est. 60,000 USD
About 10a Labs: 10a Labs is the safety and threat-intelligence layer trusted by frontier AI labs, AI unicorns, Fortune 10 companies, and leading global technology platforms. Our adversarial red teaming, model evaluations…
Est. 52,000 USD
About 10a Labs: 10a Labs is the safety and threat-intelligence layer trusted by frontier AI labs, AI unicorns, Fortune 10 companies, and leading global technology platforms. Our adversarial red teaming, model evaluations…
Est. 147,500 USD
About 10a Labs: 10a Labs is the safety and threat-intelligence layer trusted by frontier AI labs, AI unicorns, Fortune 10 companies, and leading global technology platforms. Our adversarial red teaming, model evaluations…
Est. 135,000 USD
About 10a Labs: 10a Labs is the safety and threat-intelligence layer trusted by frontier AI labs, AI unicorns, Fortune 10 companies, and leading global technology platforms. Our adversarial red teaming, model evaluations…
Est. 127,000 USD
About 10a Labs: 10a Labs is the safety and threat-intelligence layer trusted by frontier AI labs, AI unicorns, Fortune 10 companies, and leading global technology platforms. Our adversarial red teaming, model evaluations…
Est. 92,500 USD
About 10a Labs: 10a Labs is the safety and threat-intelligence layer trusted by frontier AI labs, AI unicorns, Fortune 10 companies, and leading global technology platforms. Our adversarial red teaming, model evaluations…
Est. 120,000 USD
Orion Innovation is a premier, award-winning, global business and technology services firm. Orion delivers game-changing business transformation and product development rooted in digital strategy, experience design, and…
Est. 120,000 USD
About 10a Labs: 10a Labs is the safety and threat-intelligence layer trusted by frontier AI labs, AI unicorns, Fortune 10 companies, and leading global technology platforms. Our adversarial red teaming, model evaluations…
Est. 165,000 USD
About 10a Labs: 10a Labs is the safety and threat-intelligence layer trusted by frontier AI labs, AI unicorns, Fortune 10 companies, and leading global technology platforms. Our adversarial red teaming, model evaluations…
Est. 362,500 USD
About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of co…
Est. 205,600 GBP
About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of co…
Est. 60,000 GBP
About 10a Labs: 10a Labs is the safety and threat-intelligence layer trusted by frontier AI labs, AI unicorns, Fortune 10 companies, and leading global technology platforms. Our adversarial red teaming, model evaluations…
Est. 80,600 GBP
About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of co…
Est. 134,400 USD
Are you a cybersecurity expert eager to shape the future of AI? Large‑scale language models are evolving from clever chatbots into powerful engines of digital defense. With high‑quality training data, tomorrow’s AI can d…
Est. 165,000 USD
We are searching for an AI Safety Specialist who will play a crucial role in enhancing the security and robustness of language models. You will ensure the safe deployment of AI systems by conducting adversarial testing,…
Est. 620,000 USD
About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of co…
Est. 80,000 GBP
About 10a Labs: 10a Labs is the safety and threat-intelligence layer trusted by frontier AI labs, AI unicorns, Fortune 10 companies, and leading global technology platforms. Our adversarial red teaming, model evaluations…
Est. 190,000 USD
We are searching for an AI Safety Specialist who will play a crucial role in enhancing the security and robustness of language models. You will ensure the safe deployment of AI systems by conducting adversarial testing,…
Est. 445,000 USD
About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of co…
Est. 80,600 GBP
About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of co…
About 10a Labs: 10a Labs is the safety and threat-intelligence layer trusted by frontier AI labs, AI unicorns, Fortune 10 companies, and leading global technology platforms. Our adversarial red teaming, model evaluations…
Est. 228,950 USD
Reddit is a community of communities. It’s built on shared interests, passion, and trust, and is home to the most open and authentic conversations on the internet. Every day, Reddit users submit, vote, and comment on the…
Est. 352,500 USD
About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of co…
Are you an AI QA expert eager to shape the future of AI? Large-scale language models are evolving from clever chatbots into enterprise-grade platforms. With rigorous evaluation data, tomorrow’s AI can democratize world-c…
Est. 165,000 USD
We are searching for an AI Safety Specialist who will play a crucial role in enhancing the security and robustness of language models. You will ensure the safe deployment of AI systems by conducting adversarial testing,…
Est. 402,500 USD
About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of co…
Est. 362,500 USD
About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of co…
Est. 80,600 GBP
About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of co…
Est. 425,000 USD
About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of co…