Engineering Manager (Red Team)

far.ai · Remote (International) · $170k–$250k

Posted about 16 hours ago

Apply in seconds with Jobply
One-click apply on Workday, Greenhouse, Lever & 50+ ATS systems
Apply with Jobply →

Skills

PythonLLM APIsLocal LLM frameworksEvaluation frameworksLLM agentsCloud infrastructureKubernetesSlurmCoding agentsAgentic red-teaming systems

Benefits

  • Catered lunch and dinner
  • Work-related travel
  • Equipment expenses

Job description

About Us

FAR.AI is a non-profit AI research institute dedicated to ensuring advanced AI is safe and beneficial for everyone. Our mission is to facilitate breakthrough AI safety research, advance global understanding of AI risks and solutions, and foster a coordinated global response.

Since our founding in July 2022, we've grown quickly to 50+ staff, producing over 40 influential academic papers, and establishing leading AI Safety events. Our work is recognized globally, with publications at premier venues such as NeurIPS, ICML, and ICLR, and features in WIRED, Financial Times, Guardian, Washington Post, and MIT Technology Review. Additionally, we help steer and grow the AI safety field through developing research roadmaps with renowned researchers such as Yoshua Bengio; running FAR.Labs, an AI safety-focused co-working space in Berkeley housing 40 members; and supporting the community through targeted grants to technical researchers.

About Red-Teaming at FAR.AI

FAR.AI’s red team is building toward a simple outcome: materially raising the bar for safety and security of the most widely deployed and capable AI systems in the world. We intend to be the tip of the spear in AI safety: the team that consistently finds the failures others miss, resulting in real mitigations, and setting the standard that labs and governments converge on. We also leverage our in-depth understanding of weaknesses in frontier models to advise frontier developers on mitigations, to guide our own research and grantmaking for improving model security, and to inform the public of key AI risks.

We are already one of the leading independent red-teaming organizations. Our work has helped most Western frontier model developers improve safeguards through pre- and post-deployment testing (e.g., we have directly influenced safeguards at major frontier developers like OpenAI and Anthropic), and are increasingly embedded in high-leverage government efforts (e.g., leading a consortium building CBRN evaluations for the European Commission/EU AI Office, and collaborating with the UK AI Security Institute).

"FAR.AI's pre-deployment testing of GPT-5 series models identified failure modes and mitigations, improving the security of our model releases." – Senior Technical Program Manager, OpenAI

FAR.AI have been a trusted and thoughtful collaborator for us, and they have progressed the state of frontier red-teaming through research like STACK. We expect this to be a high impact role and are excited to explore collaborations with the successful candidate.” – Xander Davies, Technical Lead, Red Team at UK AISI

In 2026, we are scaling from a strong team with standout wins into a new level of impact for any AI red team globally:

  • Red-teaming all major frontier model releases (closed and open-weight) within days/weeks of release;

  • Expanding strategic engagements with governments and conducting pre-deployment testing with most frontier labs;

  • Deepening our testing of key risk areas like CBRN, cyber, and agents, and exploring new ones like AI control and alignment;

  • Building tools, agents, and insights that raise the global standard for red-teaming.

About the Role

As engineering manager on the FAR.AI Red Team, you will be the senior technical owner of our engineering, reporting to Kellin Pelrine with a dotted line to Edward Yee. You will build the engineering team and the systems to test critical safeguards of the next generation of AI models. Success looks like vulnerabilities fixed, safeguards strengthened, and global standards shifted – working with leading frontier labs and governments globally to make this happen. Systems your team builds will enable our high-stakes engagements, accelerate cutting-edge AI security research, and create the tools/products/services that let us and our partners red-team the frontier continuously and comprehensively. The current red team will expand in size significantly, and your engineering team will lay the scalable foundation for a relentless and high velocity division to make advanced AI systems safer.

We expect this role to start with a small team and include both substantial hands-on engineering and substantial teambuilding and management, e.g., 40% hands-on 60% management. The balance will increasingly evolve towards the management side as we scale the team. In practice, this role spans:

  • Building a scalable red-teaming engine: tooling, products, services, evals, agents, and a self-improving workflow that multiplies output:

    • Build and scale internal/external red-teaming tools and products that improve speed, coverage, and severity of findings;

    • Develop agentic red-teaming systems to automatically and systematically explore the attack space;

    • Develop agentic systems to identify new public jailbreaks and other releases, and automatically integrate them into our systems;

    • Advance state-of-the-art attacker simulation and output harmfulness evaluation;

    • Maintain and evolve internal vulnerability databases, statistical analysis tools, and reporting infrastructure.

  • Build systems to empower high-stakes red-teaming engagements with frontier AI companies and governments:

    • Support public reports, benchmarks, and leaderboards that shape industry norms

    • Contribute to red-teaming of frontier models (closed- and open-weight), improving testing systems both within particular engagements and turning insights from engagements into advances in our overall systems;

    • Build agentic workflows that accelerate the red team broadly.

  • Managing, mentoring, and supporting a growing elite team of red-teaming ICs with extremely high standards for velocity, rigor, judgment, and impact;

    • Plan sprints and lead engineering execution;

    • Mentor red team ICs in engineering skills, both in direct 1-1 settings and by building resources for rapid skill growth and knowledge transfer;

    • Manage ICs, supporting and amplifying their impact, setting high standards and prioritization week-to-week, and creating the environment and paths for rapid growth;

    • Create processes that scale high-tempo engagements without sacrificing quality;

    • Design and be hiring manager for pipelines to scale the engineering team, in partnership with others on the red team and the org’s recruiting team.

  • Contribute to the overall technical and product strategy of the red team;

  • Partner closely with the rest of the division to translate technical ideas and findings into real-world impact.

This role would be a great fit if you:

  • Are excited by high-stakes, real-world technical work where success is measured by impact, not revenue, clicks, or papers published;

  • Want to work with a range of governments, leading AI companies, and academics. We’re a lean organization, and seek to leverage our impact through strategic partnerships;

  • Get energy from working at the frontier of AI and shaping its path forward;

  • Enjoy combining hands-on technical work with managing, mentoring, and architecting systems on an elite high-impact and high-performing team;

  • Are excited to build the team from the hiring side, identifying the skillsets and roles we need, building new pipelines and assessments for them where existing ones fall short, and working with our recruiters and others on the broader team to execute;

  • Are a senior leader who wants to spend some time building a new team and working close to code again, potentially as you transition to a new area;

  • Are a junior leader who wants to continue growing in an entrepreneurial mindset team;

  • Are excited to leverage AI agents (e.g., coding agents) at the very frontier of the field, continually building your skills and judgment to achieve superhuman results;

  • Enjoy working with mission-driven colleagues to achieve real-world impact;

  • Care deeply about AI safety and impacting how advanced AI systems are deployed;

  • Are ready to get shit done and willing to do whatever it takes to change the world;

  • Thrive in fast-moving, ambiguous environments with shifting threat models and defenses.

This role would be a poor fit if you:

  • Do not have leadership experience – a Senior Engineer or similar position on our team may be a better fit;

  • Want to work only at the management, strategy, or architecture levels, without hands-on coding and engineering. The role will evolve over time to increasingly focus on management as the team scales, but for the first 6 months candidates are expected to also dive deep hands-on;

  • Want to start with a large team day 1, or for others to manage all hiring. Part of the role is building the team, by identifying and filling the gaps between our existing team members and recruiting pipelines and the new capacity and skills the team needs;

  • Want to set technical direction unilaterally. You will own engineering execution and week-to-week sprint planning, and have substantial input on broader technical direction, working to chart it collaboratively with other team leads and division leadership;

  • Want to prioritize open-ended research, academic metrics of success like papers published, or other interests not directly connected to red-teaming success – our Senior Research Engineer position in the Research division may be a better fit;

  • Are primarily motivated by equity upside or compensation;

  • Are not willing to move at the velocity we need / want a slow-moving, highly structured environment with fixed problem definitions. As an organization, we’re planning to double in size in the next 12-18 months, and the red team will grow even faster. A lot will change, so navigating uncertainty is a core part of the role to capitalize on all the opportunities we will encounter.

About You

Strong candidates for this role typically have many (but not necessarily all) of the following:

  • Have a strong track record in software engineering, AI, or another computer engineering discipline (e.g. cybersecurity, MLOps);

  • Have a strong track record of managing, growing, and leading technical teams;

  • Have experience with parts of the tech stack we commonly use. For example, coding agents (e.g., Claude Code, Codex, Cursor), Python, LLM APIs (e.g. OpenAI, Anthropic, Google), local LLM frameworks (e.g., vLLM), evaluation frameworks (e.g., Inspect), LLM agents (e.g., OpenClaw), cloud infrastructure (e.g. GCP), compute clusters (e.g. Kubernetes, Slurm), etc.

  • Have thrived in rapidly evolving environments;

  • Demonstrated drive for mission/impact and desire to create a real impact on frontier AI systems;

  • Can effectively communicate technical solutions to both technical and non-technical audiences;

  • Demonstrated relentlessness in achieving ambitious goals.

It is a plus (but not required) if you have:

  • Experience building or red-teaming frontier LLMs or agentic systems;

  • Experience building technical teams and products in an entrepreneurial environment;

  • Demonstrated ability to discover non-obvious, high-severity vulnerabilities in complex systems;

  • Other hands-on experience in closely related areas like adversarial ML or security;

  • Prior collaboration with AI labs, security teams, or government safety institutes;

  • Published work in AI safety, security, or robustness.

If you are earlier in your career, we encourage you to consider our other roles within the red team. We are likewise also open to more senior versions of this role, simply apply or reach out to [email protected].

Logistics

If based in the USA or Singapore, you will be an employee of FAR.AI (501(c)(3) research non-profit / non-profit CLG). Outside the USA or Singapore, you will be employed via an EOR organization on behalf of FAR.AI or as a contractor.

  • Location: Remote globally. We can sponsor US or Singapore visas. The current red team is based in UTC+1, UTC+8, and UTC-4, so some timezone flexibility is expected.

  • Hours: Full-time (40 hours/week).

  • Compensation: $170,000-$250,000/year depending on experience and location, with the potential for additional compensation for exceptional candidates. We will also pay for work-related travel and equipment expenses. We offer catered lunch and dinner at our offices in Berkeley.

  • Application process: Apply by submitting a CV and answering our questions around your interest. From there: a short screening call, work tests, interviews, and finally a paid work trial. If you are not available for a work trial we may be able to find alternative ways of testing your fit.

If you're uncertain whether your background fits but are excited by the mission and challenges, we encourage you to apply – we're looking for excellence and potential, not a perfect resume match.

If you have any questions about the role, please get in touch at [email protected].

Otherwise, if you don't have questions, the best way to ensure a proper review of your skills and qualifications is by applying directly via the application form. Please don't email us to share your resume (it won't have any impact on our decision). Thank you!

If you have any questions about the role, feel free to contact us at [email protected]. Otherwise, if you don't have questions, the best way to ensure a proper review of your skills and qualifications is by applying directly via the application form. Please don't email us to share your resume (it won't have any impact on our decision). Thank you!

Stop filling out the same form 100 times.

Install the free Jobply Chrome extension and auto-apply to Engineering Manager (Red Team) and 300,000+ other live jobs across Workday, Greenhouse, Lever, and 50+ other ATS systems.

Apply with Jobply — Free
✓ Free forever✓ No credit card✓ 4.8★ from 12k+ users