--- title: 'Research Lead at FAR.AI' canonical: 'https://feeny.ai/job/research-lead-far-ai-berkeley-rj833s1kwdxz' type: 'job' last_seen: '2026-09-16' --- # Research Lead at FAR.AI - **Company:** FAR.AI - **Location:** Berkeley, CA - **Compensation:** $170k–$270k - **Employment:** full-time - **Work type:** remote - **Posted:** 2026-04-24 - **Last confirmed live:** 2026-09-16 - **Apply:** https://jobs.ashbyhq.com/far.ai/a7b0dbea-e1ac-45c3-9e74-e400f65913c6/application **Skills:** Artificial Intelligence, Machine Learning, Statistical Analysis, Research Leadership, Team Management, AI Safety Research, Grant Writing, Publication Writing > The Research Lead will define and execute a research agenda to mitigate catastrophic risks from advanced AI, building and leading a team to focus on empirically grounded, scalable ML safety work. This role involves driving research direction, mentoring staff, and collaborating with governments and academia to ensure... ## Job description [FAR.AI](http://FAR.AI) is hiring a Research Lead to develop and lead a research agenda that reduces catastrophic risks from advanced AI. You'll build and lead a team executing this agenda — setting research direction, mentoring Members of Technical Staff to scale your vision, and remaining hands-on enough to write code and run experiments yourself. What counts is whether AI labs and governments actually change how they act; publications are useful but aren't the measure. Beyond your team, you can shape [FAR.AI](http://FAR.AI)'s broader work by directing millions of dollars in grants to external researchers extending your agenda, convening the people who can act on it, and influencing our independent testing and advising of AI companies and governments. This role suits you if you want high autonomy in an impact-driven environment, pursuing empirically grounded, scalable ML safety work. ## About Us [FAR.AI](http://FAR.AI) is a non-profit AI research institute working to ensure advanced AI is safe and beneficial for everyone. Our mission is to facilitate breakthrough AI safety research, advance global understanding of AI risks and solutions, and foster a coordinated global response. We’re structured to support that work from early research through real-world adoption: Independent by design. We can pursue what's most impactful based on our theory of change and share what we find publicly. A portfolio approach. Rather than focus on one single direction, we run diverse bets across the safety stack. We take promising ideas from initial experiments to deployment, informed by red-team partnerships with frontier labs and governments. Serious infrastructure for ambitious research. A dedicated engineering team runs our compute cluster and experiment-scaling stack, so researchers spend their time on research instead of on infra. Setting the standard. Our events convene key decision makers; our red-team works with frontier developers and governments; and our communications inform the public. Together, this drives adoption and sets the new standard in safety. Since our founding in July 2022, we've grown to [50+ staff](https://www.far.ai/about/team), published [40+ academic papers](https://scholar.google.com/citations?user=FVJ24k8AAAAJ), and convened leading [AI safety events](https://far.ai/events/). Our work is recognized globally, with publications at premier venues such as NeurIPS, ICML including a [Best Paper Honorable Mention in 2026](https://icml.cc/virtual/2026/oral/71065), and ICLR, and features in the [Financial Times](https://www.ft.com/content/175e5314-a7f7-4741-a786-273219f433a1), [Nature News](https://www.nature.com/articles/d41586-024-02218-7), [Wired Magazine](https://www.wired.com/story/jailbreaking-ai-models-google-anthropic-openai-spacexai/) and [MIT Technology Review](https://www.technologyreview.com/2020/02/28/905615/reinforcement-learning-adversarial-attack-gaming-ai-deepmind-alphazero-selfdriving-cars/). We conduct pre-deployment testing on behalf of frontier developers such as OpenAI and independent evaluations for governments [including the EU AI Office](https://www.far.ai/news/far-ai-selected-to-lead-eu-ai-act-cbrn-risk-consortium) and publish the [AI Security Leaderboard](https://leaderboard.far.ai/) based on our red-teaming expertise. We help steer and grow the AI safety field through [developing](https://arxiv.org/abs/2405.06624) [research](https://arxiv.org/abs/2506.20702) [roadmaps](https://www.researchgate.net/publication/396910034_Open_Technical_Problems_in_Open-Weight_AI_Model_Risk_Management) with renowned researchers such as Yoshua Bengio; running [FAR.Labs](https://www.far.ai/programs/far-labs), an AI safety-focused co-working space in Berkeley housing 40+ members; and supporting the community through [targeted grants](https://www.far.ai/programs/grantmaking) to technical researchers. ## About FAR.Research We explore promising research directions in AI safety and scale up only those showing a high potential for impact. Once the core research problems are solved, we work to scale them to a minimum viable prototype, demonstrating their validity to AI companies and governments to drive adoption. Our recent and ongoing research includes: Adversarial Robustness: working to rigorously solve security problems through building a science of security and robustness for AI, from [demonstrating superhuman systems can be vulnerable](https://far.ai/post/2023-07-superhuman-go-ais/), to [scaling laws for robustness](https://www.far.ai/news/does-robustness-improve-with-scale) and [jailbreaking constitutional classifiers](https://arxiv.org/abs/2506.24068). Mechanistic Interpretability: [finding](https://arxiv.org/abs/2502.12892) [issues](https://arxiv.org/abs/2508.16560) [with](https://arxiv.org/abs/2505.11756) Sparse Autoencoders, probing deception using [AmongUs](https://arxiv.org/abs/2504.04072), understanding [learned planning](https://far.ai/post/2024-07-learned-planners/) in SokoBan, and interpretable data attribution. Red-teaming: conducting pre- and post-release adversarial evaluations of frontier models (e.g. [Claude 4 Opus](https://x.com/ARGleave/status/1926138376509440433), [ChatGPT Agent](https://cdn.openai.com/pdf/839e66fc-602c-48bf-81d3-b21eacc3459d/chatgpt_agent_system_card.pdf), [GPT-5](https://cdn.openai.com/gpt-5-system-card.pdf)); developing [novel attacks](https://www.far.ai/news/defense-in-depth) to support this work. Evals: developing evaluations for new threat models, e.g. [persuasion](https://arxiv.org/abs/2506.02873) and [tampering risks](https://arxiv.org/abs/2507.11630). Mitigating AI deception: studying when [lie detectors induce honesty or evasion](https://www.far.ai/news/avoiding-ai-deception), and developing approaches to deception and sandbagging. We are particularly looking to add Research Leads in the following pod shapes: - Applied Interpretability — using interpretability to tackle concrete safety problems (better probes, backdoor detection, deception monitoring), aiming for fast feedback loops, often in collaboration with our other pods. A new pod, greenfield. - Scalable Oversight / Alignment — methods that keep oversight robust as models become more capable than their supervisors: recursive reward modeling, debate, weak-to-strong generalization, process-based supervision. - Adversarial Robustness —extending our independent-testing work into deployed-system protection: better safety guardrails, pre-training safety interventions (initially CBRN misuse, especially for open-weight models), backdoor detection and mitigation, realistic cybersecurity evaluations, and loss-of-control deception evaluations. - Auditing / Evals — safety and alignment auditing: evaluation awareness (construct validity, safety-relevance, hyper-realistic evals), CoT monitorability and faithfulness training, black-box monitoring as a complement to our existing white-box work. - Persuasion / Epistemic Risks — science of epistemic risks and intervention points, persuasion's role in loss of control risks, evaluations and independent testing, connections to broader harmful manipulation, solutions and epistemic uplift. Building on our existing work and shaping your own agenda in the area. - Bring Your Own Agenda — an open track for senior researchers with a strong vision outside the pods above. ## About the Role Research Leads define and own a research workstream end-to-end. Day-to-day, that means: - Articulate a research agenda with a clear theory of change for mitigating catastrophic risks from human-level or superhuman AI systems, and/or vastly increasing the upside of such systems. - Grow and lead a team of technical staff in pursuit of this agenda, either directly or in partnership with an engineering co-lead. - Lead novel research projects where there may be unclear markers of progress or success. - Share your research findings through written content (e.g. academic publications, blog posts) and presentations (e.g. ML conferences, policymaker briefings) to drive adoption and change. - Mentor and coach junior team members in research skills and ML engineering. - Contribute to the [FAR.AI](http://FAR.AI) intellectual environment, for example by giving feedback on early-stage proposals. - Build a research field around your agenda through [FAR.AI](http://FAR.AI)'s grantmaking and events, and connect it to real-world deployments through our independent testing and government advising. This role would be a great fit if you: - Want to work on the most impactful research directions, alongside mission-driven colleagues who'll push them forward with you. - Wish to pursue empirically grounded, scalable research directions that lean, technically strong teams can drive forward. - Value the ability to speak freely. We don't censor our researchers — we just ask that you protect confidential information and make clear when you're speaking personally or on behalf of the organization. - Want to advise and collaborate with governments, leading AI companies, and academics. We're a small organization that punches above its weight by working closely with these partners — through red-teaming, technical standards work, and research collaborations. This role would be a poor fit if you: - Prefer solo IC research to leading a team toward a shared agenda. Some people can do great research that way, but in this role we're looking for someone whose research direction is strong enough that other excellent researchers want to build it with them. - Prioritize novelty and intellectual elegance over impact. We care about both — a mathematically elegant solution to AI safety would be wonderful — but when we have to choose, we choose what makes AI safer in practice. - Can only work with the largest compute clusters available at industry labs or need to be compensated with equity in a rapidly growing startup. We offer competitive salaries and sizable compute budgets on a cluster that we manage, but if you value these things over having a positive impact on the future, then you may be more suited to a for-profit lab. ## About You To be a strong candidate for the Research Lead role, you likely: - Have a strong existing research track record in AI or another highly technical subject (e.g. CS, math, physics). - Have a clear view of which safety research directions are likely to matter most over the next few years, and why. - Have either (a) a clear research agenda you'd pursue at [FAR.AI](http://FAR.AI), with a theory of change explaining why it's valuable, or (b) a strong track record and a research space you'd sharpen into an agenda over your first months. We assess both paths against the same bar — depth of articulation at application is itself a signal about expected runway. - Have led a team, mentored graduate students, or supported early-career researchers through fellowship programs. Informal leadership in flatter organizations counts, as we’re more interested in experience than job titles. - Can effectively communicate novel methods and solutions to both technical and non-technical audiences. - Are not a new entrant to AI safety. We don't require a PhD or specific years of experience, but you should have engaged substantively with the field — through prior research, employment, or sustained independent contribution. It is preferable if you: - Have an established publication record in AI safety. - Are comfortable writing grant proposals and navigating collaborations with other organizations or external research groups. If you are missing key leadership experience or are earlier in your career, we encourage you to consider the open [Research Scientist](https://far.ai/careers/research-scientist-39dfd?ashby_jid=1bda4204-bfef-4a47-b72b-3562ec0bb3f9) pathway and invite you to contribute to one of our existing agendas. We're also open to more senior versions of this role; simply apply or reach out to talent@far.ai. ## Benefits* - 🩺 Health Insurance - 94% of Insurance premium paid by Organization commencing within 1 month after your start date - 💰 Retirement - 401(k) plan with up to 2% match - 🏝️ PTO - 25 days Paid Time Off per year, accrued weekly and up to 10 days of paid sick leave per year - 🚼 Paid Leave - Paid Bereavement, Family, Medical and Pregnancy Disability Leave - 🖥️ WFH Stipend & Equipment - Work computer and stipend provided for eligible employees - 🍽️ Catered Meals (Berkeley Office Only) - Catered lunches and dinners on workdays at our office *(Available only to full-time employees located in the US) Logistics If based in the USA or Singapore, you will be an employee of [FAR.AI](http://FAR.AI) (501(c)(3) research non-profit / non-profit CLG). Outside the USA or Singapore, you will be employed via an EOR organisation on behalf of [FAR.AI](http://FAR.AI) or as a contractor. - Location: Both remote and in-person (Berkeley, CA or Singapore) are possible. We sponsor visas for in-person employees, and can hire remotely in most countries. - Hours: Full-time (40 hours/week). - Application materials: Expect ~1–2 hours of preparation; most carries forward from prior job searches. We ask for a CV, a short research direction statement (the form supports both fully-formed agendas and developing ones), 2–3 selected works with a brief note on your personal contribution, and a short note on why [FAR.AI](http://FAR.AI) is a good home for your direction. If you advance to portfolio review, we'll ask for a full research direction statement (1–2 pages, with a theory of change to real-world implementation; ~1.5–2 hours, due within about a week). - Process: From application: a portfolio review (async), a 60-minute bilateral fit call, a research deep-day (~3.5 hours live, including an open talk to FAR research staff and two interview sessions), a 5-day paid work trial, structured reference calls, and a final decision panel. Typical elapsed time: 4–6 weeks. Total candidate time end-to-end is ~50 hours, with the paid work trial being the bulk. If a 5-day block isn't feasible for you, reach out — we can discuss alternatives. If you have any questions about the role, please do get in touch at talent@far.ai. If you have any questions about the role, feel free to contact us at talent@far.ai. Otherwise, if you don't have questions, the best way to ensure a proper review of your skills and qualifications is by applying directly via the application form. Please don't email us to share your resume (it won't have any impact on our decision). Thank you! ## About FAR.AI ## Company Overview - **One-liner**: FAR.AI is a technical AI safety research non-profit dedicated to ensuring advanced AI systems are safe and beneficial for everyone through in-house research, grantmaking, and global coordination events. - **Entity Type**: Private (Non-profit, fiscally sponsored project) - **Headquarters**: Berkeley, California, USA - **Founded**: July 2022 (incorporated October 2022) - **Founders**: Adam Gleave (CEO) and Karl Berzins (President) ## Core Business - **Primary industry**: AI Safety Research, Research Services - **Target customers**: Policymakers, industry leaders, academic researchers, and the broader AI safety ecosystem (primarily B2B/Institutional) - **Mission**: To ensure advanced AI is safe and beneficial for everyone. The organization is motivated by the potential risks posed by rapid advances in AI capabilities. ## Products & Services - **FAR.Research**: In-house technical research team exploring early-stage agendas for AI safety, including adversarial robustness, AI control, and scaling issues. Publishes influential papers and open-source tools. - **FAR.Labs**: A collaborative co-working space in Berkeley for AI safety researchers and organizations, hosting over 40 active members and fostering a thriving community. - **FAR.Futures (formerly FAR.Grants)**: A targeted grantmaking program supporting academics and independent researchers developing innovative solutions to critical AI risks. Has directed millions of dollars in funding. - **Events & Conferences**: Organizes high-impact events including the Alignment Workshop series (global), Berkeley ControlConf 2026 (AI control), and the Technical Innovations for AI Policy (TIAP) Conference, connecting policymakers with leading AI technical experts. ## Market Standing - **Valuation/Market Cap**: Not applicable (non-profit) - **Key Metric**: Total Funding — Not disclosed (non-profit, fiscally sponsored). Key output: 30+ research publications, 1000+ attendees across 10+ events. - **Notable Investors/Partners**: Collaborations with leading think tanks, academic groups (e.g., CHAI, MATS, Apart Research), and major media outlets (Nature, MIT Technology Review, Financial Times, NYT). Fiscally sponsored project of IDAIS. - **Growth Signals**: Headcount grew 53.1% YoY (from ~17 to 34 employees as of mid-2026). Expanded to 8 countries (US, UK, Germany, Spain, Switzerland, Singapore, Mexico, Australia). LinkedIn followers grew 171.4% yearly. Research cited in US Senate hearings (Stuart Russell testimony). ## Competitive Advantages - **Technical Breakthroughs**: Published influential work on adversarial attacks on superhuman Go AIs (featured in Nature), multi-agent adversarial policies, and red teaming leading language models for frontier labs. - **Ecosystem Centrality**: Runs the only dedicated AI safety coworking space (FAR.Labs) and a premier grantmaking program, positioning FAR.AI as a hub connecting academia, industry, and policy. - **Policy Influence**: TIAP Conference directly connects policymakers with technical experts; research cited in congressional testimony. - **Top Talent Magnet**: Attracts researchers from top AI labs (Anthropic, Google DeepMind) and academic safety hubs (Cambridge AI Safety Hub, CHAI). ## Strategic Focus - **Field Building**: Scaling the global AI safety ecosystem through grants, events, and coworking space. - **Technical Innovation**: Continuing to explore early-stage, high-impact research agendas (e.g., AI control, robustness, adversarial training) until they can be adopted by the broader community. - **Policy Engagement**: Deepening connections between technical experts and policymakers to ensure safety techniques are adopted. - **Talent Development**: Expanding the team with world-class researchers, engineers, and operations staff to tackle critical AI safety challenges. ## Why Work Here - **Mission-Driven Culture**: Employees report high satisfaction (5.0/5.0 on Culture, 5.0/5.0 on Work-Life on LinkedIn). The mission to make advanced AI safe is described as "one of the most critical challenges of our time." - **Work Environment**: Hybrid/remote-first with a physical HQ in Berkeley. Offers both remote and onsite roles. The Berkeley office (FAR.Labs) fosters a collaborative, startup-like atmosphere. - **Team Composition**: Small, high-agency team of ~34 people with a flat structure. Heavy emphasis on research (18% of staff) and project management (12%). Senior leadership includes former Anthropic and DeepMind researchers. - **Compensation & Perks**: Rated 4.0/5.0 on Compensation. As a non-profit, likely offers competitive (but not market-maxing) salaries with strong mission alignment. Perks include the coworking space, events, and a highly collaborative environment. - **Career Growth**: High growth trajectory (53% YoY headcount increase) means rapid advancement opportunities. Employees often move to top AI labs (Anthropic, DeepMind, Stability AI) after their tenure. - **Notable**: Currently hiring for Jailbreaking Lead (Red Team), Executive Operations Assistant, and Technical Project Manager (Red Team). ## Sources 1. [far.ai](https://www.far.ai/) 2. [far.ai/about](https://www.far.ai/about) 3. [far.ai/careers](https://www.far.ai/careers) 4. [linkedin.com/company/far-ai](https://www.linkedin.com/company/far-ai) 5. [builtin.com/company/farai](https://builtin.com/company/farai) ## Other roles at FAR.AI - [Tech Lead Manager, GPU Cluster Infrastructure](https://feeny.ai/job/tech-lead-manager-gpu-cluster-infrastructure-far-ai-global-qcbffm40ehej) — Global - [Senior Software Engineer, GPU Cluster Infrastructure](https://feeny.ai/job/senior-software-engineer-gpu-cluster-infrastructure-far-ai-global-ps8k924d9yph) — Global - [IT & Security Manager](https://feeny.ai/job/it-security-manager-far-ai-united-states-gn70wqjc2gbh) — United States - [Research Lead - Pre-training Safety](https://feeny.ai/job/research-lead-pre-training-safety-far-ai-berkeley-z2xbmrwva3e6) — Berkeley, CA - [Engineering Manager (Red Team)](https://feeny.ai/job/engineering-manager-red-team-far-ai-global-0rrs6jb6d92f) — Global - [Technical Program Manager, Research](https://feeny.ai/job/technical-program-manager-research-far-ai-united-states-xwtezz6qryxw) — United States - [General Counsel](https://feeny.ai/job/general-counsel-far-ai-united-states-z6d5d2mpgz2q) — United States - [Head of People](https://feeny.ai/job/head-of-people-far-ai-united-states-w5w1pw55j7de) — United States - [Senior Programs & Strategy Manager](https://feeny.ai/job/senior-programs-strategy-manager-far-ai-united-states-3gjmjqaj5fcy) — United States - [Senior Project Manager, Events](https://feeny.ai/job/senior-project-manager-events-far-ai-united-states-zqdbs22ykftw) — United States