AI攻防动态负责人研究员
AI Offense-Defense Dynamics Lead Researcher
职位简介
随着人工智能能力的迅速发展,我们面临一个根本性的知识空白:我们尚未完全理解决定人工智能系统,甚至其个别能力,主要对社会构成威胁还是保护的因素之间的复杂动态。在此职位中,您将领导研究以解码这些进攻与防御的动态关系,分析人工智能技术的具体属性如何影响其增强社会安全或加剧风险的倾向。您将运用跨学科方法,开发定量和定性框架,分析人工智能能力如何作为保护性或有害性应用在社会中扩散,为开发者、评估者、标准机构和政策制定者提供可操作的见解,以预见并减轻风险。该职位提供了独特的机会,塑造社会如何评估和管理日益强大的人工智能系统,直接对全球努力最大化人工智能利益的同时最小化风险产生影响。该职位为100%远程办公,但需要偶尔出差。
关于CARMA
AI风险管理和对齐中心(CARMA)致力于帮助社会应对日益强大的人工智能系统带来的复杂且可能造成灾难性后果的风险。我们的使命是具体降低从变革性人工智能中对人类和生物圈造成的风险。
我们专注于将人工智能风险管理建立在严谨分析的基础上,开发直面超级智能(AGI)的政策框架,推进技术安全方法,并促进持久安全的全球视角。通过这些互补的方法,CARMA旨在为社会提供关键支持,以在高级人工智能带来的巨大风险显现之前进行管理。
CARMA是Social & Environmental Entrepreneurs, Inc.(一家501(c)(3)非营利公共福利公司)财政资助的项目。
职责
- 开发量化系统动力学模型,捕捉技术、社会和制度因素之间相互关系,这些因素影响人工智能风险格局
- 设计详细的分析模型和模拟,识别政策干预的关键杠杆点,以改变进攻与防御平衡,实现更安全的结果
- 扩展并实施我们当前的进攻/防御动态分类法和初步框架,开发指标和模型,以预测特定人工智能系统特性是否有利于进攻性或防御性应用
- 构建基于实证的分析框架,使用文档
查看英文原文
Job Summary
As AI capabilities rapidly advance, we face a fundamental knowledge gap: we don't yet fully understand the complex dynamics that determine whether AI systems, or even individual capabilities of them, predominantly threaten or protect society. In this role, you'll lead research to decode these offense-defense dynamics, examining how specific attributes of AI technologies influence their propensity to either enhance societal safety or amplify risks. You'll apply interdisciplinary methods to develop quantitative and qualitative frameworks that analyze how AI capabilities proliferate through society as either protective or harmful applications, producing actionable insights for developers, evaluators, standards bodies, and policymakers to anticipate and mitigate risks. This position offers a unique opportunity to shape how society evaluates and governs increasingly powerful AI systems, with direct impact on global efforts to maximize AI's benefits while minimizing risks. This role is 100% remote but requires occasional travel.
About CARMA
The Center for AI Risk Management & Alignment (CARMA) works to help society navigate the complex and potentially catastrophic risks arising from increasingly powerful AI systems. Our mission is specifically to lower the risks to humanity and the biosphere from transformative AI.
We focus on grounding AI risk management in rigorous analysis, developing policy frameworks that squarely address AGI, advancing technical safety approaches, and fostering global perspectives on durable safety. Through these complementary approaches, CARMA aims to provide critical support to society for managing the outsized risks from advanced AI before they materialize.
CARMA is a fiscally-sponsored project of Social & Environmental Entrepreneurs, Inc., a 501(c)(3) nonprofit public benefit corporation.
Responsibilities
- Develop quantitative system dynamics models capturing the interrelationships between technological, social, and institutional factors that influence AI risk landscapes
- Design detailed analytical models and simulations to identify critical leverage points where policy interventions could shift offense-defense balances toward safer outcomes
- Expand and operationalize our current offense/defense dynamics taxonomy and nascent framework, developing metrics and models to predict whether specific AI system features favor offensive or defensive applications
- Build empirically-informed analytical frameworks using documented cases of AI misuse and beneficial deployed uses to validate theoretical models
- Research how specific technical characteristics (capabilities breadth/depth, accessibility, adaptability, etc.) interact with sociotechnical contexts to determine offense-defense balances
- Build public understanding of offense-defense dynamics through blog posts, articles, conference talks, and media engagement
- Create tools and methodologies to assess new AI models upon release for their likely offense-defense implications
- Draft evidence-based guidance for AI governance that accounts for complex interdependencies between technological capabilities and deployment contexts
- Translate research findings into actionable guidance for key stakeholders including policymakers, AI developers, security professionals, and standards organizations
Requirements
- A M.Sc. or higher in either Computer Science, Cybersecurity, Criminology, Security Studies, AI Policy, Risk Management, or a related field
- Demonstrated experience with complex systems modeling, risk assessment methodologies, or security analysis
- Strong understanding of dual-use technologies and the factors that influence whether capabilities favor offensive or defensive applications
- Deep understanding of modern AI systems, including large language models, multimodal models, and autonomous agents, with ability to analyze their technical architectures and capability profiles
- Experience in any of the following: Security mindset, Security studies research, Cybersecurity, Safety engineering, AI governance, Operational risk management, Systems dynamics modeling, Network theory, Complexity science, Adversarial analysis, or Technical standards development
- Ability to develop both qualitative frameworks and quantitative models that capture sociotechnical interactions, and comfort creating semi-quantitative semi-empirical models also grounded in logic
- Record of relevant publications or research contributions related to technology risk, governance, or security
- Exceptional analytical thinking with ability to identify non-obvious path dependencies and feedback loops in complex systems
Pluses
- PhD in a relevant field
- Experience with system dynamics modeling, hypergraph techniques, or other complex network analysis methods
- Skills in developing interactive tools or dashboards for risk visualization and communication
- Background in interdisciplinary research bridging technical and social science domains
- Demonstrated aptitude in top-down techniques and first-principles thinking
- Experience with the quantification of qualitative risk factors or developing proxy metrics for complex phenomena
- Background in compiling and analyzing incident databases or case studies for pattern recognition
- Familiarity with empirical approaches to technology assessment and impact prediction
- Knowledge of international relations theory as it applies to technology proliferation dynamics
CARMA/SEE is proud to be an Equal Opportunity Employer. We will not discriminate on the basis of race, ethnicity, sex, age, religion, gender reassignment, partnership status, maternity, or sexual orientation. We are, by policy and action, an inclusive organization and actively promote equal opportunities for all humans with the right mix of talent, knowledge, skills, attitude, and potential, so hiring is only based on individual merit for the job. Our organization operates through a fiscal sponsor whose infrastructure only supports persons authorized to work in the U.S. as employees. Candidates outside the U.S. would be engaged as independent contractors with project-focused responsibilities. Note that we are unable to sponsor visas at this time.
Originally posted on Himalayas