BlueDot Impact's AI governance discussion group. Regular virtual meetings for community members to discuss AI policy, regulation, and governance challenges. Part of BlueDot's governance education programming.
Past events
Sorted by date, newest first. Useful as a memory of what's happened in the community.
BlueDot Impact's weekly AI safety evaluations paper reading group. Community-focused discussion of recent research on AI safety evaluation methods and techniques. Part of BlueDot's ongoing educational programming for the AI safety community.
United Nations Institute for Disarmament Research conference addressing critical space security challenges. International AI policy and disarmament-related conferences in scope per the safety community's broader governance interests. Registration open. UNIDIR hosts approximately 20K+ event participants annually.
Hackathon focused on translating AI safety research and concepts into accessible content and communications. Organized by London Interface for AI Safety. Participants work on making technical AI safety work more accessible to broader audiences.
BlueDot Impact's 4th anniversary celebration bringing together the AI safety community in San Francisco. Networking event for BlueDot's 8,000+ alumni network working at organizations like OpenAI, Anthropic, and Google DeepMind.
Grant fund awarding at least $200,000 in grants and prizes for corrigibility research in 2026. Typical awards range from $5,000-$35,000. Two rounds: Round 1 deadline August 23, Round 2 deadline October 31. Prize awards of $40,000 by end of September and at least $60,000 mid-December (no application required for prizes). Research spanning pure theory (formal models, decision-theoretic analysis) to empirical work (training experiments, evaluations) qualifies. Work must be legible to frontier AI labs and avoid accelerating capabilities development.
Regional EA Global conference with heavy AI safety attendance. Berkeley location draws alignment researchers, interpretability practitioners, and governance community. In scope for social and community networking reasons.
EAGx (Effective Altruism Global x) regional conference in Berkeley. Part of the EA Global conference series organized by Centre for Effective Altruism. Heavy AI safety attendance and programming as Berkeley is a major hub for alignment research.
16-day virtual incubator for AI safety researchers to develop clearer articulation of research questions and directions. Structure includes three weekend sessions with lightning talks, one-on-ones, and feedback cycles. Participants work toward either a well-scoped research direction for potential pursuit at next AI Safety Camp or refined set of research questions. Focus on honest inquiry and examining assumptions in AI safety discourse. At least 10 hours per week commitment.
Apart Research sprint focused on designing and running experiments on AI preferences and welfare. Participants examine how frontier models express values and internal states, exploring digital sentience and AI consciousness measurement. Hybrid format allows both online and in-person participation.
3rd New England Mechanistic Interpretability workshop bringing together academic and industry researchers. Topics include interpretability of neural circuits, probe-based analysis, feature attribution methods, model simplification, actionable interp, and interp for science. Four keynote speakers (Tom McGrath, Natalie Shapira, Belinda Li, Stephen Casper), student oral presentations, interactive poster sessions, panel discussion. Encourages submissions from rising researchers at New England universities.
Weekly paper reading club focused on AI safety evaluations research. Organized by BlueDot Impact with Andreas Turanski and Mmachukwu Osisioma. Recurring sessions every Tuesday at 4:00 PM UTC to discuss recent papers in the evals space. Community-driven discussion format accessible globally via Zoom.
3-week ML upskilling bootcamp for AI safety focusing on interpretability and RL, based on ARENA curriculum. Run by Cambridge Boston Alignment Initiative. Includes housing and meals, 24/7 office access in Harvard Square, dedicated teaching assistants, and travel support. Prerequisites: Python familiarity, multivariable calculus, and linear algebra.
3-week intensive program in applied mathematics for AI alignment by Iliad. Based in Berkeley at Lighthaven. Provides $5,000 support. Selection based on estimated mathematical strength. Part of Iliad's series of intensive and fellowship programs.
9th annual Cognitive Computational Neuroscience conference bringing together researchers in cognitive science, neuroscience, and AI. Topics include brain information processing, AI representational competencies, ML applications to brain modeling, DNN interpretability, perception/cognition modeling. Keynote speakers from UC Berkeley, Princeton, University of Alberta, OIST, and Stanford. Single-track format with keynotes, oral presentations, and poster sessions.
5-day, multi-track unconference with 100+ participants focused on theoretical AI alignment. Unconference format where attendees propose and lead sessions. Research areas include Singular Learning Theory, Agent Foundations, Causal Incentives, Computational Mechanics, and Safety-by-Debate. Free attendance, limited on-site bedrooms available, financial assistance for travel/accommodations on needs-basis.
5-day multi-track unconference with 100+ researchers focused on theoretical AI alignment. Participants propose and lead own sessions. Research areas: Singular Learning Theory, Agent Foundations, Causal Incentives, Computational Mechanics, Safety-by-Debate, Scalable Oversight. Free to attend, limited financial support for travel/accommodation available. Third ILIAD conference.
Summit bringing together academic leaders, entrepreneurs, AI experts, venture capitalists, and policymakers to discuss the future of AI and Agentic AI. Call for Papers and Startup Spotlight applications open. In-person and livestream available.
BlueDot Impact launch event in San Francisco bringing together the AI safety community. Part of BlueDot's expansion to build the workforce needed to safely navigate AGI. Community gathering for networking and connection among local AI safety practitioners, researchers, and fellows.
Part 2 of the Breaking Barriers to AI Safety series, featuring a career fair connecting AI safety job seekers with organizations, plus a builders summit for technical collaboration. Organized by Jen Ba and Robin Goins (Mox) through BlueDot Impact.
Apart Research hackathon focused on AI safety research. Part of the 55+ sprints series with 6,000+ participants across 200+ global locations.
Weekly AI safety evaluations paper reading club organized by BlueDot Impact. Virtual session on Zoom. Part of BlueDot's regular community programming building the workforce needed to safely navigate AGI.
Part 1 of a two-part AI safety event series organized by BlueDot Impact. This hackathon focuses on building AI safety tools and breaking barriers to entry in the field. Followed by Part 2 (Career Fair and Builders Summit) on July 24. Organized by Joshua Landes and Jen Ba.
Technical workshop bringing together top talent to address bottlenecks in secure AI advancement. Organized by Foresight Institute, a 40-year-old organization focused on transformative technology. Registration available.
A community event organized by Singapore AI Safety Hub (SASH) focused on getting into AI safety. Provides networking and information for those interested in entering the AI safety field.
Second Workshop on Agents in the Wild focusing on safety and security of AI agents deployed in real-world environments. Addresses challenges in ensuring safe and secure operation of autonomous agents. Part of ICML 2026 workshop track.
The workshop addresses how to integrate diverse perspectives, values, and expertise into pluralistic AI alignment frameworks. Examines multi-objective alignment approaches, preference elicitation methods, and human-AI interaction workflows that reflect pluralistic values across diverse communities.
Workshop on pluralistic AI alignment addressing diversity of human values. Over 95 accepted papers spanning philosophy, ML, HCI, social sciences, policy, and applications. Topics include value pluralism frameworks, annotation disagreement methods, pluralistic evaluation metrics, human-AI interaction design, consensus mechanisms, and deployment policies. Four keynote speakers (Alice Oh, Xiaoyuan Yi, Mitchell Gordon, Atoosa Kasirzadeh), panel discussions, poster presentations.
Second Workshop on Technical AI Governance Research at ICML 2026, focusing on technical approaches to AI governance, policy, and regulation. Part of the main conference workshop track.
Workshop at ICML 2026 focused on identifying, diagnosing, and fixing failure modes in agentic AI systems. Covers reproducible triggers for failures, diagnostic tracing methods, and verified repair approaches. Highly relevant to AI safety and robustness.
Workshop bringing together diverse perspectives from the community to discuss recent advances in mechanistic interpretability, build common understanding and chart future directions. Addresses developing principled methods to analyze and understand model internals (weights and activations) to gain insight into behavior and underlying computation. Received 2.6x submissions from previous year.
Annual mechanistic interpretability workshop at ICML bringing together diverse perspectives to discuss recent advances, build common understanding, and chart future directions. Study of understanding neural network internals through principled analysis methods. 23 spotlight presentations plus poster sessions. Addresses how understanding neural network mechanisms can predict behavior, ensure reliability, and detect adversarial patterns.
Annual mechanistic interpretability workshop at ICML. High-quality venue for mech interp research organized by leading researchers in the field.
Interdisciplinary forum grounded in the science of AI safety. Brings together researchers, government officials, industry leaders, and civil society. 79 speakers across 75 sessions with 330 participants. Topics include technical AI safety, governance, risk assessment, evaluation frameworks, and cross-sector collaboration.
3-week ML upskilling bootcamp for AI safety focusing on interpretability and RL, based on ARENA curriculum. Run by Cambridge Boston Alignment Initiative. Provides housing, meals, 24/7 office access, dedicated teaching assistants, and travel support to participants.
10-week Cambridge-based fellowship on existential AI risk for researchers and entrepreneurs at any career stage. Provides competitive stipend, meals during working hours, transport, visa coverage, and lodging. Rolling application process via expression of interest. Focuses on AI safety and governance research.
10-week Cambridge-based fellowship on existential AI risk. Fellows paid a competitive stipend, with meals during working hours, transport, visas, and lodging covered for the duration. Welcomes talented individuals from around the world at any career stage motivated to contribute to AI safety and governance research. 30+ events hosted over the fellowship period, with weekly mentorship and dedicated research management support.
International Conference on Machine Learning. Premier gathering for ML professionals showcasing cutting-edge research on all aspects of machine learning, AI, statistics, and data science, plus applications in computer vision, computational biology, speech recognition, and robotics. Structure: July 6 for Expo/Tutorial, July 7-9 main conference, July 10-11 workshops. Safety-related workshops extracted separately.
Part of ongoing Alignment Workshop series convening global leaders in academia and industry to deepen collective understanding of potential risks from Artificial General Intelligence. Organized by FAR AI.
Workshop on practical tensions between model developers and evaluation researchers. Surfaces practical insights from across the evaluation ecosystem examining challenges in AI evaluation methodology, infrastructure costs, and sociotechnical impacts. Topics include evaluation methodology and measurement theory, infrastructure/costs/stakeholder relationships, and sociotechnical impact assessments across bias, privacy, labor, and environmental dimensions.
AI safety conference in Christchurch, New Zealand. Organized by BlueDot Impact community. In-person event bringing together the New Zealand AI safety community.
29th Annual Meeting of the Association for the Scientific Study of Consciousness in Santiago de Chile. Academic society promoting rigorous research on understanding the nature, function, and underlying mechanisms of consciousness. Relevant for ACO practitioners working on consciousness studies and measurement science applicable to interpretability work. Includes members from cognitive science, medicine, neuroscience, philosophy, and related disciplines.
Nine-week AI safety research fellowship with potential 6-month extensions (70-90% of fellows receive extensions). ยฃ6,000-ยฃ8,000 stipend, travel coverage, ยฃ2,400 housing for non-London residents, computational resources. Seven cohorts completed with 129 alumni. Four-stage selection: written application, video interview, mentor-specific work task, personal interview.
AI safety research fellowship based in London. Runs June 29 to August 28, 2026, hosted at LISA workspace. In-person participation. Base stipend ยฃ6,000-ยฃ8,000 (Senior Fellows receive ยฃ8,000), additional coverage for travel to London, ยฃ2,400 housing for non-London residents, compute resources, and weekday meals at LISA. Accepts anyone 18 or older regardless of academic background.
AI safety research fellowship based in London. Quarterly cohorts. Q3 cohort runs June 29 to August 28, 2026. ยฃ6,000-ยฃ8,000 stipend plus travel, housing, meals & compute. Up to 6-month extensions for strong projects (70-90% of recent fellows received extensions). Weekly one-on-one mentorship with established researchers. In-person co-working space with meals included. Application deadline May 3rd, decision date May 28th (delayed one week due to high application volume).
AI policy hackathon focused on governance challenges of deployed AI systems. Organized by BlueDot Impact community in Toronto, bringing together participants to work on practical policy frameworks for AI governance in real-world deployment contexts.
Regional AI safety hackathon for Latin America, Africa, and Asia. Participants build AI safety tools, evaluations, and policy research. Regional competition structure with pipeline to fellowship and placement opportunities. Funded by Schmidt Sciences.
United Nations Institute for Disarmament Research global conference on AI governance, security, and ethics. Part of UNIDIR's Centre of Excellence on AI, Peace and Security programming. International AI policy and disarmament-related conference in scope per the safety community's broader governance interests.
Two-week intensive summer school at Columbia University covering machine learning topics including mechanistic interpretability, alignment/safety, RAG & agents, and LLM systems. Approximately 200 PhD students participate alongside faculty and industry speakers. In-scope due to dedicated alignment and mechanistic interpretability tracks.
Bringing together multiple faith perspectives on AI governance and security implications, exploring ethical dimensions of advanced AI through inter-religious dialogue.
Nine-week AI safety research fellowship in Cambridge, MA. $10,000 stipend, housing through Harvard dorms and Airbnb, free weekday meals, 24/7 office access in Harvard Square, and up to $10,000 in compute credits per fellow. Pairs fellows with established researchers from Harvard, MIT, Northeastern, and Boston University for 1-2 hours/week mentorship. Approximately 30 fellows accepted for AI Safety track. Rolling applications.
Nine-week research fellowship with 30 AI safety fellows and 15 AIxBiosecurity fellows. Provides $10,000 stipend, housing through Harvard dorms, free weekday meals, and up to $10,000 in compute support (API credits and GPU access). In-person participation required; must have valid US work authorization.
Three-month bipartisan fellowship designed to launch or accelerate impactful careers in American AI governance and policy. Participants deepen understanding of the field, connect with network of experts, and build skills and professional profile. $21,000 stipend. Alumni have secured positions at leading AI companies (DeepMind, OpenAI, Anthropic).
Three-month fellowship where fellows conduct independent research on AI governance topic of their choice with mentorship from leading experts. ยฃ12,000 stipend. GovAI was founded to help decision-makers navigate the transition to advanced AI through rigorous research and talent fostering. Alumni have secured positions at DeepMind, OpenAI, Anthropic.
AI Risk Content Hackathon organized by BlueDot Impact in London. Focus on creating AI safety and risk communication content. Part of BlueDot's broader mission to build the workforce needed to safely navigate AGI.
Foresight Institute flagship event gathering leading scientists, entrepreneurs, funders, and policymakers to explore the frontiers of science and technology. Multiple tracks including AI safety topics.
Flagship conference where leading scientists, entrepreneurs, funders, and policymakers convene to explore frontier technology and plan for beneficial futures. Includes AI safety track as part of broader focus on transformative technology. 40-year-old organization focused on beneficial technology development.
Foresight Institute Vision Weekend with frontier science and technology tracks including AI safety. 40-year-old organization focused on transformative technology. Three-day event in London featuring AI safety programming alongside other frontier tech tracks.
Five-day intensive programme for AI safety founders going from idea to funded. Successful pitches receive ยฃ50k in equity-free seed funding. Part of BlueDot Impact's incubator and rapid-funding initiatives supporting concrete AI safety work.
Three-month research program investigating how advanced AI affects society and examining institutions and policies that could help communities respond effectively to these changes. Organized by Center for AI Safety.
Largest MATS program to date with 120 fellows and 100 mentors. Fellows connected with mentors or organizational research groups such as Anthropic's Alignment Science team, UK AISI, Redwood Research, ARC, and LawZero to collaborate on research projects over the summer. ML Alignment & Theory Scholars is one of the most established AI safety research fellowships.
EA Global conference series organized by Centre for Effective Altruism. Speakers present research on effective altruism including heavy AI safety programming. Features talks, workshops, and networking opportunities. In scope for social and community event reasons with significant AI safety attendance.
In-person bootcamp covering AI safety fundamentals, mechanistic interpretability, and reinforcement learning. Program covers travel, visas, accommodation, and meals. Duration: 4-5 weeks typical for ARENA bootcamps. Online curriculum available for independent study.
8-week virtual reading group run by MIT AI Alignment (MAIA), meeting 2 hours per week. Explores why AI safety matters and current mitigation approaches including AI trajectory, misalignment risks, technical safety solutions, policy, and career paths. No prior AI background required. Led by small groups facilitated by MAIA team members.
Hackathon focused on secure program synthesis by Apart Research. Hybrid format with online participation and in-person hubs.
5-week project-based course for engineers and early researchers to work with an AI safety expert on a contribution to AI safety research or engineering. Includes mentorship, regular check-ins, and a published write-up. Covers alignment, mechanistic interpretability, evaluations, red-teaming, AI control, and scalable oversight.
3-week ML upskilling bootcamp for AI safety focusing on interpretability and RL, based on ARENA curriculum. Run by Cambridge Boston Alignment Initiative. Provides housing, meals, 24/7 office access, dedicated teaching assistants, and travel support to participants.
Workshop on infrastructure for secure and verifiable AI, co-organized with Center for AI Safety. Brings together experts across ML, hardware security, systems, cryptography, and computer security to identify promising technical approaches and foster collaborations.
FAR.AI and Center for AI Safety workshop on infrastructure for secure and verifiable AI, bringing together researchers, builders, and funders across ML, hardware security, systems, cryptography, and computer security to identify the most promising technical approaches and spark concrete collaborations.
Free, one-day AI safety event. Third iteration of TAIS conference, first time hosted in UK. Organized by Oxford Martin AI Governance Initiative and Noeon Research. Supported by MATS, Apart Research, IASEAI, and Foresight Institute, with volunteer assistance from OAISI. Welcomes participants from all backgrounds with no prior research experience required. Previous talks available on YouTube.
A free, accessible workshop hosted by AI Safety Awareness Group Oakland exploring AI's trajectory and societal impact. No technical background required. Features live demonstrations of current AI systems, interactive forecasting activities, and discussions about AI's implications for work, relationships, and society over the next 1-5 years.
Two-day conference on global cooperation for cybersecurity resilience and stability. Organized by United Nations Institute for Disarmament Research. Addresses international frameworks for cyber governance and security cooperation.
AI safety research fellowship designed to accelerate AI safety research and foster research talent. Focus on steering and controlling future powerful AI systems, understanding and evaluating risks. Runs multiple cohorts per year.
13-week part-time remote fellowship designed to develop Black researchers, practitioners, and leaders in AI Safety, AI Security, and AI Governance. Program runs in two phases: Weeks 1-5 cover foundations and training curriculum, while Weeks 6-13 focus on research and project development. Organized by Black in AI Safety & Ethics (BASE) to empower Black researchers in the AI safety community.
A community-organized regional Effective Altruism conference targeting the policy, research, and public-interest communities across the Washington DC, Maryland, and Virginia area. Designed to help professionals connect with practitioners, explore high-impact career paths, and engage with organizations focused on global challenges. Key topics include global health, AI governance, public policy, and biosecurity.
AI safety mixer for professionals exploring a move into AI safety, hosted by the London Initiative for Safe AI (LISA). A low-pressure evening to explore the field, understand the part you could play, and meet people who are already working in AI safety. Designed for professionals interested in transitioning into the AI safety field.
Workshop bringing together researchers from ML, neuroscience, and cognitive science to explore representational alignment. 2026 edition focuses on what affordances alignment makes possible, specifically examining neural control and downstream behavior applications. Includes Re-Align Challenge with leaderboard submissions. Short papers (up to 5 pages) and long papers (up to 10 pages) accepted. Non-archival format.
Three-day hybrid hackathon focusing on AI and biosafety. Organized by Apart Research with in-person hubs in London, Berlin, and San Francisco.
Two-day conference focused on AI control: reducing risks from AI misalignment through interventions robust even when AI models attempt to undermine safeguards. Features speaker talks, fireside chats, and breakout discussions. Organized by Redwood Research and FAR.AI. Pre-conference workshop on April 17 for newcomers. Featured speakers from Redwood Research, Anthropic, METR, and CMU. Free to attend, application-based registration.
Second annual Technical Innovations in AI Policy Conference organized by FAR.AI in collaboration with leading think tanks. Focuses on technical innovations that enable AI policy implementation.
A three-day hackathon focused on developing safety measures for potentially misaligned AI systems. Participants work on control protocols and evaluation tools to keep autonomous AI agents contained, addressing oversight challenges as AI systems become more autonomous. Co-organized by Apart Research and Redwood Research (founder of the AI control field), with 700+ participants submitting 126 projects across main competition and specialty tracks.
FAR.AI hosted event bringing together more than 200 researchers, policymakers, and industry experts to discuss AI alignment topics. Featured presentations on honeypot-based methods for detecting scheming, and training against interpretability-based deception detectors. Victoria Krakovna (Google DeepMind) and Stefan Heimersheim (Google DeepMind) presented.
Interdisciplinary workshop examining how generative AI models like LLMs are transforming information creation and access while addressing manipulation and public discourse integrity concerns. Features three expert panel discussions on AI manipulation definitions, measurement methodologies, and societal impacts, plus poster session showcasing community research.
First AI, Manipulation, and Information Integrity workshop addressing how generative AI models impact information creation and access, with focus on manipulation, deception, and public discourse integrity. Interdisciplinary workshop examining AI's persuasive capabilities and societal implications. Program features three panel discussions by leading experts, poster session, and interactive discussions.
Half-day workshop at IASEAI'26 bringing together researchers from computer science, cognitive science, philosophy, political science, and policy to examine AI's manipulative capabilities and information integrity concerns. Features three panel discussions with leading researchers, plus a poster session showcasing 24 accepted posters and 5 hackathon-winning projects addressing definitions and taxonomies of persuasion, manipulation, and deception.
A part-time, remote research fellowship enabling aspiring AI safety and policy researchers to collaborate with professionals on impactful projects addressing risks from artificial intelligence. Participants commit 5-40 hours weekly for approximately 3 months, culminating in a Demo Day presentation. Mentors include professionals from Google DeepMind, RAND, Apollo Research, MATS, UK AISI.
SPAR (Supervised Program for Alignment Research) Spring 2026 mentee track. Part-time remote research program pairing aspiring researchers with experienced mentors from Google DeepMind, RAND, Apollo Research, MATS, UK AISI for three-month projects. Mentees commit 5-40 hours per week. Research period: February 16 - May 16. Mentee decisions were sent out February 2-6. Applications for Spring 2026 have closed.
Part-time remote research fellowship pairing aspiring researchers with 130+ experienced mentors from Google DeepMind, RAND, Apollo Research, MATS, UK AISI, and other top organizations for three-month AI alignment projects. Participants commit 5-40 hours weekly. Covers project expenses including compute and API/LLM access. Culminates in virtual Demo Day with prizes totaling $7,000. Optional continuation beyond May 16. Mentor application track: experienced researchers from Google DeepMind, RAND, Apollo Research, MATS, UK AISI etc. apply to mentor a project. Mentor application deadline 2025-12-05 (passed).
A three-day conference bringing together the effective altruism community to share new thinking and research, coordinate on global projects, and network. Features keynote speakers, talks, workshops, and social activities focused on addressing pressing global problems including AI safety. Friday evening opening reception, full-day Saturday and Sunday programming.
150-hour academic programme combining technical depth with policy and governance perspectives on AI evaluation. Includes 90 hours of online work, 20 hours of hands-on courses, and a 40-hour in-person capstone week in Valencia. Participants gain expertise in AI evaluation, capabilities assessment, and safety considerations. Fully funded scholarships for 40 global participants. Awarded 15 ECTS Expert Diploma by ValgrAI.
A three-day hackathon bringing together builders to create verification systems, compliance infrastructure, and coordination tools for international AI governance. Five focus tracks: Hardware Verification & Attestation, Compliance Infrastructure & Privacy-Preserving Proofs, Risk Thresholds & Compute Verification, International Verification & Coordination, and Research Governance & Dual-Use Detection. Organized by Apart Research in partnership with MIRI Technical Governance Team and Lucid Computing.
Third international AI governance workshop focusing on alignment, morality, law, and design. Addresses governance challenges from generative and agentic AI systems. Full-day hybrid event with keynotes, technical sessions, industry panels, lightning talks, poster sessions, and interactive discussions. Expected attendance of 120-150 participants.
Twelve-week intensive AI safety research fellowship with 25+ hours per week commitment. Two-stage selection: 60 participants accepted for trial week (Jan 19-23) on foundational coursework in RLHF, interpretability, SAEs, and adversarial robustness, then 30 selected as Research Fellows based on performance. Fully funded by Open Philanthropy, covers tuition, computing infrastructure access, mentorship, and limited conference travel support.
Three-month online AI safety research program where participants form teams to work on pre-selected projects. Opening weekend January 10-11, projects run through April 19, final presentations April 24-27. Features 27 projects across six themes: Stop/Pause AI (5), Policy/Governance (5), Evaluate Risks from AI (5), Mech-Interp (2), Agent Foundations (4), Alternative LLM Safety (5), and Safe by Design AIs (1). Participants work 10 hours per week with weekly team meetings. Seven-year track record with alumni founding 10 organizations and securing 43 jobs in AI Safety. Application deadline November 23, 2025.
3-month long online program where participants form teams to work on pre-selected AI Safety projects. Opening weekend January 10-11 through April 19 with final presentations April 24-27. Three tracks available covering alignment research, governance, policy, and stop/pause advocacy. Team member applications deadline was November 23rd.
10-week hybrid fellowship with 7 weeks in-person, 2 days per week in-person attendance. 5-10 hours of project work weekly plus discussions, mentorship, and speaker sessions. Mandatory Saturday sessions (10am-6pm) plus optional weekday coworking. Customized discussions, speaker presentations, dedicated coworking space, social events, compute resources, mentorship, and networking. Potential flight reimbursement for top candidates (capped at regional flight costs). Organized by Chris Leong and Jack Payne.
Fully funded, in-person programme pairing senior advisors with emerging talent on 5-month technical, governance, strategy, and field-building projects. Two streams: Empirical (ML research in technical safety areas including alignment, control, evals, scalable oversight) and Strategy/Governance. Monthly stipend of $8,400, approximately $15,000 monthly research budget per empirical fellow for compute. Weekly mentorship from industry experts, visa support for international applicants.
Cambridge Bootcamp for Research in Interpretability and Alignment: 3-week ML upskilling bootcamp for AI safety, focusing on interpretability and RL. Based on ARENA curriculum. Run by Cambridge Boston Alignment Initiative in Cambridge, Massachusetts. In-person intensive programme.
Part-time, remote-first, 12-week fellowship by Future Impact Group for students and early-career researchers. Minimum 8+ hours/week on research projects in AI policy, philosophy for safe AI (technical safety and ethical foundations), or AI sentience. Provides co-working sessions, issue troubleshooting, career guidance, networking, and guest speakers.
Part-time, remote-first, 12-week research fellowship where participants work as research associates on specific projects under experienced supervision. Focus areas: AI governance, technical AI safety, and digital sentience. Time commitment: 8+ hours per week. Includes co-working sessions, issue troubleshooting, career guidance, opening and closing events, networking opportunities, research sprints, and guest speakers.