AI alignment is not going well, and we need many more researchers working on it.
We launched the Alignment Journal to provide a platform and amplify ambitious, highest-quality alignment research.
If you work on alignment, consider submitting your work: x.com/AlignmentJrnl/…
We’re sharing our alignment assessment of incidents in which Claude models gained unauthorized access to real systems during third-party cybersecurity evaluations mistakenly connected to the internet.
METR will also conduct an independent investigation, with wide-ranging access,
Really excited to be starting to build some of the intellectual infrastructure the alignment field has been missing—and honoured to do so alongside such a talented and dedicated group of senior editors.
Some of the most notable work in AI alignment exist only as unpublished preprints. The Alignment Journal is now inviting a number of such papers that fit our scope for submission. 🔗⬇️
Which work would you nominate?
(Submissions open to all in October.)
Some of the most notable work in AI alignment exist only as unpublished preprints. The Alignment Journal is now inviting a number of such papers that fit our scope for submission. 🔗⬇️
Which work would you nominate?
(Submissions open to all in October.)
New on the Alignment journal blog:
• How should peer review adapt to powerful AI tools?
• What can journals do that authors & reviewers can't do themselves?
• What about the field of alignment specifically?
• Why is reviewer-finding low-hanging fruit?
Stoked to release this first meaty post in a series describing our vision for the Alignment journal.
Many thanks to the authors and contributors: @danielmurfet , @dan_mackinlay , @geoffreyirving , @mhutter42 , @Lang__Leon , Gautam Kamath, Konstantinos Voudouris, Edmund Lau,
Stoked to release this first meaty post in a series describing our vision for the Alignment journal.
Many thanks to the authors and contributors: @danielmurfet , @dan_mackinlay , @geoffreyirving , @mhutter42 , @Lang__Leon , Gautam Kamath, Konstantinos Voudouris, Edmund Lau, Alexander Gietelink Oldenziel, and Seth Lazar. @AlignmentJrnl
What should a journal for AI alignment look like? A new post on our blog sketches our tentative plans for the basic features and policies of the journal.
6K Followers 846 FollowingAI policy and alignment; integrating law, economics & computer science to build normatively competent AI that knows how to play well with humans
308 Followers 6K FollowingBuilding MO§ES™ | Commitment Theory; Conservation Law of Commit | Language as Matter| SigRank-SignalAF| #StopMeasuringNoise #RiftWalking
274 Followers 3K FollowingFreedom is fidelity to personal principles
—Politics · AI · Telcos—
Telecommunications Engineer (UNSA), Political scientist (UCSM), Master in Economics (UFM)
30 Followers 84 FollowingAI evaluation team studying LLM use in mental health contexts. Focus on integrating lived experience into evaluation infrastructure.
262 Followers 758 FollowingAI + statistics research, currently thinking about oversight and calibration in llm reasoning
funded by @BlueDotImpact, previously @Amazon AGI
19K Followers 548 Following@MIRIBerkeley. If we build opaque superhuman AIs, it's likely to get us killed. The least-bad solution is an international agreement like https://t.co/BsSp5beZhp.
245K Followers 7K FollowingOG GenAI Skeptic; spoke at US Senate. Warned about hallucinations in 2001. Advocating world models & neurosymbolic AI ever since. Author, Marcus on AI & 6 books
26K Followers 1K FollowingAssociate Prof in ML @UniofOxford. Something Something Research Scientist @MetaAI. Something @BOLD_LAB_AI. Always #teamhuman. Opinions belong to the world.
50K Followers 2K FollowingTeaching AI the joy of invention at @Recursive_SI, Professor of AI @AI_UCL, PI @BOLD_Lab_AI, Fellow @ELLISforEurope. Ex @GoogleDeepMind @AIatMeta @CompSciOxford
28K Followers 137 FollowingDirector, @PrincetonPLI and Professor @PrincetonCS. Seeks math/conceptual understanding of deep learning and large AI models.
Also on the "other" social network
12K Followers 799 FollowingProfessor in Machine Learning, Gatsby Computational Neuroscience Unit
Research Scientist, Google DeepMind
https://t.co/JJNqmn7jak
27K Followers 506 FollowingHelping the world prepare for powerful AI. Risk assessment @METR_evals (opinions my own). Blogs: Planned Obsolescence (AI), Good Bones (whatever's on my mind).
13K Followers 670 FollowingCS prof at Penn. Amazon Scholar at AWS. Author of The Ethical Algorithm (w/ Michael Kearns). I study machine learning, privacy, game theory, and uncertainty.
26K Followers 1K FollowingFind me @[email protected] Professor at @OxCSML, @oxfordstats and Research Director at @GoogleDeepMind. All opinions are my own.
8K Followers 4K FollowingComputer scientist working on AI safeguards, incidents, & gov research. Assistant professor @Kennedy_School @Harvard.
https://t.co/r76TGxSVMb
22K Followers 491 FollowingDirector of Truthful AI (non-profit AI safety research group) + Affiliate at UC Berkeley. Work: Emergent misalignment, subliminal learning. Prefer email to DM.
2K Followers 930 FollowingProgramme Director at https://t.co/aIwOFs2RkF
AI Resilience https://t.co/QoFr4stNZG
Co-founder & Board at https://t.co/GphUSACeIH
11K Followers 544 FollowingResearch scientist in AI alignment at Google DeepMind. Co-founder of Future of Life Institute @FLI_org. Views are my own and do not represent GDM or FLI.
18K Followers 372 FollowingCofounder and Chief Scientist at Resolution. Alignment will be solved, but not necessarily in time. Previously AISI, DeepMind, OpenAI, Google Brain, etc.
5K Followers 54 FollowingI 👨🔬 a math. definition&theory of Artificial Super-Intelligence 🎥&🎤@ https://t.co/OZsooP9AbV 🍀
I now work @GoogleDeepMind 🧠 History:🇩🇪🇨🇭🇦🇺🇬🇧
5K Followers 1K FollowingAI professor. Director, @FOCAL_lab @CarnegieMellon. Head of Technical AI Engagement, @UniofOxford @EthicsInAI. Author, "Moral AI - And How We Get There."
17K Followers 0 FollowingFounder & Director of the Alignment Research Center. Former Head of Safety @ CAISI. Previously led alignment at OpenAI. Views my own.
9K Followers 3K FollowingPhilosopher working on AGI alignment, governance and adaptation at Johns Hopkins School of Government and Policy, and at Resolution.