Alexandr Wang
@alexandr_wang
We believe strongly in the necessity to invest into alignment.
1. People and businesses will only use agents that are aligned with their intent and values. If we do not build models aligned with people and businesses, then they will move to more aligned options.
2. Every lab should have a strong governance framework across training and deployment. This should include external evaluators, which are best practice for transparency, and independent oversight on things like safety criteria for model launches.
3. Every lab will need to operate within institutional protections of democratic countries. This means labs face significant liability if their models cause harm. This will push the ecosystem in the right ways.
4. Advances in the field are ultimately downstream of compute and resource allocation. Racing on recursive self-improvement is one of the riskiest pathways for potential loss of control to powerful models. Meta is committing the significant majority of our compute towards serving people rather than racing on RSI, and other labs can choose to do the same.
AI is a very powerful technology, and there is immense responsibility in developing it safely alongside the right checks and balances.