
— People won’t want to use agents that are misaligned with them and that don’t do what they ask, so labs have a strong natural incentive to make their models more aligned, he says, while not engaging with the innovation pressure and catastrophic risk at the high-end frontier labs.
— My view is that trust and alignment are quickly becoming the most important capabilities that will differentiate agents and models. Any lab that doesn’t focus on alignment will fall behind, he adds.
He also says that Meta delayed their new Muse agent for «several months» in order to «focus on safety and security,» says they engage with «independent evaluators and advisors» as «industry best practice,» and says matter-of-factually that «other labs can just do this too.»
He then adds that a majority of industry compute should go towards «serving people rather than racing towards recursive self-improvement,» like Meta says they do and that «other labs can do this as well.»
On a more positive note for the ongoing debate, he does say that «it would be helpful for there to be a larger and more diverse ecosystem of evaluators,» which might hint at support for an industry standards body.
Zuckerberg also links up his August The Future is for Everyone-manifesto that says «Any policy that slows American model releases — even by a month — could add significant risk to American leadership while letting foreign models race ahead.»
Read more: Zuckerberg’s X post, Reuters, Business Insider and CNBC.











