Twitter/X

@naval (posted 2026-07-26) asserts that if open-weight models contained backdoors…

Brief

Naval argues that market and competitive incentives mean closed-weight (proprietary) labs would uncover and disclose any backdoors or biases in open-weight models. He implies secrecy wouldn’t let such flaws remain hidden because rival organizations would benefit from finding and revealing them, framing disclosure as an expected outcome rather than a risk.

Why it matters

@naval (posted 2026-07-26) asserts that if open-weight models contained backdoors or hidden biases, closed-weight/closed-source labs would detect and publicly reveal them.

Key details

  • The claim rests on the incentive that proprietary labs have strong motives to find and expose flaws in open models rather than keep them hidden.
Source evidence

If there were backdoors and biases hidden in open-weight models, you could count on closed-weight labs to find and reveal them.