If there were backdoors and biases hidden in open-weight models, you could count on closed-weight labs to find and reveal them.
@naval (posted 2026-07-26) asserts that if open-weight models contained backdoors…
Brief
Naval argues that market and competitive incentives mean closed-weight (proprietary) labs would uncover and disclose any backdoors or biases in open-weight models. He implies secrecy wouldn’t let such flaws remain hidden because rival organizations would benefit from finding and revealing them, framing disclosure as an expected outcome rather than a risk.
Why it matters
@naval (posted 2026-07-26) asserts that if open-weight models contained backdoors or hidden biases, closed-weight/closed-source labs would detect and publicly reveal them.
Key details
- The claim rests on the incentive that proprietary labs have strong motives to find and expose flaws in open models rather than keep them hidden.