[ gpt-oss-20b-finance weights on Hugging Face ] [Try the playground ] [ LineageEval ] [Explore the data on GitHub ] β +45.45 DeepSeek V4 Flash censorship gap on China-sensitive prompts vs matched controls Β· 76 pairs Β· four judges 83.61% CTGT GPT-OSS-120B on FinanceReasoning at 8k budget Β· above Kimi K3 at 81.93% and Inkling at 65.13% 62Γ Lower cost per query than Inkling at the same budget Β· 160Γ lower than Kimi K3 β The affordability and accessibility of open frontier models has led to their widespread usage among American developers and enterprises. While this has enabled the benefits of AI to be reaped by more people, concerns have mounted over models influenced by foreign actors, namely the Chinese Communist Party. The worry expressed in Washington and regulated industries is that values, censorship or viewpoints at odds with American ideals are intrinsically transferred along with the gains in intelligence. We wanted to rigorously examine this phenomenon under a controlled scenario.β¦