The very newly released Deepseek R1 “reasoning model” from China beats OpenAI’s o1 model on multiple areas, it seems – and you can even see all the steps of the pre-answering “thinking” that’s hidden from the user in o1. It’s a huge model, but it (and the paper about it) will probably positively impact future “open source” models in general, now the “thinking” cat’s outta the bag. Though, it can’t think about Tiananmen Square or Taiwan’s autonomy – but many derivative models will probably be modified to effectively remove such Chinese censorship.
The very newly released Deepseek R1 “reasoning model” from China beats OpenAI’s o1 model on multiple areas, it seems – and you can even see all the steps of the pre-answering “thinking” that’s hidden from the user in o1. It’s a huge model, but it (and the paper about it) will probably positively impact future “open source” models in general, now the “thinking” cat’s outta the bag. Though, it can’t think about Tiananmen Square or Taiwan’s autonomy – but many derivative models will probably be modified to effectively remove such Chinese censorship.