DeepSeek releases v4.1-Flash, a 763B-parameter encoder-decoder model with vision
DeepSeek has introduced v4.1-Flash, a large-scale model with 763B total parameters built on a new causal encoder-decoder design that also handles vision input. Commentary from Latent Space and Sebastian Raschka argues the release is significant enough that it should have been branded as a new major version rather than a point update. Details on training, availability and licensing were not included in the report.