LaSER-Qwen3-8B: Alibaba NLP's 8B dense retriever with latent reasoning released on Hugging Face
Original titleAlibaba-NLP/LaSER-Qwen3-8B
AISummary
Alibaba NLP released LaSER-Qwen3-8B, an 8B-parameter dense retriever built on Qwen/Qwen3-8B that internalizes explicit reasoning into latent space through continuous latent thinking tokens.
The model scores 29.3 nDCG@10 on the BRIGHT benchmark, ahead of the rewrite-then-retrieve pipeline's 28.1, and carries a 4096-dimension embedding with an 8192-token maximum sequence length.
It is licensed under MIT and adds about 1.7× latency over standard single-pass dense retrievers.
Source: Alibaba NLP (Tongyi) · new models on Hugging Face · huggingface.co