Skip to content

LaSER-Qwen3-8B: Alibaba NLP's 8B dense retriever with latent reasoning released on Hugging Face

Original titleAlibaba-NLP/LaSER-Qwen3-8B

AISummary

Alibaba NLP released LaSER-Qwen3-8B, an 8B-parameter dense retriever built on Qwen/Qwen3-8B that internalizes explicit reasoning into latent space through continuous latent thinking tokens.

The model scores 29.3 nDCG@10 on the BRIGHT benchmark, ahead of the rewrite-then-retrieve pipeline's 28.1, and carries a 4096-dimension embedding with an 8192-token maximum sequence length.

It is licensed under MIT and adds about 1.7× latency over standard single-pass dense retrievers.

Read the original huggingface.co

Source: Alibaba NLP (Tongyi) · new models on Hugging Face · huggingface.co