Text / tencent

WeMM Embedding 2B

WeMM Embedding 2B from tencent. Source-based hardware guidance from its published configuration.

EstimatedRepository opened Aug 25, 2026Source checked 9/23/2026Version: bbd6cd4b
LOCALRENTED GPUOPEN WEIGHTS

At a glance

Parameters
2.72B
Architecture
qwen3_5
License
apache-2.0
Software
SGLang, vLLM
View the model source
MY HARDWARE

Will it run on your machine?

Save your machine to see a personalized rating and its reasoning.

Add my hardware

Ways to run it

SGLang · Linux

Repository-specific command in publisher documentation; confirm dependencies and hardware in the source.

python -m sglang.launch_server --model-path "$MODEL_PATH" --is-embedding --enable-precise-embedding-interpolation
Official instructions

vLLM · Linux

Repository-specific command in publisher documentation; confirm dependencies and hardware in the source.

vllm serve "$MODEL_PATH" --runner pooling --chat-template "$MODEL_PATH/embedding_chat_template.jinja"
Official instructions

Explore its uses

SAFETENSORS

Keep exploring

WeMM Embedding 2B: hardware, VRAM & setup | YouRunAI