Octen / sentence-similarity updated 7 months ago

Octen-Embedding-8B

Octen-Embedding-8B is a text embedding model developed by Octen for semantic search and retrieval tasks. This model is fine-tuned from Qwen/Qwen3-Embedding-8B and supports multiple languages, providing high-quality embeddings for various applications.

Params
7.6 B
Context
40,960
Downloads 30d
929 K
Likes
192
Commercial use: allowed apache-2.0 Not gated SAFETENSORS 3 languages View on Hugging Face ↗

Download history

daily snapshots · 10 days
929 K453 K
Sep 11Sep 14Sep 17Sep 20

Can you run it?

Estimated VRAM at 8K context unless noted. Pick your hardware to see the verdict per quantization.

FileQuantSizeEst. VRAMVerdict on RTX 4090 · 24 GB
model.safetensors bf16 15.1 GB 18.3 GB ✅ Runs comfortably
model.safetensors (bf16, full) bf16 + 40K ctx 15.1 GB 22.8 GB ⚠️ Tight — reduce context
Estimate: file size × 1.1 + KV cache at 8K + 0.5 GB overhead. Not a benchmark — how we calculate this.

Specifications

Architecture
Qwen3Model
Parameters
7.6 B
Tensor type
BF16
Context length
40,960
Vocabulary
151,665
Layers / heads
36 / 32
Licence
apache-2.0
First seen on the Hub
2025-12-23
Base model
Qwen3-Embedding-8B
Training datasets
undisclosed
Added to our catalog
2026-09-11

Family

Base model and the most-downloaded derivatives in the catalog.

Compare with any sentence-similarity model