tencent/R3-rerank-0.6b
Primitive: /score · Score ·
Qwen3
The latest agent skill reranking model at the 0.6B scale. R3-Reranker is the cross-encoder (rerank) stage of R3-Skill's two-stage retriever for query-conditional agent skill retrieval. It scores each (query, skill) pair jointly, paired with R3-Embedding-0.6B for recall.
View on Hugging Face → Fine-tuned from Qwen/Qwen3-Reranker-0.6B
Overview
Hardware: — drives latency, throughput & cost
| Size | 596M params |
|---|---|
| Tasks | /score |
| License | apache-2.0 |
| Latency | 449 ms |
| Throughput | 2.1K tok/s |
| Cost | $0.103 /1M tok |
Cost is approximate — computed from list GPU prices; your actual price depends on the provider you deploy SIE with.
Scoring
| Inputs | text |
|---|---|
| Max sequence length | 4,096 |
Benchmarks
R3Skill (candidates: R3-embedding-0.6b, k=20)
Agent skill retrieval: match user requests to the SKILL.md that solves them (Tencent R3 release set)
R3Skill (candidates: R3-embedding-0.6b, k=20, limit 384)
Agent skill retrieval: match user requests to the SKILL.md that solves them (Tencent R3 release set)