Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
majdalkawaas
/
matryoshka-hypencoder
like
0
Feature Extraction
Safetensors
microsoft/ms_marco
jfkback/hypencoder-msmarco-training-dataset
sentence-similarity
retrieval
hypencoder
matryoshka
msmarco
arxiv:
2607.17457
arxiv:
2502.05364
License:
apache-2.0
Model card
Files
Files and versions
xet
Community
Copy to bucket
new
main
matryoshka-hypencoder
800 MB
Ctrl+K
Ctrl+K
1 contributor
History:
4 commits
majdalkawaas
Added link to the accompanying paper
67e6c4a
verified
8 days ago
.gitattributes
Safe
1.52 kB
initial commit
about 2 months ago
README.md
5.05 kB
Added link to the accompanying paper
8 days ago
config.json
Safe
945 Bytes
Upload model from checkpoint-13100
about 2 months ago
model.safetensors
558 MB
xet
Upload model from checkpoint-13100
about 2 months ago
optimizer.pt
241 MB
xet
Upload model from checkpoint-13100
about 2 months ago
rng_state.pth
14.6 kB
xet
Upload model from checkpoint-13100
about 2 months ago
scheduler.pt
1.47 kB
xet
Upload model from checkpoint-13100
about 2 months ago
special_tokens_map.json
Safe
125 Bytes
Upload model from checkpoint-13100
about 2 months ago
tokenizer.json
Safe
711 kB
Upload model from checkpoint-13100
about 2 months ago
tokenizer_config.json
Safe
314 Bytes
Upload model from checkpoint-13100
about 2 months ago
trainer_state.json
Safe
8.25 kB
Upload model from checkpoint-13100
about 2 months ago
training_args.bin
4.88 kB
xet
Upload model from checkpoint-13100
about 2 months ago
vocab.txt
Safe
232 kB
Upload model from checkpoint-13100
about 2 months ago