---
tags:
- sentence-transformers
- sentence-similarity
- feature-extraction
- dense
- generated_from_trainer
- dataset_size:705905
- loss:MultipleNegativesSymmetricRankingLoss
base_model: sentence-transformers/all-MiniLM-L6-v2
widget:
- source_sentence: gerber baby food fruits apples bananas & cereal
sentences:
- world of sweets puzzle
- baby food
- baby food
- source_sentence: granville original one bite original rice crispy squares
sentences:
- ' one bite rice crispy '
- sweet
- bounty wafer rolls
- source_sentence: rosa / porcelain us andalusia mug
sentences:
- mug
- ' rosa mug'
- melamine small plate - teal
- source_sentence: cetaphil sunscreen spf 50+ cream 89 ml
sentences:
- sunscreen
- ' cetaphil sunscreen cream'
- garnier intensity (6.60) intense ruby
- source_sentence: italian dolce provolone
sentences:
- trident - gum strawberry flavor - 5 per pack
- experience the authentic taste of italy with our italian dolce provolone. indulge
in its creamy texture, delicate flavors, and versatility in both simple and sophisticated
culinary creations.
- dairy
pipeline_tag: sentence-similarity
library_name: sentence-transformers
metrics:
- cosine_accuracy
model-index:
- name: SentenceTransformer based on sentence-transformers/all-MiniLM-L6-v2
results:
- task:
type: triplet
name: Triplet
dataset:
name: Unknown
type: unknown
metrics:
- type: cosine_accuracy
value: 0.9678199887275696
name: Cosine Accuracy
---
# SentenceTransformer based on sentence-transformers/all-MiniLM-L6-v2
This is a [sentence-transformers](https://www.SBERT.net) model finetuned from [sentence-transformers/all-MiniLM-L6-v2](https://huggingface.co/sentence-transformers/all-MiniLM-L6-v2). It maps sentences & paragraphs to a 384-dimensional dense vector space and can be used for semantic textual similarity, semantic search, paraphrase mining, text classification, clustering, and more.
## Model Details
### Model Description
- **Model Type:** Sentence Transformer
- **Base model:** [sentence-transformers/all-MiniLM-L6-v2](https://huggingface.co/sentence-transformers/all-MiniLM-L6-v2)
- **Maximum Sequence Length:** 256 tokens
- **Output Dimensionality:** 384 dimensions
- **Similarity Function:** Cosine Similarity
### Model Sources
- **Documentation:** [Sentence Transformers Documentation](https://sbert.net)
- **Repository:** [Sentence Transformers on GitHub](https://github.com/huggingface/sentence-transformers)
- **Hugging Face:** [Sentence Transformers on Hugging Face](https://huggingface.co/models?library=sentence-transformers)
### Full Model Architecture
```
SentenceTransformer(
(0): Transformer({'max_seq_length': 256, 'do_lower_case': False, 'architecture': 'BertModel'})
(1): Pooling({'word_embedding_dimension': 384, 'pooling_mode_cls_token': False, 'pooling_mode_mean_tokens': True, 'pooling_mode_max_tokens': False, 'pooling_mode_mean_sqrt_len_tokens': False, 'pooling_mode_weightedmean_tokens': False, 'pooling_mode_lasttoken': False, 'include_prompt': True})
(2): Normalize()
)
```
## Usage
### Direct Usage (Sentence Transformers)
First install the Sentence Transformers library:
```bash
pip install -U sentence-transformers
```
Then you can load this model and run inference.
```python
from sentence_transformers import SentenceTransformer
# Download from the 🤗 Hub
model = SentenceTransformer("LamaDiab/v2MiniLM-V18Data-256ConstantBATCH-SemanticEngine")
# Run inference
sentences = [
'italian dolce provolone',
'experience the authentic taste of italy with our italian dolce provolone. indulge in its creamy texture, delicate flavors, and versatility in both simple and sophisticated culinary creations.',
'trident - gum strawberry flavor - 5 per pack',
]
embeddings = model.encode(sentences)
print(embeddings.shape)
# [3, 384]
# Get the similarity scores for the embeddings
similarities = model.similarity(embeddings, embeddings)
print(similarities)
# tensor([[1.0000, 0.8368, 0.2176],
# [0.8368, 1.0000, 0.2259],
# [0.2176, 0.2259, 1.0000]])
```
## Evaluation
### Metrics
#### Triplet
* Evaluated with [TripletEvaluator](https://sbert.net/docs/package_reference/sentence_transformer/evaluation.html#sentence_transformers.evaluation.TripletEvaluator)
| Metric | Value |
|:--------------------|:-----------|
| **cosine_accuracy** | **0.9678** |
## Training Details
### Training Dataset
#### Unnamed Dataset
* Size: 705,905 training samples
* Columns: anchor, positive, and itemCategory
* Approximate statistics based on the first 1000 samples:
| | anchor | positive | itemCategory |
|:--------|:----------------------------------------------------------------------------------|:---------------------------------------------------------------------------------|:---------------------------------------------------------------------------------|
| type | string | string | string |
| details |
mango nos nos small | milk chocolate ganache cake | sweet |
| lux soap creamy perfection 165 gm | soap | hand soap |
| grey deo original | classic deodrant | women's deodorant |
* Loss: [MultipleNegativesSymmetricRankingLoss](https://sbert.net/docs/package_reference/sentence_transformer/losses.html#multiplenegativessymmetricrankingloss) with these parameters:
```json
{
"scale": 20.0,
"similarity_fct": "cos_sim",
"gather_across_devices": false
}
```
### Evaluation Dataset
#### Unnamed Dataset
* Size: 9,509 evaluation samples
* Columns: anchor, positive, negative, and itemCategory
* Approximate statistics based on the first 1000 samples:
| | anchor | positive | negative | itemCategory |
|:--------|:---------------------------------------------------------------------------------|:----------------------------------------------------------------------------------|:---------------------------------------------------------------------------------|:---------------------------------------------------------------------------------|
| type | string | string | string | string |
| details | pilot mechanical pencil progrex h-127 - 0.7 mm | office supplies | scary halloween skull mask | pencil |
| superior drawing marker -pen - set of 12 colors - 2 nib | superior | coloring and writing book 21 x 29.7 cm 100 gsm 18 pages number subtraction ma4014 | marker |
| first person singular author: haruki murakami | haruki murakami book | buried secrets | literature and fiction |
* Loss: [MultipleNegativesSymmetricRankingLoss](https://sbert.net/docs/package_reference/sentence_transformer/losses.html#multiplenegativessymmetricrankingloss) with these parameters:
```json
{
"scale": 20.0,
"similarity_fct": "cos_sim",
"gather_across_devices": false
}
```
### Training Hyperparameters
#### Non-Default Hyperparameters
- `eval_strategy`: steps
- `per_device_train_batch_size`: 256
- `per_device_eval_batch_size`: 256
- `learning_rate`: 2e-05
- `weight_decay`: 0.01
- `num_train_epochs`: 6
- `warmup_ratio`: 0.2
- `fp16`: True
- `dataloader_num_workers`: 1
- `dataloader_prefetch_factor`: 2
- `dataloader_persistent_workers`: True
- `push_to_hub`: True
- `hub_model_id`: LamaDiab/v2MiniLM-V18Data-256ConstantBATCH-SemanticEngine
- `hub_strategy`: all_checkpoints
#### All Hyperparameters