inference.py
Model Overview
A xlarge-scale implementation of the mae architecture, built for retrieval tasks.
Architecture
- Architecture: mae
- Scale: xlarge
- Attention: multi query
- Fusion strategy: co attention
- Task head: retrieval
- Activation: swish
- Normalization: batchnorm
- Initialization: kaiming normal
Training
- Optimizer: lion
- LR scheduler: cosine
Files
inference.py— main artifact of this repository
License
See the license field above.
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support