inference.py
Model Overview
A nano-scale implementation of the deit architecture, built for matching tasks.
Architecture
- Architecture: deit
- Scale: nano
- Attention: flash
- Fusion strategy: co attention
- Task head: matching
- Activation: gelu
- Normalization: batchnorm
- Initialization: kaiming normal
Training
- Optimizer: adamw
- LR scheduler: step
Files
inference.py— main artifact of this repository
License
See the license field above.
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support