Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
thinletter
/
harrier-0.6b-vq-clients
Like
0
Follow
thinletter
1
Feature Extraction
English
embeddings
retrieval
quantization
vector-quantization
webgpu
query-encoder
asymmetric-retrieval
License:
mit
Model card
Files
Files and versions
xet
Community
3
Copy to bucket
new
New discussion
New pull request
Resources
PR & discussions documentation
Code of Conduct
Hub documentation
All
Discussions
Pull requests
View closed (1)
Sort:Β Recently created
The query prompt's K/V in full precision: +0.01β0.04 nDCG@10 at 1.6β1.8 bits per weight for 2 MB of data (measured, not yet in the files)
2
#3 opened 1 day ago by
honza-rosecky
Vector-quantised harrier-0.6b query clients at 1.8β2.1 bits per weight (105β127 MiB) β summary and links
#2 opened 1 day ago by
honza-rosecky