Models for the paper Cottention: Linear Transformers With Cosine Attention https://arxiv.org/abs/2409.18747
Gabriel Mongaras
gmongaras
AI & ML interests
None yet
Recent Activity
updated
a model
about 1 month ago
gmongaras/Softmax_Attention_BERT
updated
a model
about 1 month ago
gmongaras/Cosine_Attention_BERT
updated
a model
about 1 month ago
gmongaras/Cosine_Attention_GPT_1.2B
Organizations
Collections
5
Papers
1
models
18
gmongaras/Softmax_Attention_BERT
Feature Extraction
•
Updated
•
6
gmongaras/Cosine_Attention_BERT
Feature Extraction
•
Updated
•
6
gmongaras/Cosine_Attention_GPT_1.2B
Feature Extraction
•
Updated
•
5
gmongaras/Cosine_Attention_GPT_300M
Feature Extraction
•
Updated
•
2
gmongaras/Softmax_Attention_GPT_1.2B
Feature Extraction
•
Updated
•
3
gmongaras/Softmax_Attention_GPT_300M
Feature Extraction
•
Updated
•
3
gmongaras/Yann_UWU
Text Generation
•
Updated
•
10
gmongaras/Meta-Llama-3.1-8B
Text Generation
•
Updated
•
11
gmongaras/reddit_negative_v1_13B
Text Generation
•
Updated
•
11
•
1
gmongaras/Wizard_7B_Squad_v2
Text Generation
•
Updated
•
12
datasets
21
gmongaras/Elon_Tweets_Score
Viewer
•
Updated
•
5.9k
•
35
gmongaras/Elon_Tweets
Viewer
•
Updated
•
5.9k
•
44
gmongaras/Pile_Llama_Tokenized
Updated
•
9
gmongaras/Anime_Subtitle_data2
Viewer
•
Updated
•
1.91M
•
47
gmongaras/Anime_Subtitle_data
Viewer
•
Updated
•
14.6M
•
41
gmongaras/BERT_Base_Cased_128_Dataset_Mapped
Viewer
•
Updated
•
132M
•
814
gmongaras/BERT_Base_Cased_128_Dataset
Viewer
•
Updated
•
134M
•
262
gmongaras/Yann_LeCun_Tweets
Viewer
•
Updated
•
406
•
49
gmongaras/dummy_text_dataset
Viewer
•
Updated
•
2.05k
•
86
gmongaras/EleutherAI_the_pile_deduplicated
Viewer
•
Updated
•
134M
•
210
•
2