Models
Datasets
Spaces
Posts
Docs
Enterprise
Pricing
Log In
Sign Up

Collections

Discover the best community collections!

Collections including paper arxiv:2307.08621

fla-hub/retnet-1.3B-100B

Text Generation • Updated Aug 26 • 73
fla-hub/retnet-2.7B-100B

Text Generation • Updated Aug 26 • 7 • 1
Retentive Network: A Successor to Transformer for Large Language Models

Paper • 2307.08621 • Published Jul 17, 2023 • 170

abacusai/Smaug-72B-v0.1

Text Generation • Updated Feb 23 • 2.86k • 467
Running on A10G

826

📚

ReplaceAnything
miqudev/miqu-1-70b

Updated Feb 4 • 6.13k • 979
fka/awesome-chatgpt-prompts

Viewer • Updated Sep 3 • 170 • 6.64k • 6.63k

Running on A100

897

🪁

Zephyr Chat
Retentive Network: A Successor to Transformer for Large Language Models

Paper • 2307.08621 • Published Jul 17, 2023 • 170
fka/awesome-chatgpt-prompts

Viewer • Updated Sep 3 • 170 • 6.64k • 6.63k

Generative AI Family

The collection of generative ai with research papers, datasets and models

Attention Is All You Need

Paper • 1706.03762 • Published Jun 12, 2017 • 49
Training Generative Adversarial Networks with Limited Data

Paper • 2006.06676 • Published Jun 11, 2020
A survey of Generative AI Applications

Paper • 2306.02781 • Published Jun 5, 2023
Running on CPU Upgrade

10.8k

🔥

Stable Diffusion 2-1

LLM architecture

The Impact of Depth and Width on Transformer Language Model Generalization

Paper • 2310.19956 • Published Oct 30, 2023 • 9
Retentive Network: A Successor to Transformer for Large Language Models

Paper • 2307.08621 • Published Jul 17, 2023 • 170
RWKV: Reinventing RNNs for the Transformer Era

Paper • 2305.13048 • Published May 22, 2023 • 15
Attention Is All You Need

Paper • 1706.03762 • Published Jun 12, 2017 • 49

Detecting Pretraining Data from Large Language Models

Paper • 2310.16789 • Published Oct 25, 2023 • 10
Let's Synthesize Step by Step: Iterative Dataset Synthesis with Large Language Models by Extrapolating Errors from Small Models

Paper • 2310.13671 • Published Oct 20, 2023 • 18
AutoMix: Automatically Mixing Language Models

Paper • 2310.12963 • Published Oct 19, 2023 • 14
An Emulator for Fine-Tuning Large Language Models using Small Language Models

Paper • 2310.12962 • Published Oct 19, 2023 • 14

Retentive Network: A Successor to Transformer for Large Language Models

Paper • 2307.08621 • Published Jul 17, 2023 • 170

Interesting Musings

Retentive Network: A Successor to Transformer for Large Language Models

Paper • 2307.08621 • Published Jul 17, 2023 • 170
TabLib: A Dataset of 627M Tables with Context

Paper • 2310.07875 • Published Oct 11, 2023 • 8

Stuff I (TheProjectsGuy) have summarized (for time pass). Mostly papers. I do not guarantee that the summaries are fully correct (as I am no expert).

SIMPL: A Simple and Efficient Multi-agent Motion Prediction Baseline for Autonomous Driving

Paper • 2402.02519 • Published Feb 4
Mixtral of Experts

Paper • 2401.04088 • Published Jan 8 • 158
Optimal Transport Aggregation for Visual Place Recognition

Paper • 2311.15937 • Published Nov 27, 2023
GOAT: GO to Any Thing

Paper • 2311.06430 • Published Nov 10, 2023 • 14

Retentive Network: A Successor to Transformer for Large Language Models

Paper • 2307.08621 • Published Jul 17, 2023 • 170
Llama 2: Open Foundation and Fine-Tuned Chat Models

Paper • 2307.09288 • Published Jul 18, 2023 • 243

Previous
1
2
Next

Company

TOS Privacy About Jobs

Website

Models Datasets Spaces Pricing Docs