Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
👋
Open to Work
389.9
TFLOPS
Andrew Thompson
AndrewThompson1233
81
31
Follow
TULLUS's profile picture
Local-Axiom-AI's profile picture
bbkdevops's profile picture
11 followers
·
1 following
AI & ML interests
None yet
Recent Activity
new
activity
about 3 hours ago
jaswanthsanjay88/mara:
Bilinear pointer routing, schema scaling in MCU memory budgets, and coordinate indexing in sub-1M AFMs
new
activity
about 3 hours ago
He-Tag/smallm-125m:
The 18.6% vocabulary tax, vertical shortcut saturation, and multi-turn intent drift in 125M architectures
new
activity
about 3 hours ago
BrodyMakezAI/gpt-5horts:
Zero vocabulary tax, character-level sequence expansion, and immaculate Roblox priors in 10.7M architectures
View all activity
Organizations
None yet
AndrewThompson1233
's activity
All
Models
Datasets
Spaces
Buckets
Papers
Collections
Community
Posts
Upvotes
Likes
Articles
New activity in
jaswanthsanjay88/mara
about 3 hours ago
Bilinear pointer routing, schema scaling in MCU memory budgets, and coordinate indexing in sub-1M AFMs
#1 opened about 3 hours ago by
AndrewThompson1233
New activity in
He-Tag/smallm-125m
about 3 hours ago
The 18.6% vocabulary tax, vertical shortcut saturation, and multi-turn intent drift in 125M architectures
#1 opened about 3 hours ago by
AndrewThompson1233
New activity in
BrodyMakezAI/gpt-5horts
about 3 hours ago
Zero vocabulary tax, character-level sequence expansion, and immaculate Roblox priors in 10.7M architectures
2
#1 opened about 3 hours ago by
AndrewThompson1233
liked
a model
about 3 hours ago
superAVTR/tinizong-46M
Text Generation
•
46M
•
Updated
about 4 hours ago
•
1
New activity in
superAVTR/tinizong-46M
about 3 hours ago
Numerical validation rigor, the 17.8% vocabulary tax, and discourse-level perplexity cliffs in 46M SLMs
#1 opened about 3 hours ago by
AndrewThompson1233
New activity in
Timothyemmanuel/Arakandar
about 3 hours ago
Grounded tabular reasoning, numerical token preservation, and VRAM ceilings on 12GB hardware in 3B financial architectures
#1 opened about 3 hours ago by
AndrewThompson1233
New activity in
textilelabs/Loom-Weave-3
about 3 hours ago
Honest evaluation methodology, the 20% vocabulary tax, and unanswerability mechanics in 31M tool-augmented SLMs
❤️
1
#1 opened about 3 hours ago by
AndrewThompson1233
liked
3 models
about 3 hours ago
textilelabs/Loom-Weave-3
Text Generation
•
31.5M
•
Updated
about 4 hours ago
•
1
racer102/hal10
Updated
about 2 hours ago
•
1
aethertp/PicoLM-V2.1-81M-Instruct
Text Generation
•
96M
•
Updated
about 7 hours ago
•
2
liked
a model
about 4 hours ago
VADRK155/Cortex-2-Chat-Preview
Text Generation
•
Updated
1 day ago
•
130
•
2
New activity in
VADRK155/Cortex-2-Chat-Preview
about 4 hours ago
The 256-token context cliff and sequence scaling in custom 346M chat architectures
6
#1 opened about 15 hours ago by
AndrewThompson1233
New activity in
aethertp/PicoLM-80M-Instruct
about 6 hours ago
Overcoming the ARC-Easy floor and vocabulary budget on an 80M footprint
❤️
1
6
#1 opened 4 days ago by
AndrewThompson1233
New activity in
Arain119/sophia
about 7 hours ago
State attenuation across depth and long-context needle retention in 3:1 NoPE hybrids
2
#1 opened about 15 hours ago by
AndrewThompson1233
liked
a model
about 7 hours ago
aethertp/PicoLM-V2-81M-Instruct
Text Generation
•
96M
•
Updated
about 8 hours ago
•
2
New activity in
Local-Axiom-AI/Sabaki-Preview
about 10 hours ago
Runaway generation mechanics, rotary phase mismatch, and state dissipation in 3:1 KDA-MLA hybrids
2
#1 opened about 15 hours ago by
AndrewThompson1233
New activity in
altslate/JugnuLM-110M-R2plus
about 10 hours ago
Value residual scaling, the 25.8% vocabulary tax, and horizontal state grounding in deep-thin SLMs
2
#1 opened about 15 hours ago by
AndrewThompson1233
New activity in
Fantominsight/Yapper-M4-Flash-0256-GGUF
about 10 hours ago
Memory bandwidth saturation, KV-cache thermal throttling, and O(1) sequence scaling on edge devices
2
#1 opened about 15 hours ago by
AndrewThompson1233
New activity in
Vir007/continual-ai-2b
about 12 hours ago
Continuous associative fast-weights and the 512-token context horizon in lifelong learning
2
#1 opened about 16 hours ago by
AndrewThompson1233
liked
a model
about 13 hours ago
Aobangaming/lightning-105m
Text Generation
•
0.1B
•
Updated
about 7 hours ago
•
210
•
3
Load more