David Aylward

AI tinkerer in South Africa. I take language models apart on two 8GB gaming GPUs to see what's inside, then put smaller ones back together. Everything I break and everything I learn ships public: weights, measurements, mistakes and all.

Current projects

Qwen3.8-Whittle-16B A 27B whittled to 16.8B with a logit lens and a pricing table, healed with one A100 evening. 36/39 on the field battery at 20 tok/s on consumer GPUs. Research preview, v2.
The un-repaired cut The raw surgery artifact: full measurement history, every pricing run, every script. The damage profile is the science.

How this works

No lab, no cluster. Measurements run on a Ryzen 2700X with an RTX 4060 and an RTX 3050. Training runs on rented A100 hours. Every experiment publishes its wins, its bugs, and its dead ends on the model cards themselves.

Support the tinkering

☕ ko-fi.com/davida81328

Every donation becomes A100 hours, and every A100 hour ends up as a public model or a public measurement.

Downloads last month

-

Downloads are not tracked for this model. How to track
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support