Michael Min Wah Leung
Notes on post-training, sequence modelling, and the occasional brain. Code-first, results-honest, written by an ML engineer with a research background in neuroscience.
Writing
Readable shallow, steerable deep: the same split in three transformers
interpretability
steering
probing
transformers
Does a single direction mediate refusal? A small reproduction
interpretability
safety
steering
transformers
Patient-specific filters as biomarkers
neuroscience
signal-processing
interpretability
No matching items