Kwanghee Choi @juice500ml.bsky.social · 16/03/2026📄 Paper 1 (submitted to Jan ARR): "[b] = [d] − [t] + [p]: Self-supervised Speech Models Discover Phonological Vector Arithmetic" We show how phone(me)s are encoded in S3Ms: as a linear combination of phonological feature vectors. arxiv.org/abs/2602.18899 100
Kwanghee Choi @juice500ml.bsky.social · 16/03/2026📄 Paper 2 (submitted to IS): "Self-Supervised Speech Models Encode Phonetic Context via Position-dependent Orthogonal Subspaces" We further show how sequences of phone(me)s can be encoded, i.e., contextualize, in a single S3M frame. arxiv.org/abs/2603.12642 110
Kwanghee Choi @juice500ml.bsky.social · 16/03/2026🧵 Together, both papers take a step beyond the usual "what info do S3Ms encode" probing paradigm. We aim to answer how is that info actually encoded geometrically? Come see for yourself Thursday! 👀 Slides: docs.google.com/presentation...docs.google.comSelf-supervised Speech Models are Phonological Vector MachinesSelf-supervised Speech Models are Phonological Vector Machines Kwanghee Choi kwanghee@utexas.edu 100