I remember a tweet from @rmcelreath.bsky.social saying he sometimes modeled categorical effects as hierarchical (I assume drawn from a normal with shared scale). I reckon modeling them this way, with a sum-to-zero constraint to retain a nice intercept, *has* to be the most principled strategy.