Ending AI Slop — Thais Castello Branco, Taste Labs
For subjective agent output, replace vague requests for quality with decomposed brand constraints, then reserve human preference data for style and creativity that resist deterministic checks.
Subjective quality depends on audience, context, and time. Brand adherence becomes more testable when split into **colors, typography, motion, and textures**; Taste Labs says its contributor community includes **over 1,000 experts** across media and styles.
Give coding agents explicit brand components instead of asking for something generally good. Verify alignment and typography directly, while treating style fit and creativity as preference problems that need carefully selected human data.
Subjective quality depends on audience, context, and time. Brand adherence becomes more testable when split into **colors, typography, motion, and textures**; Taste Labs says its contributor community includes **over 1,000 experts** across media and styles. Give coding agents explicit brand components instead of asking for something generally good. Verify alignment and typography directly, while treating style fit and creativity as preference problems that need carefully selected human data. An LLM judge can hallucinate or invite reward hacking, but expert consensus is not universally reliable either. Disagreement about aesthetics may represent valid preferences rather than bad labels, so averaging judgments can erase useful distinctions.
This turns interface quality from a single taste score into a mixed evaluation problem: some brand properties can be checked directly, while style and creativity require preference data that preserves legitimate disagreement. It supports encoding design guidance for coding agents, but warns that neither an LLM judge nor averaged expert consensus is a sufficient quality oracle.