Neural Transparency: A Window into AI Before Using It
A team from the MIT Media Lab has developed 'neural transparency,' a technique that translates a neural network's internal activations into intuitive visualizations so users can anticipate traits like empathy, toxicity, or sycophancy before the chatbot speaks. The study reveals that people overestimate positive traits and underestimate negative ones when designing their personalized assistants.
