UX Research · AI Fairness

WeAudit TAIGA

Redesigning the forum and engagement features of TAIGA - a crowdsourced platform for auditing AI-generated images for bias - to increase user participation and make audit contributions more meaningful.

My Role
UX Researcher & Designer
Platform
WeAudit / TAIGA
Methods
Survey Analysis · Heuristic Evaluation · Ideation
Tools
Figma · FigJam

Crowdsourcing AI fairness - and making it worth participating in.

WeAudit is a platform that harnesses the crowd to improve machine learning and algorithmic fairness. TAIGA (Tool for Auditing Images Generated by AI) is one of its core tools - users explore AI-generated image prompts, write audit reports on the biases they observe, and share their findings in a community forum.

The three-step flow - Explore Prompts → Author an audit report → Share and discuss - is simple. But participation was uneven and the forum lacked the engagement features needed to sustain a community around bias auditing. Our work focused on understanding who was participating, what they needed, and redesigning the platform to support broader, more sustained engagement.

"To audit AI fairly, you need a diverse crowd. The design of the platform itself shapes who shows up and who stays."
TAIGA platform - how it works

TAIGA: Explore prompts → Author an AI audit report → Share your report and discuss.

Who participates, and what shapes their perception of bias?

We analyzed survey data from existing TAIGA users to understand how gender, familiarity with algorithmic systems, and bias category affected how people perceived and rated AI-generated images.

Survey Data Analysis Sentiment Analysis Affinity Diagramming Heuristic Evaluation

Finding 01 - Gender shapes harm perception

Males were significantly more likely to rate AI-generated images as not harmful. Females and LGBTQ+ participants were more likely to rate the same images as harmful - a consistent gap that suggests the auditing crowd's composition directly affects what gets flagged as a problem.

Gender vs harm perception chart

Males more likely to rate images not harmful; females and LGBTQ+ more likely to rate them harmful.

Finding 02 - Algorithmic literacy correlates with bias awareness

Users familiar with algorithmic systems were 94% aware of societal bias issues. Even users who were not familiar showed strong awareness (70%) - suggesting participation in auditing builds understanding regardless of prior knowledge.

Algorithmic familiarity vs bias awareness

Familiar users: 94% aware · Neither familiar: 76% aware · Not familiar: 70% aware.

Finding 03 - Bias category affects sentiment differently

Sentiment scores varied significantly across bias categories. Sexuality-related prompts produced the highest scores (0.095), while gender bias prompts scored lowest (0.044) - revealing that different types of bias provoke different emotional responses in auditors.

Sentiment scores by bias category

Sentiment scores: Neutral 0.084 · Gender 0.044 · Sexuality 0.095 · Race 0.047.

Synthesis

Diversity in the auditing crowd is a design problem

Who participates shapes what gets identified as bias. A platform designed only for technically-savvy users systematically underweights the perspectives of people who experience bias most directly. Broadening participation isn't just a community goal - it's a fairness requirement.

Evaluating TAIGA against Nielsen's 10 heuristics

Our team conducted a structured heuristic evaluation of the TAIGA platform, rating each of Nielsen's 10 usability heuristics by the number of follows and violations found - then weighting each by severity. The evaluation revealed a platform that handles system feedback and error prevention well, but struggles significantly with user control, consistency, and navigational clarity.

Heuristic
Follows
Violations
Weight
Visibility of System Status
5
1
+4
Match Between System and Real World
0
2
−2
User Control and Freedom
0
2
−2
Consistency and Standards
0
2
−2
Error Prevention
2
0
+2
Recognition Rather than Recall
1
0
+1
Flexibility and Efficiency of Use
0
0
0
Aesthetic and Minimalist Design
0
1
−1
Help Users Recognize & Recover from Errors
1
1
0
Help and Documentation
2
3
−1
Strongest area

Visibility of System Status (+4)

Loading indicators, button feedback, and progress cues consistently inform users what the system is doing - including a "Generating from Stable Diffusion... Don't look away!" message with a loading circle while images generate.

Weakest areas

User Control, Consistency, and Real-World Match (−2 each)

No way to undo a submitted post. Inconsistent button wording ("Create Thread" vs "+ New Thread"). Jargon like "Stable Diffusion" and "Google mode" unfamiliar to non-technical users. Two search bars with no explanation of the difference.

Six Prioritized Issues

Beyond the heuristic framework, we identified six specific usability issues rated by frequency, impact, and persistence.

Positive

#1 Visual Design of the Interface

The front page is simple and easy to navigate - clear layout, organized buttons, and numbered directions help first-time users understand TAIGA's features and direction for use. Consistently helpful for new users.

Heuristic 1 - Visual Design
Positive

#2 Concise and Relevant Information

The "Show Example" toggle brings up relevant, concise prompt examples that help users understand what to generate. Common enough to have a significant impact on user experience, and easy to act on repeatedly.

Heuristic 2 - Concise information
Positive

#3 Consistent Prompt History

The prompt history bar efficiently retrieves previous prompts and image results - formatted in a timely, intuitive manner. Users regularly rely on it while trying new prompts and generating posts, and can favorite or pin prompts for quicker access.

Heuristic 3 - Prompt History
Major

#4 Long Form Discourages Completion

The audit report form is too long and the text boxes are too large - users report feeling overwhelmed after thinking through too much information at once. Shortening text boxes or converting to a multi-step form would significantly improve completion rates.

Heuristic 4 - Long form
Major

#5 No Search Bar Parameters

When asked to "insert prompt here," users interpret the directions as inputting an entire phrase like "show me politicians" - unaware that just the subject should be entered. Adding example adjectives or nouns after "Insert prompt here" would prime users on the correct format.

Heuristic 5 - No search parameters
Major

#6 Selected Dropdown Not Displayed

After selecting a "Types of harms" dropdown option, the selected value isn't shown in the box - causing users to think the system didn't register their response. Displaying the selected option in the white box would eliminate this persistent point of confusion.

Heuristic 6 - Dropdown display

Three directions for enhancing the TAIGA platform

We used a Why/How ladder to frame the design space across three strategic directions - each addressing a different root problem identified in the research.

Why/How ladder diagram

Why/How ladder: three directions for enhancing TAIGA - trust and inclusivity, user empowerment, and interface clarity.

Direction 01

Trustworthy, inclusive, equitable

Conduct regular audits and involve a wide range of users in design and testing phases. Establish user panels for platform review and improvement recommendations - making the platform itself accountable to the community it serves.

Direction 02

Empower everyday users

Implement interactive tutorials and resources that educate users about AI bias. Integrate a system where user feedback on biases is directly reflected in AI training and model updates - giving participants a visible stake in outcomes.

Direction 03

More visually intuitive and clear

Enhance platform design with accessibility features - screen reader support, alternative text for images, unambiguous icons, and consistent visual hierarchy. Lower the barrier to entry for users unfamiliar with algorithmic systems.

Crazy 8s - Forum Feature Sketches

We ran Crazy 8s ideation sessions focused on the forum and social engagement layer - the part of TAIGA where findings get shared, discussed, and built upon.

Crazy 8s sketches

Crazy 8s sketches exploring richer forum interactions, live reactions, and content categorization.

Forum features that reward contribution and sustain engagement

The redesign focused on enriching the forum experience with features that give users more ways to interact with findings, recognize active contributors, and make participation feel meaningful over time.

Six Feature Areas

Live interaction

Real-time reactions on audit posts - hearts, thumbs up - making the forum feel active and responsive rather than static.

Share findings with others

Easy ways to highlight and share specific bias findings with the broader WeAudit community and beyond.

Recognition for active users

Visible indicators of contribution level - badges, levels, and audit counts - that acknowledge users who participate consistently.

Incentive to contribute

Rewards tied to participation quality - encouraging users to post findings, engage with others' reports, and return over time.

Nuanced interactions

More diverse interaction tools - save, upvote, downvote - giving users a fuller range of ways to signal agreement, disagreement, or interest.

Categorisation of posts

Tag and filter audit reports by bias type (gender, race, sexuality) so users can navigate findings by the areas most relevant to them.

Design concept cards

Six feature concept cards generated from Crazy 8s and research synthesis.

Gamification - Audit Bits and Quests

A lightweight gamification layer rewards audit quality through "Audit Bits" - an in-platform currency earned by finding biases that receive community upvotes. Users level up through auditing milestones and unlock quests like "Post 3 examples of gender bias in Meta's latest model."

Gamification - Audit Bits profile

Profile screen: Level 3 user with 300 Audit Bits and active quests - completed and in-progress.

WeAudit forum

The parent WeAudit platform provides the community layer - a forum where audit reports become discussion threads, users follow topics, and findings accumulate into a shared knowledge base about AI bias patterns.

WeAudit forum

WeAudit forum: crowdsourced bias findings become community discussions.

Affinity Synthesis - Walking the Wall

Walking the wall affinity diagram

Module 5 deliverable: team affinity synthesis across participation, trust, interaction, and interface themes.