Rosenverse
[Demo] How to re-categorize content at scale using LLMs

This video is only accessible to Gold members. Log in or register for a free Gold Trial Account to watch.

Log in Register

Most conference talks are accessible to Gold members, while community videos are generally available to all logged-in members.

[Demo] How to re-categorize content at scale using LLMs

Gold
Wednesday, June 5, 2024 • Designing with AI 2024
Share the love for this talk
[Demo] How to re-categorize content at scale using LLMs
Speakers: Jorge Arango
Link:

Summary

Large Language Models (LLMs) are to language as spreadsheets are to numbers: tools for modeling, exploration, and development. Among their many capabilities, LLMs can alleviate chores related to the design and implementation of information architectures. But doing so requires venturing beyond chat-based interfaces. In this brief demonstration, we'll see how to use OpenAI's API and a few open source command line tools to re-categorize content in a 1,000+ page website. The techniques demonstrated can be extended to other common content organization tasks.

Key Insights

  • Manual retagging of 1,200 blog posts would take about 10 hours, but leveraging GPT-4 reduced active human time to about 2 hours.

  • Using GPT-4 via command line and shell scripts enables automated tagging outside typical chat interfaces.

  • An organically grown taxonomy over 20 years contained unclear acronyms and inconsistent tag forms that GPT initially struggled with.

  • Cleaning and standardizing the taxonomy before prompting GPT is critical for effective AI assistance.

  • A review step of AI-suggested tags in CSV format allows human correction to avoid hallucinations entering production.

  • GPT-4 can propose new and useful tags outside the original taxonomy, enriching content classification.

  • The four-step GRU framework (Gather, Review, Update, Wrap up) balances automation with human oversight.

  • Storing blog content as markdown files simplifies integrating AI workflows via scripting and file manipulation.

  • The approach is adaptable and scalable to other CMS platforms by replacing scripting with API calls.

  • Taxonomies should use clear, unambiguous terms to improve both human and AI understanding.

Notable Quotes

"Some of the older content has discoverability problems, which is typical with blogs."

"Doing this tagging manually would have taken me around 10 hours of mind-numbing work."

"I’m actually using GPT-4, but not via the chat interface—I'm calling it from the Mac’s command line."

"I had to clean the taxonomy up because GPT wouldn’t know what to do with acronyms like TAOI."

"I save the proposed tags to a CSV file so I can preview and edit them before applying the changes."

"A middle review step prevents hallucinations from making it into the production site."

"GPT-4 functioned as an assistant not just in retagging but also in improving the taxonomy itself."

"The entire process took about three hours from start to finish, about a fifth of the manual time."

"Use clear and obvious terms in taxonomies—unusual acronyms won’t make sense to GPT or others."

"You need to review proposed changes before committing them to production, otherwise errors sneak in."

Ask the Rosenbot
Caroline Jarrett
Have fun with statistics?
2024 • Rosenfeld Community
Uday Gajendar
Making the Invisible Visible: The most critical deliverable isn't always the design
2026 • Rosenfeld Community
Spencer L. A. Stultz
Why Social Justice Frameworks are Necessary for Successful DEI/JEDI Initiatives
2023 • DesignOps Summit 2023
Gold
Anil Dash
Designing with power in the age of AI
2026 • Designing with AI 2026
Conference
Ryan Matthew
DesignOps without Boundaries: Building More with What You Have
2025 • DesignOps Summit 2025
Gold
Chris Moses
Stretching the Definition of DesignOps with Product Development
2018 • DesignOps Summit 2018
Gold
Dan Willis
Theme 3: Intro
2024 • Enterprise Experience 2020
Gold
Nathan Shedroff
Double Your Mileage: Use Your Research Strategically
2020 • Advancing Research 2020
Gold
Theme 3 Intro
2022 • Advancing Research 2022
Gold
Sam Proulx
SUS: A System Unusable for Twenty Percent of the Population
2021 • Civic Design 2021
Gold
Mac Smith
Measuring Up: Using Product Research for Organizational Impact
2021 • Advancing Research 2021
Gold
Luca Rager
Empowering Gaming at Scale: How Xbox Builds Powerful, Automated, and Distributed Design Systems with Sketch
2021 • DesignOps Summit 2021
Gold
Dorelle Rabinowitz
The Magic Word is Trust
2018 • Enterprise Experience 2018
Gold
Melinda Belcher
Bridging the Gap: Making the Most of the Differences Between Agency and Enterprise
2024 • Enterprise Experience 2020
Gold
Eric Shumake
An AMA on UX's Role in Healthcare
2026 • Rosenfeld Community
Louis Rosenfeld
Coffee with Lou #3: What Makes for a Successful UX Conference Presentation?
2024 • Rosenfeld Community

More Videos

Jayne Engle

"Sacred civics requires us to hold higher order accountabilities to ecosystems and future generations."

Jayne Engle Tanya Chung-Tiam-Fook

Civic Design for the Next Seven Generations—A Discussion on Sacred Civics

August 25, 2022

Marc Fonteijn

"Remote work is less frequent now than it was two years ago among service designers, except for freelancers who mostly work remotely."

Marc Fonteijn

First Insights from the 2025 Service Design Salary(+) Report

December 4, 2024

Steve Portigal

"I’ve started giving up on product management and focusing more on product strategy and senior executives."

Steve Portigal Alba Villamil Sam Ladner

The Future of Research: Bridging the Gaps

July 29, 2021

Tricia Wang

"Web3 offers a unique chance to get involved early before some of the ethical challenges of Web2 take root."

Tricia Wang

The most popular design thinking strategy is BS

January 27, 2022

Farid Sabitov

"Design operations maturity has five levels: awareness, standardized, integrated, accelerated, and innovated."

Farid Sabitov

Theme Four Intro

September 9, 2022

Elena Naids

"Government culture is not like tech culture; IT constraints mean tools like Sketch are still a win."

Elena Naids Liza McRuer

The Power of Difficult Conversations: A Case Study on How We Introduced Design Ops in the Federal Government Space

October 2, 2023

Kristin Wisnewski

"We’re the voice of the employee—we’re arbiters of truth, defenders of experience, and sometimes validators against manipulation."

Kristin Wisnewski

Measuring What Matters

October 23, 2019

Caitlyn Hampton

"Compass gave me a one-month runway to transition from content strategist to product designer, and that support was incredible."

Caitlyn Hampton Monica Lee Jina Yoon

Compass 101: Growing Your Career In A Startup World

June 11, 2021

JD Buckley

"Conducting the first benchmark study always leaves you questioning, and the first comparison study is less of a big bang and more likely to be a subtle sigh of relief."

JD Buckley

Communicating the ROI of UX within a large enterprise and out on the streets

June 14, 2018