How We Built a Data Pipeline: From Raw 2.8M Messy Tweets to Training Data
The Problem
We wanted to build a customer support agent powered by real conversation data. We found a dataset of 2.8 million Twitter support conversations on Kaggle. Sounds perfect, right?
Not exactly
blog.realdev.club3 min read