Skip to main content
All Questions

Predict Harmful Text

Hard

Dataset

The dataset comprises a collection of tweets, each annotated to indicate whether it includes harmful content. The label '1' signifies harmful content, while '0' denotes content that is not harmful. To proceed, download the dataset and employ it within your .ipynb (Jupyter Notebook) environment to train and refine your model.

Here is a preview of the dataset structure:

  • The 'Text' column contains the tweet text.
  • The 'Target' column contains the label, where '1' corresponds to "harmful" and '0' corresponds to "not harmful".
IndexTextTarget
0@user #cnn calls #michigan middle school 'build the wall' chant "#tcot1
1it's unbelievable that in the 21st century we'd need something like this. #neverump #xenophobia1
2bihday your majesty0
3#model i love u take with u all the time in ur0
4we won!!! love the land!!! #allin #cavs #champions #cleveland #clevelandcavaliers0