FYDP_Dataset

Published: 12 July 2026| Version 1 | DOI: 10.17632/frbkwttk6c.1
Contributor:
Shadab Yasar Aronno

Description

Final Balanced 8K Curated Dataset is a curated text classification dataset developed for the Final Year Design Project (FYDP). The dataset contains 8,000 balanced text samples collected from multiple sources and organized for supervised machine learning research. Each record includes the text content, corresponding class label, source information, and title (where available). The dataset was cleaned, curated, and balanced to reduce class imbalance and improve the reliability of model training and evaluation. It is intended for research on natural language processing (NLP), text classification, and artificial intelligence. This dataset can be used for training, validation, and benchmarking machine learning and deep learning models while supporting reproducible academic research.

Files

Institutions

Categories

Artificial Intelligence, Natural Language Processing, Machine Learning, Large Language Model

Licence