Official data release to reproduce Confident Learning paper results
-
Updated
Feb 17, 2022 - Jupyter Notebook
Official data release to reproduce Confident Learning paper results
Use Large Language Models like OpenAI's GPT-3.5 for data annotation and model enhancement. This framework combines human expertise with LLMs, employs Iterative Active Learning for continuous improvement, and integrates CleanLab (Confident Learning) to ensure high-quality datasets and better model performance
Empirical investigation of when data quality remediation outperforms architectural upgrades in NLP text classification, across 135 experimental conditions spanning five corpora, three architectures, and four annotation cost tiers.
Add a description, image, and links to the confident-learning topic page so that developers can more easily learn about it.
To associate your repository with the confident-learning topic, visit your repo's landing page and select "manage topics."