LLMs1 min read
Weakly Supervised Dataset Extraction Framework Using LLMs
Researchers developed a framework to identify dataset mentions in forced displacement and FCV documents. The system uses a lightweight model and a large language model to refine labels, achieving high precision and recall on a benchmark of 1,706 text passages.
From arXiv cs.CL
