AnalysisAI ModelsJuly 15, 2026

LLM knowledge distillation papers on RAG, data distillation, detection

Researchers fine-tune LLaMA 3 (8B) as a cross-encoder for RAG reranking via knowledge distillation. Other proposals include a text dataset distillation framework to reduce corpora size, and a reference-based method to detect whether an LLM was trained on outputs from stronger third-party models.

4 sources