Evaluation of Contextual Understanding in Large Language Models
Evaluation of Contextual Understanding in Large Language Models. It centres on Benchmarks, and also names Perplexity. Reported by arXiv. Bharat Hunt files it under AI Research and AI Models — the section covering papers, benchmarks, evaluations, interpretability and safety results.