Evaluating the Diversity of AI-Generated Content with Diversity Profiles
Diversity is a fundamental criterion for evaluating generative artificial intelligence (AI) systems, yet its measurement remains inherently ambiguous. Existing approaches typically represent generated samples in an embedding space, compute pairwise distances or similarities, and aggregate them into a single scalar score. Such scalar summaries are convenient, but they often encode different…
We haven't written up this one. arXiv cs.AI has the full story — the link below goes straight to it.