SEEK: Skill-Routed Evaluation with Evolvable Knowledge for Industrial Search
Search quality evaluation provides essential supervision and diagnostic signals for the development and iteration of industrial search systems. Although large language models (LLMs) offer a scalable alternative to manual assessment, reliable automatic evaluation remains challenging: users experience search results at the page level, while the applicable evaluation criteria are multi-dimensional…
We haven't written up this one. arXiv cs.AI has the full story — the link below goes straight to it.