EADC: Evaluation of Advanced and Deep-level Compliance in Large Language Models
Large Language Models (LLMs) have been used in various industries. However, ensuring their compliance with complex laws and regulatory frameworks remains a great challenge. Existing evaluation paradigms mainly rely on static benchmarks that suffer from three severe limitations: First, the compliance rules being used do not comply with the requirements of Artificial Intelligence (AI) laws and…
We haven't written up this one. arXiv cs.AI has the full story — the link below goes straight to it.