UNICC and EPFL Lab Advance Responsible AI Through Joint Evaluation of LLM Safety

The collaboration delivers a practical evaluation framework that can be adapted to assess the safety and reliability of AI systems.

20 July, 2026

...
Credit: UNICC

GENEVA – 20 July 2026

Following the AI for Good Global Summit 2026, the United Nations International Computing Centre (UNICC) and the Machine Learning and Optimization Laboratory of EPFL ( École Polytechnique Fédérale de Lausanne) are pleased to announce the publication of a new white paper as a first outcome of their joint efforts to advance responsible AI. Facilitated by the International Computation and AI Network (ICAIN), the publication presents a practical framework for evaluating large language models (LLMs) to support their safe, reliable, and responsible use in institutional settings.

The collaboration under ICAIN brings together UNICC’s operational expertise in digital foundations for the UN system and EPFL’s leading research capabilities in artificial intelligence, reflecting a shared commitment to advancing trustworthy AI through rigorous research, practical testing and knowledge sharing.

The newly released white paper “Safety Evaluation of an Institutional LLM-RAG Deployment: A Four-layer AI Safety Audit Framework” presents a structured approach to assessing the behaviour of Apertus, an open-source large language model developed by the Swiss AI Initiative. Rather than focusing solely on the model’s technical performance, the research examines how an AI system behaves when deployed in a real institutional environment where accuracy, safety, and reliability are essential.

“The project evaluated the system across multiple dimensions, including its resistance to harmful prompts, its ability to recognize when it should not answer, the accuracy of its responses when using external knowledge, and its susceptibility to biased or misleading interactions”, explained Associate Professor Martin Jaggi, Head of the Machine Learning and Optimization Laboratory of EPFL.

Beyond the findings presented in the white paper, a practical evaluation framework was developed that the public sector, the UN system, and other international organizations can adapt to assess the safety and reliability of their AI systems.

“This collaboration shows what international organizations and academic institutions can achieve together”, said Anusha Dandapani, Chief of the UNICC AI Hub. “UNICC brings the operational context, while EPFL brings research rigour. Together, we have evaluated open-weight models in practice rather than in theory. That is what building institutional capacity for responsible AI looks like, and the knowledge generated extends well beyond this partnership”.

This publication forms part of UNICC’s broader efforts to foster innovation through partnerships with academia, research institutions and technology communities; accelerating learning, promoting responsible innovation and strengthening the digital capabilities available to the UN system and other international organizations.

As artificial intelligence becomes increasingly integrated into public sector operations, initiatives such as this demonstrate the importance of combining scientific research with operational experience to help ensure AI technologies remain safe and trustworthy.

This publication was made possible through financial support provided by the Republic and Canton of Geneva to the UNICC AI Hub.

About the white paper

The white paper, Safety Evaluation of an Institutional LLM-RAG Deployment: A Four-Layer Audit of the UNICC Apertus System, documents the joint evaluation framework developed by UNICC and EPFL to assess the safety and robustness of institutional AI systems. In addition to presenting the findings, it introduces practical methodologies that can support future AI assurance efforts across the UN system and beyond.

Read the full white paper here.

For media inquiries, please contact: [email protected].