Addressing benchmarking gaps in large language models for health and medicine with dynamic red-teaming
We subject a diverse panel of 15 state-of-the-art LLMs (both proprietary and open-source) to our DAS red-teaming protocol,…
Browsing Tag