#safety fine-tuning
Safety Fine-Tuning: 1 AI articles covering safety fine-tuning news, analysis, and research
Articles
Benchmarking Argumentative Behaviour of LLMs: A Study ofNEW⭐9
A study benchmarking how large language models use and defend against character attacks in political debate, comparing their argumentative strategies with human...
