Determine the effect of fine-tuning on SafeAlert evaluation performance
Determine whether fine-tuning produces different results from prompt-level interventions for the generation-safety and Nigerian fintech communication-classification tasks evaluated by SafeAlert.
References
The three conditions tested here are prompt-level interventions; whether fine-tuning produces different results is a separate question this evaluation does not address.
— Context-Aware Pre-Deployment Evaluation of AI Systems: A Regulatory Framework for Nigerian Fintech
(2609.24016 - Uduimoh et al., 21 Sep 2026) in Section 4, subsection “Limitations”