baumhaus.digital/Miscellanous/Symposia/AE55/Symposium on Moral and Legal AI Alignment/From 'Benevolence' to 'Nature': Moral Ordinals, Axiometry and Alignment of Values in Small Instruct Language Models
baumhaus.digital/Miscellanous/Symposia/AE55/Symposium on Moral and Legal AI Alignment/From 'Benevolence' to 'Nature': Moral Ordinals, Axiometry and Alignment of Values in Small Instruct Language Models/Introduction Kategorien erforschen
baumhaus.digital/Miscellanous/Symposia/AE55/Symposium on Moral and Legal AI Alignment/From 'Benevolence' to 'Nature': Moral Ordinals, Axiometry and Alignment of Values in Small Instruct Language Models/Introduction Kategorien erforschen
baumhaus.digital/Miscellanous/Symposia/AE55/Symposium on Moral and Legal AI Alignment/From 'Benevolence' to 'Nature': Moral Ordinals, Axiometry and Alignment of Values in Small Instruct Language Models/Introduction Audiotext hören
baumhaus.digital/Miscellanous/Symposia/AE55/Symposium on Moral and Legal AI Alignment/From 'Benevolence' to 'Nature': Moral Ordinals, Axiometry and Alignment of Values in Small Instruct Language Models/Introduction listen to audiotext
baumhaus.digital/Miscellanous/Symposia/AE55/Symposium on Moral and Legal AI Alignment/From 'Benevolence' to 'Nature': Moral Ordinals, Axiometry and Alignment of Values in Small Instruct Language Models/Introduction Memory spielen
baumhaus.digital/Miscellanous/Symposia/AE55/Symposium on Moral and Legal AI Alignment/From 'Benevolence' to 'Nature': Moral Ordinals, Axiometry and Alignment of Values in Small Instruct Language Models/Introduction play memory game with primer knot 2070
baumhaus.digital/Miscellanous/Symposia/AE55/Symposium on Moral and Legal AI Alignment/From 'Benevolence' to 'Nature': Moral Ordinals, Axiometry and Alignment of Values in Small Instruct Language Models/Alignment Kategorien erforschen
baumhaus.digital/Miscellanous/Symposia/AE55/Symposium on Moral and Legal AI Alignment/From 'Benevolence' to 'Nature': Moral Ordinals, Axiometry and Alignment of Values in Small Instruct Language Models/Alignment Kategorien erforschen
baumhaus.digital/Miscellanous/Symposia/AE55/Symposium on Moral and Legal AI Alignment/From 'Benevolence' to 'Nature': Moral Ordinals, Axiometry and Alignment of Values in Small Instruct Language Models/Alignment Audiotext hören
baumhaus.digital/Miscellanous/Symposia/AE55/Symposium on Moral and Legal AI Alignment/From 'Benevolence' to 'Nature': Moral Ordinals, Axiometry and Alignment of Values in Small Instruct Language Models/Alignment listen to audiotext
baumhaus.digital/Miscellanous/Symposia/AE55/Symposium on Moral and Legal AI Alignment/From 'Benevolence' to 'Nature': Moral Ordinals, Axiometry and Alignment of Values in Small Instruct Language Models/Alignment Memory spielen
baumhaus.digital/Miscellanous/Symposia/AE55/Symposium on Moral and Legal AI Alignment/From 'Benevolence' to 'Nature': Moral Ordinals, Axiometry and Alignment of Values in Small Instruct Language Models/Alignment play memory game with primer knot 2072
baumhaus.digital/Miscellanous/Symposia/AE55/Symposium on Moral and Legal AI Alignment/From 'Benevolence' to 'Nature': Moral Ordinals, Axiometry and Alignment of Values in Small Instruct Language Models/Method Kategorien erforschen
baumhaus.digital/Miscellanous/Symposia/AE55/Symposium on Moral and Legal AI Alignment/From 'Benevolence' to 'Nature': Moral Ordinals, Axiometry and Alignment of Values in Small Instruct Language Models/Method Kategorien erforschen
baumhaus.digital/Miscellanous/Symposia/AE55/Symposium on Moral and Legal AI Alignment/From 'Benevolence' to 'Nature': Moral Ordinals, Axiometry and Alignment of Values in Small Instruct Language Models/Method Audiotext hören
baumhaus.digital/Miscellanous/Symposia/AE55/Symposium on Moral and Legal AI Alignment/From 'Benevolence' to 'Nature': Moral Ordinals, Axiometry and Alignment of Values in Small Instruct Language Models/Method listen to audiotext
baumhaus.digital/Miscellanous/Symposia/AE55/Symposium on Moral and Legal AI Alignment/From 'Benevolence' to 'Nature': Moral Ordinals, Axiometry and Alignment of Values in Small Instruct Language Models/Method Memory spielen
baumhaus.digital/Miscellanous/Symposia/AE55/Symposium on Moral and Legal AI Alignment/From 'Benevolence' to 'Nature': Moral Ordinals, Axiometry and Alignment of Values in Small Instruct Language Models/Method play memory game with primer knot 2071
baumhaus.digital/Miscellanous/Symposia/AE55/Symposium on Moral and Legal AI Alignment/From 'Benevolence' to 'Nature': Moral Ordinals, Axiometry and Alignment of Values in Small Instruct Language Models/Pre-Alignment Results Kategorien erforschen
baumhaus.digital/Miscellanous/Symposia/AE55/Symposium on Moral and Legal AI Alignment/From 'Benevolence' to 'Nature': Moral Ordinals, Axiometry and Alignment of Values in Small Instruct Language Models/Pre-Alignment Results Kategorien erforschen
baumhaus.digital/Miscellanous/Symposia/AE55/Symposium on Moral and Legal AI Alignment/From 'Benevolence' to 'Nature': Moral Ordinals, Axiometry and Alignment of Values in Small Instruct Language Models/Pre-Alignment Results Audiotext hören
baumhaus.digital/Miscellanous/Symposia/AE55/Symposium on Moral and Legal AI Alignment/From 'Benevolence' to 'Nature': Moral Ordinals, Axiometry and Alignment of Values in Small Instruct Language Models/Pre-Alignment Results listen to audiotext
baumhaus.digital/Miscellanous/Symposia/AE55/Symposium on Moral and Legal AI Alignment/From 'Benevolence' to 'Nature': Moral Ordinals, Axiometry and Alignment of Values in Small Instruct Language Models/Pre-Alignment Results Memory spielen
baumhaus.digital/Miscellanous/Symposia/AE55/Symposium on Moral and Legal AI Alignment/From 'Benevolence' to 'Nature': Moral Ordinals, Axiometry and Alignment of Values in Small Instruct Language Models/Pre-Alignment Results play memory game with primer knot 2073
baumhaus.digital/Miscellanous/Symposia/AE55/Symposium on Moral and Legal AI Alignment/From 'Benevolence' to 'Nature': Moral Ordinals, Axiometry and Alignment of Values in Small Instruct Language Models/Discussion Kategorien erforschen
baumhaus.digital/Miscellanous/Symposia/AE55/Symposium on Moral and Legal AI Alignment/From 'Benevolence' to 'Nature': Moral Ordinals, Axiometry and Alignment of Values in Small Instruct Language Models/Discussion Kategorien erforschen
baumhaus.digital/Miscellanous/Symposia/AE55/Symposium on Moral and Legal AI Alignment/From 'Benevolence' to 'Nature': Moral Ordinals, Axiometry and Alignment of Values in Small Instruct Language Models/Discussion Audiotext hören
baumhaus.digital/Miscellanous/Symposia/AE55/Symposium on Moral and Legal AI Alignment/From 'Benevolence' to 'Nature': Moral Ordinals, Axiometry and Alignment of Values in Small Instruct Language Models/Discussion listen to audiotext
baumhaus.digital/Miscellanous/Symposia/AE55/Symposium on Moral and Legal AI Alignment/From 'Benevolence' to 'Nature': Moral Ordinals, Axiometry and Alignment of Values in Small Instruct Language Models/Discussion Memory spielen
baumhaus.digital/Miscellanous/Symposia/AE55/Symposium on Moral and Legal AI Alignment/From 'Benevolence' to 'Nature': Moral Ordinals, Axiometry and Alignment of Values in Small Instruct Language Models/Discussion play memory game with primer knot 2074
baumhaus.digital/Miscellanous/Symposia/AE55/Symposium on Moral and Legal AI Alignment/From 'Benevolence' to 'Nature': Moral Ordinals, Axiometry and Alignment of Values in Small Instruct Language Models/Post-Alignment Results Kategorien erforschen
baumhaus.digital/Miscellanous/Symposia/AE55/Symposium on Moral and Legal AI Alignment/From 'Benevolence' to 'Nature': Moral Ordinals, Axiometry and Alignment of Values in Small Instruct Language Models/Post-Alignment Results Kategorien erforschen
baumhaus.digital/Miscellanous/Symposia/AE55/Symposium on Moral and Legal AI Alignment/From 'Benevolence' to 'Nature': Moral Ordinals, Axiometry and Alignment of Values in Small Instruct Language Models/Post-Alignment Results Audiotext hören
baumhaus.digital/Miscellanous/Symposia/AE55/Symposium on Moral and Legal AI Alignment/From 'Benevolence' to 'Nature': Moral Ordinals, Axiometry and Alignment of Values in Small Instruct Language Models/Post-Alignment Results listen to audiotext
baumhaus.digital/Miscellanous/Symposia/AE55/Symposium on Moral and Legal AI Alignment/From 'Benevolence' to 'Nature': Moral Ordinals, Axiometry and Alignment of Values in Small Instruct Language Models/Post-Alignment Results Memory spielen
baumhaus.digital/Miscellanous/Symposia/AE55/Symposium on Moral and Legal AI Alignment/From 'Benevolence' to 'Nature': Moral Ordinals, Axiometry and Alignment of Values in Small Instruct Language Models/Post-Alignment Results play memory game with primer knot 2096
⬆️
Thema
baumhaus.digital
Miscellanous
Symposia
AE55
Symposium on Moral and Legal AI Alignment
From 'Benevolence' to 'Nature': Moral Ordinals, Axiometry and Alignment of Values in Small Instruct Language Models
⁙
Main result: All models displayed their ability to properly understand the instruction to return a sorted list of randomly shuffled concepts provided in their input.
This article first presents a high-level, language-based method for axiometric exploration of moral value representations infused in diverse small language models. The method is based around the idea of "moral ordinals" - a list of items from a value lexicon which the model is prompted to sort according to its own intrinsic "morality" criterion. After presenting the method, the lexicon based on Schwartz's ``basic value theory'' is used to explore dominance of different value representations in 6 small (<4 milliard parameter) language models. For most models, ``benevolence'' is consistently ranked at the highest position and there is no statistically significant difference between rankings obtained at minimal and default inference temperatures. Across all models, the distribution of aggregate moral-ranking scores was well approximated by a Beta distribution (K–S $p > 0.3$), revealing consistent yet model-specific patterns of moral weighting. Subsequently, foundational models are subjected to a sort of ``minimalist alignment'' whereby they undergo 7 epochs of performance-efficient fine-tuning with synthetically generated 80-instruction codex directed towards sustainability and nature protection. Finally, such minimally aligned models are explored once again with the ``moral ordinals'' method, providing insights into axiological drift induced by the mini-alignment process.
explore & evaluate with Moral Ranking Method (MoRM)
align with Low Rank Adaptation
MoRM-explore&evaluate the aligned model
Discussion
Your browser does not support the audio format.