This model is intended for research on multi-agent alignment and instruction following. It is part of the MAHALS (Multi-Agent Hierarchical Alignment) research project.
Limitations
Trained on 10% of data (reduced capability vs full Tulu-3)
English only
May exhibit biases present in training data
Not suitable for production without further evaluation