Views
No views yet
tau_pos = 1.0, tau_neg = 1.05).nvidia/NVIDIA-Nemotron-3-Nano-4B-BF16 (thinking enabled).| Base model | nvidia/NVIDIA-Nemotron-3-Nano-4B-BF16 |
| Algorithm | SAPO |
| KL coefficient | 0.002 |
| Learning rate | 1e-6 (constant w/ warmup, 10 step warmup) |
| Weight decay | 0.01 |
| Global prompt batch size | 64 |
| Precision | bf16 |
| Num generations per prompt | 4 |
1@article{tamber2026privacyaligncontextualprivacyalignment,
2 title={PrivacyAlign: Contextual Privacy Alignment for LLM Agents},
3 author={Manveer Singh Tamber and Abhay Puri and Marc-Etienne Brunet and Perouz Taslakian and Jimmy Lin and Spandana Gella},
4 year={2026},
5 eprint={2606.21710},
6 archivePrefix={arXiv},
7 primaryClass={cs.CL},
8 url={https://arxiv.org/abs/2606.21710},
9}