Rishub Tamirisa
发表论文 6 篇 · 总被引 678 次 · h-index 4
代表论文
- The WMDP Benchmark: Measuring and Reducing Malicious Use With Unlearning (2024 · International Conference on Machine Learning · 被引 510)
- Tamper-Resistant Safeguards for Open-Weight LLMs (2024 · International Conference on Learning Representations · 被引 142)
- Utility Engineering: Analyzing and Controlling Emergent Value Systems in AIs (2025 · Neural Information Processing Systems · 被引 68)
- FedSelect: Personalized Federated Learning with Customized Selection of Parameters for Fine-Tuning (2024 · Computer Vision and Pattern Recognition · 被引 67)
- FedSelect: Customized Selection of Parameters for Fine-Tuning during Personalized Federated Learning (2023 · arXiv.org · 被引 3)
- T OWARD R OBUST U NLEARNING FOR LLM S