Scholar
Guanxu Chen
Google Scholar ID: 2TzpwC0AAAAJ
Shanghai Jiao Tong University
Trustworthy AI
Interpretability
Follow
Google Scholar
↗
Citations & Impact
All-time
Citations
3
H-index
1
i10-index
0
Publications
3
Co-authors
0
Publications
15 items
Imprint Reader: From Weight-Update Readout to Behavioral Intervention
2026
Cited
0
SafeEvolve: Harness-Policy Co-Evolution from Agent Experience for Safety Alignment
2026
Cited
0
AgentDoG 1.5: A Lightweight and Scalable Alignment Framework for AI Agent Safety and Security
2026
Cited
0
Preference-aware Influence-function-based Data Selection Method for Efficient Fine-Tuning
2026
Cited
0
Frontier AI Risk Management Framework in Practice: A Risk Analysis Technical Report v1.5
2026
Cited
0
DeepSight: An All-in-One LM Safety Toolkit
2026
Cited
0
AgentDoG: A Diagnostic Guardrail Framework for AI Agent Safety and Security
2026
Cited
0
Loop as a Bridge: Can Looped Transformers Truly Link Representation Space and Natural Language Outputs?
2026
Cited
1
Load more