Leo Park
leopark42
AI & ML interests
AI alignment, jailbreak detection, red teaming, model robustness, safety evaluation
Recent Activity
upvoted a paper about 7 hours ago
What Gradients Add to Text Leakage in Split Language Models, Counted per Token and per Document upvoted a paper 1 day ago
World Embedding Benchmark upvoted a paper 3 days ago
Architect-Ant: Editable Automatic Furnishing of Architectural Floor PlansOrganizations
None yet