Joaquín Herrera
joaco-h
·
AI & ML interests
AI alignment, jailbreak detection, red teaming, model robustness, safety evaluation
Recent Activity
upvoted a paper 10 days ago
StudentSim: Training LLM-based Student Simulators liked a dataset 12 days ago
tourist800/LLM-Hallucination-Detection-complex-mathematics liked a dataset 12 days ago
TrustAIRLab/in-the-wild-jailbreak-promptsOrganizations
None yet