Asen Dotsinski

Asen Dotsinski
       Sofia, Bulgaria

Hi there!

I’m Asen, a PhD student in LLM security at INSAIT under the supervision of Dr. Yuxia Wang.

I’m currently interested in safeguarding open-weight LLMs against malicious use. I’ve mostly been exploring adversarial robustness against jailbreaking, but I’m curious to learn more about tamper resistance and unlearning.

If you would like to collaborate, feel free to drop me an email!

news

Aug 03, 2026 Excited to be joining INSAIT to start my PhD on LLM security, advised by Dr. Yuxia Wang! :sparkles:
May 01, 2026 Sockpuppetting, my work on jailbreaking LLMs via prefilling and optimization, was accepted at ICML’s Second Workshop on Agents in the Wild: Safety, Security, and Beyond. Thank you to anyone who stopped by my poster! [paper]
Dec 07, 2025 I had a great time at EurIPS this past week, where I presented my reproducibility study on competitions of mechanisms in LLMs! [paper]

selected publications

  1. sockpuppetting_preview.png
    Sockpuppetting: Jailbreaking LLMs by Combining Prefilling with Optimization
    Asen Dotsinski and Panagiotis Eustratiadis
    Second Workshop on Agents in the Wild: Safety, Security, and Beyond, 2026
  2. comp_mech_preview.png
    On the Generalizability of "Competition of Mechanisms: Tracing How Language Models Handle Facts and Counterfactuals"
    Asen Dotsinski, Udit Thakur, Marko Ivanov, Mohammad Hafeez Khan, and 1 more author
    Transactions on Machine Learning Research, 2025