Nipola

...joined 1 month ago, and has 9 karma

submissions / comments / favourites

Independent researcher working on LLM evaluation and AI safety.

I test how open-weight models handle uncertainty, pressure, and adversarial prompts. Currently focused on hallucination patterns, system prompt leakage, and self-disclosure under interrogation.

All data and code: github.com/alitenes2020-sys

Not affiliated with any institution. Everything is exploratory and not peer-reviewed.