Nipola
...joined 1 month ago, and has 9 karma
submissions / comments / favourites
Independent researcher working on LLM evaluation and AI safety.
I test how open-weight models handle uncertainty, pressure, and adversarial prompts. Currently focused on hallucination patterns, system prompt leakage, and self-disclosure under interrogation.
All data and code: github.com/alitenes2020-sys
Not affiliated with any institution. Everything is exploratory and not peer-reviewed.