I'm a software engineer in Basel. I spent six years making backends behave, and lately I've been teaching large language models to do genuinely useful things. I'm so glad you're here. Scroll down and I'll show you around.
Nobody reads privacy policies — 74% of people skip them entirely, and reading every one you agree to would cost ~244 hours a year. My thesis asked a sharper question: can large language models catch the moment your data gets handed to a third party, reliably enough to trust?
Structured output + low temperature (0) was the single most reliable combo. Discipline beat cleverness.
Prompt design gave bigger gains than fine-tuning for most categories. Still, QLoRA lifted Mixtral from 0.53 to 0.70 F1 on cookie-tracking.
Models nailed concrete labels ("cookies & tracking") but stumbled on vague ones like "personal information."
Hiring, collaborating, or just want to talk LLMs and clean backends? My inbox is open and I'd love to hear from you.