I didn't believe when I first saw, but: We trained a prompt stealing model that gets >3x SoTA accuracy. The secret is representing LLM outputs *correctly* ๐ฒ Demo/blog: mattf1n.github.io/pils ๐: arxiv.org/abs/2506.17090 ๐ค: huggingface.co/dill-lab/pi... ๐งโ๐ป: github.com/dill-lab/PILS
11 likes 1 replies
?