Which LLM did you use for synthetic data generation?

#1
by kalyan-ks - opened

It's nice to see a model having real world application like handling SMS scams. Out of interest, I would like to know the following details

[1] Normally, LLMs because of safety alignment refuse to generate content (SMS) which is suspicious or dangerous. Did you use an uncensored LLM for synthetic data generation? If yes, which uncensored LLM?
[2] If no, how did you craft the prompt to make the safety aligned LLM to generate suspicious or dangerous SMS content?

Thanks in advance.

Sign up or log in to comment