Favorites - Depends on the model/provider. If you’re running this in Azure you can use...

danielbln , 6 months ago

Depends on the model/provider. If you’re running this in Azure you can use their content filtering which includes jailbreak and prompt exfiltration protection. Otherwise you can strap some heuristics in front or utilize a smaller specialized model that looks at the incoming prompts.

With stronger models like GPT4 that will adhere to every instruction of the system prompt you can harden it pretty well with instructions alone, GPT3.5 not so much.

Reply

Report

Activity

Open original URL

Copy original URL

Copy Mbin URL

Loading...

ns1 6 months ago
FractalsInfinite 6 months ago
adhocfungus 6 months ago
avapa 6 months ago
wanderingmagus 6 months ago
MBM 6 months ago
CorvidCawder 6 months ago