Red Teaming
AI Red Teaming
Security testing of AI and machine learning systems — large language models, chatbots, agents and scoring models. We try to make the model leak data it should not, ignore the rules it was given, produce harmful output, or take actions it was never meant to take. Every finding comes with the exact input that caused it and how reliably it reproduces.
