Behavioral code analysis with fact-checked AI auto-refactoring
Meta's open classifier for filtering unsafe LLM prompts/replies