- Madanapalle, Andhra Pradesh, India
- in/praneetha-n
Pinned Loading
-
ai-safety-policy-translation
ai-safety-policy-translation PublicFramework for translating LLM safety benchmark findings into deployment guidance — evidence-gated tooling with a hard safety mechanism against citing mock/simulated data as real findings. Evidence …
Python 1
-
cross-lingual-llm-safety-eval
cross-lingual-llm-safety-eval PublicA multilingual benchmark for evaluating LLM safety, robustness, and refusal behavior across languages and frontier models.
Python 1
-
-
SentinelAI
SentinelAI PublicCross-lingual LLM safety evaluation framework testing whether refusal behavior holds consistently across English, Hindi, Mandarin, Tamil, and Malay. Pilot v1 complete — includes a full technical re…
Python 1
-
technical-ai-safety-experiment
technical-ai-safety-experiment PublicSystematic evaluation of jailbreak and prompt-injection attacks and defense strategies across open-source and commercial LLMs — 9 guardrail strategies, statistically-grounded comparison (Wilson/boo…
Python 1
If the problem persists, check the GitHub status page or contact support.

