Skip to content
#

sanskrit-nlp

Here are 5 public repositories matching this topic...

SageGPT: A ~7.2M parameter Sanskrit-only decoder-only Transformer SLM trained from scratch on a 139 MB tokenized Sanskrit corpus containing 72.8M SentencePiece model-token IDs, derived from a 105.2M-character purified Sanskrit text corpus using NVIDIA DGX Spark. Architecture: 6 layers, 8 attention heads, 256 embedding dim, 1024 context, 8K vocab.

  • Updated Jul 16, 2026
  • Python

Add this topic to your repo

To associate your repository with the sanskrit-nlp topic, visit your repo's landing page and select "manage topics."

Learn more