Materials for "Local LLMs for Fast, Cheap, and Reproducible Inference"
- Do you really need a zoom lens?
- Security
- Reproducibility
- Cost
- Go Your Own Way
- Performance?
- Option Overload
- Simple Tasks
- Complex Tasks
- Agentic Tasks
- Hugging Face
- Model Hub
transformers
- vLLM
openaitorchtune- LiteLLM
dspyellmer- AI Sandbox - https://researchcomputing.princeton.edu/support/knowledge-base/ai-sandbox
- CLI
- UI
- Open OnDemand (OOD)
- Python API
- vLLM
- TigerFlow + TigerFlow-ML
- Multi-service
- Multi-node
- Quantization
- Fine-tuning
- Getting Started (OOD, CLI)
- Interactive LLMs (OOD)
- LLMs + Python (CLI) - Protest Analysis
- LLMS + Python (Client) - "Stochastic Parrot"
- Batch Translation (UI) - News Articles
- Batch OCR (CLI) - Scientific Articles
- TBD
- Resources
- Stars
- Issues
- Use-cases