Skip to content

Latest commit

ย 

History

History
39 lines (34 loc) ยท 4.3 KB

File metadata and controls

39 lines (34 loc) ยท 4.3 KB

LLM-PaperReview

Review papers of NLP, mainly LLM. NLP / LLM ๊ด€๋ จ ๋…ผ๋ฌธ ๋ฆฌ๋ทฐ ๋ ˆํฌ์ž…๋‹ˆ๋‹ค.

  • ์Šค์ผ€์ผ๋ถ€ํ„ฐ ๋ฉ”์†Œ๋“œ๊นŒ์ง€. ChatGPT์˜ ์ƒ์—…์  ์„ฑ๊ณต๊ณผ ์—ฌ๋Ÿฌ ๋ฐœ๊ตฐ์˜ ์˜คํ”ˆ์†Œ์Šค๋“ค์„ ์—…๊ณ  LLM์€ ๋‚˜๋‚ ์ด ๋น ๋ฅธ ์†๋„๋กœ ์„ฑ์žฅํ•˜๊ณ  ์žˆ์Šต๋‹ˆ๋‹ค.
  • ์Ÿ์•„์ง€๋Š” LLM ๊ด€๋ จ ๋…ผ๋ฌธ๋“ค์„ ๊นŠ๊ณ  ๋„“๊ฒŒ ๊ณต๋ถ€ํ•˜๊ณ  ํ† ๋ก ํ•˜๊ณ ์ž ํ•ฉ๋‹ˆ๋‹ค. :)
  • ๋…ผ๋ฌธ ์„ ์ •์€, Pond์— ๋ชจ์•„๋‘” ๋…ผ๋ฌธ๋“ค ์ค‘ ํ•˜๋‚˜๋ฅผ ๋ฐœํ‘œ์ž๊ฐ€ ๋ฐœํ‘œ์ผ 1์ฃผ ์ „๊นŒ์ง€ ์„ ์ •ํ•˜์—ฌ ๊ณต์ง€ํ•˜๋Š” ๊ฒƒ์œผ๋กœ ์ด๋ฃจ์–ด์ง‘๋‹ˆ๋‹ค.

Semesters

  • 2023 8/6 - 2023 10/19 Every Thursday(Finished).
  • 2023 11/23 - 2024 2/1 Every Thursday(Finished).

์ผ์ • ๋ฐ ์„ ์ • ๋…ผ๋ฌธ

2๊ธฐ

ย  Paper a.k.a Affiliation published date Speaker Youtube
11.23 REPLUG: Retrieval-Augmented Black-Box Language Models REPLUG Washington Univ. May. 2023 ๊น€ํ•œ์„ฑ LINK
11.30 Prefix-Tuning: Optimizing Continuous Prompts for Generation Prefix-Tuning Standford Univ. Jan. 2021 ์ž„์„œ์—ฐ LINK
12.7 QLoRA: Efficient Finetuning of Quantized LLMs QLoRA Washington Univ. May. 2023 ์ด์ƒ๋ฏผ LINK
12.21 Direct Preference Optimization: Your Language Model is Secretly a Reward Model DPO Stanford Univ. May. 2023 ์ฒœ์žฌ์› LINK
12.28 Efficient Streaming Language Models with Attention Sinks StreamingLLM MIT Dec. 2023 ๊น€๊ฐ€์˜ LINK
1.4 Efficient Memory Management for Large Language Model Serving with PagedAttention vLLM UC Berkeley Sep. 2023 ์ด์ฃผํ˜• LINK
1.18 Retrieval-Augmented Generation for Knowledge-Intensive NLP Tasks RAG FAIR April. 2021 ์‹ ์ค‘ํ˜„ LINK
1.25 Mixtral of Experts Mixtral Mistral.AI Jan. 2024 ์‹ ํ˜์ค€ LINK
2.1 Understanding In-Context Learning in Transformers and LLMs by Learning to Learn Discrete Functions - Oxford Univ. Oct. 2023 ๊น€ํ˜„์ˆ˜ LINK

1๊ธฐ

ย  Paper a.k.a Affiliation published date Speaker Youtube
8.17 RoFormer: Enhanced Transformer with Rotary Position Embedding RoPE Zhuiyi Technology August. 2022 ์ฒœ์žฌ์› LINK
8.24 TRAIN SHORT, TEST LONG:
ATTENTION WITH LINEAR BIASES
ENABLES INPUT LENGTH EXTRAPOLATION
ALiBiย  Facebook April. 2022 ์ด์ฃผํ˜• LINK
8.31 Finetuned Language Models Are Zero-Shot Learners FLAN Google Sep. 2021 ์ฒœ์†Œ์˜ LINK
9.7 WizardLM: Empowering Large Language Models to Follow Complex Instructions WizardLM Microsoft Jun. 2023 ๋ฐ•๊ฒฝํƒ LINK
9.14 G-Eval: NLG Evaluation using GPT-4 with Better Human Alignment G-Eval Microsoft May. 2023 ์‹ ํ˜์ค€ LINK
9.21 SimCSE: Simple Contrastive Learning of Sentence Embeddings SimCSE Princeton Univ. May. 2022 ๊น€์„ธํ˜• LINK
10.5 LLaMA: Open and Efficient Foundation Language Models LLaMA Meta Feb. 2023 ๊น€๊ฐ€์˜ LINK
10.12 LoRA: Low-Rank Adaptation of Large Language Models LoRA Microsoft Oct. 2021 ์‹ ์ค‘ํ˜„ LINK
10.19 Training language models to follow instructions with human feedback InstructGPT OpenAI March. 2022 ํ™์˜ํ›ˆ LINK