1 of 13

Large Language Models

Lecture 4

Word Representation: GloVe

Krishnendu Ghosh

2 of 13

Count-based vs Prediction-based

3 of 13

GloVe – Global Vectors

Crucial insight: Ratios of co-occurrence probabilities can encode word meaning

4 of 13

GloVe – Global Vectors

Crucial insight: Ratios of co-occurrence probabilities can encode word meaning

5 of 13

Co-occurrence Matrix

6 of 13

Learn Word Vectors

7 of 13

Learn Word Vectors

8 of 13

Learn Word Vectors

9 of 13

Learn Word Vectors

10 of 13

Weighting function

11 of 13

GloVe: Advantages

• Fast training

• Scalable to huge corpora

• Good performance even with small corpus and small vectors

12 of 13

Details About GloVe

13 of 13

Source: https://lcs2-iitd.github.io/ELL881-AIL821-2401/lectures/