Archives for PaLM - Page 2
PaLM is not only trained with the much-publicised Pathway system from Google (introduced last year), but it also avoids using pipeline parallelism, a strategy used traditionally for large language models.
PaLM showed a training efficiency of 57.8% hardware FLOPs utilisation - the highest that a large language model at this scale has reached.

