There’s much talk about Google’s PaLM 2 details being leaked. But seems the only thing “leaked” is unsurprising info: the model has 360B params trained on 3.6T tokens, rather than important details like dataset and architectural tricks. What’s the big fuss?
https://
cnbc.com/2023/05/16/goo
gles-palm-2-uses-nearly-five-times-more-text-data-than-predecessor.html
…
PaLM 2 Leak: Parameters and Tokens Revealed, But Key Details Missing
By
–