Global AI News Aggregator
About
By
–
Has anyone tried loading two models into memory w/ llama cpp? Like an 8bit Mistral 7B and a 4bit Mixtral simultaneously.
→ View original post on X — @mattlynley