vntl-llama3-8b-v2-gguf TweetSharePinShare0 Shares8B open-weight model from lmg-anon for local AI inference. Run vntl-llama3-8b-v2-gguf locally with Ollama: ollama pull vntl-llama3-8b-v2:8b Minimum 16GB RAM. Context length: 4096 tokens. TweetSharePinShare0 Shares