Llama is a family of open-weight models from Meta, not a chatbot app. Map the family, the license, and when a $20 closed chat is easier.
- 1 What is Llama? Llama is a family of weight files plus runners and hosts, not one chat app. Name the door before you paste work data, and do not confuse Groq with Grok.
- 2 License and can I use this at work? Llama at work needs the license PDF, Acceptable Use, and notice habits, not a vibes claim that open-weight means Apache. Write what you ship before the customer highlights Schedule C.
- 3 Llama sizes: laptop vs server Match Llama sizes to real RAM and VRAM. Scout-class fits many laptops; huge MoE names need servers. Write the hardware ticket from a load test, not from a parameter headline.
- 4 Hosted Llama chat vs self-host Choose hosted Llama or self-host on purpose. Local files stay local only if you do not flip a cloud toggle or paste into a logged-in host tab.
- 5 First useful Llama tasks Start Llama with bounded tasks and a verify step. Quote-check summaries, run the transform yourself, and never paste unchecked SLA numbers into a steerco slide.
- 6 Fine-tunes and community variants without the zoo Avoid the Llama variant zoo. Keep one default Instruct stack, match the model card to the file, and remember license and Acceptable Use still follow the weights.
- 7 When Claude or ChatGPT is simply easier Choose Llama when local or open weights earn the ops cost. Choose a closed $20 seat when the job is writing and the GPU project has no owner. Put the chooser on one page. Scheduled · October 2, 2026