Technology
llama
Meta's open-weights LLM family optimized for high-performance local deployment and custom fine-tuning across 8B to 405B parameter scales.
Llama 3.1 delivers state-of-the-art performance through a flagship 405B parameter model trained on 15 trillion tokens. It supports a 128k context window: ideal for analyzing massive datasets or long-form documentation. Developers utilize Llama for diverse tasks (multilingual translation, Python code generation, and complex reasoning) while maintaining data sovereignty via local hosting. The ecosystem includes the Llama Stack for agentic workflows and optimized weights for 8B and 70B models, ensuring high throughput on consumer hardware or enterprise clusters.
What builders pair with llama
Projects using both technologies. Select a pairing to see a project.
12 more pairings
Pairing: Python
Create Your Own DeepSeek Moment
Pairing: Mistral
Benchmarking 100 LLM Inference Engine Configurations
Pairing: OpenAI API
Personal assistant using low code AI tools
Pairing: vLLM
Local hosting - sometimes joy can come in small packages
Pairing: Qwen
LLMs for retrieval and recommendation
Pairing: Docker
Taskless: I Stopped Writing Tasks and Let My Notes Do It
Recent Talks & Demos
Showing 1-24 of 48