نظرة عامة
Download and run Llama, Mistral, and other models locally with a simple CLI and API.
The default local inference stack — zero config, huge model catalog, perfect dev companion.
الميزات الرئيسية
- Run Llama, Mistral, Gemma locally with one command
- OpenAI-compatible REST API on localhost:11434
- Model library with quantized variants for consumer GPUs
- Cross-platform: macOS, Linux, Windows (WSL)
حالات الاستخدام
- 01Local LLM prototyping before sending traffic to cloud APIs
- 02Privacy-sensitive workflows (legal, health, internal docs)
- 03Pair with Open WebUI or LiteLLM for production-style routing
بدء سريع
shell — quick start
curl -fsSL https://ollama.com/install.sh | sh
ollama run llama3.2توافق المكدس
Open WebUI
LiteLLM
LangChain
RunPod
الوسوم
مقالات ذات صلة
وكلاء الذكاء الاصطناعي لدعم العملاء: CrewAI مقابل AutoGen
قارن أطر العمل متعددة الوكلاء لأتمتة الدعم من المستوى الأول في SaaS مستقل.
أفضل أدوات LLM مفتوحة المصدر للمطورين المستقلين (2026)
Ollama وvLLM وLiteLLM والمزيد — ماذا تستضيف ذاتياً عند بناء ميزات الذكاء الاصطناعي.
ضبط Llama 3 على GPU اقتصادي
ضبط LoRA على RunPod بأقل من 20 دولار — متى يستحق ذلك مقابل هندسة المطالبات.