Run any open-source LLMs, such as DeepSeek and Llama, as OpenAI compatible API endpoint in the cloud.
-
Updated
Jul 20, 2026 - Python
Run any open-source LLMs, such as DeepSeek and Llama, as OpenAI compatible API endpoint in the cloud.
Running Llama 2 and other Open-Source LLMs on CPU Inference Locally for Document Q&A
LLM-PowerHouse: Unleash LLMs' potential through curated tutorials, best practices, and ready-to-use code for custom training and inferencing.
LLM (Large Language Model) FineTuning
🏗️ Fine-tune, build, and deploy open-source LLMs easily!
LLMs and Machine Learning done easily
A list of LLMs Tools & Projects
This is a PHP library for Ollama. Ollama is an open-source project that serves as a powerful and user-friendly platform for running LLMs on your local machine. It acts as a bridge between the complexities of LLM technology and the desire for an accessible and customizable AI experience.
Run Open Source/Open Weight LLMs locally with OpenAI compatible APIs
Samples on how to build industry solution leveraging generative AI capabilities on top of SAP BTP and integrated with SAP S/4HANA Cloud.
EmbeddedLLM: API server for Embedded Device Deployment. Currently support CUDA/OpenVINO/IpexLLM/DirectML/CPU
Pair Claude Desktop on Anthropic with Claude Code routed through Ollama. Visual walkthrough + copy-paste prompt that cuts your Claude Code bill ~90%.
GPU-accelerated LLaMA inference wrapper for legacy Vulkan-capable systems a Pythonic way to run AI with knowledge (Ilm) on fire (Vulkan).
md2LLM enables fine-tuning of open-source language models using personal Markdown files. It includes an observability and management layer for the fine-tuning process, allowing users to generate training data, manage models, and track and store each state of the fine-tuned model.
An autonomous agentic pipeline that finds, proves, and patches real C memory-safety vulnerabilities end-to-end using a single 7B open-source LLM (Qwen2.5-Coder) on vLLM, with AddressSanitizer as a ground-truth oracle.
Read your local files and answer your queries
Every Great Contribution Starts With a Fork - Forky 🦎
This project contains the code and documentation for an autonomous AI agent that classifies, enriches, and scores inbound business leads. It is built with a FastAPI backend, a LangGraph agent workflow powered by a local Ollama LLM, and a Streamlit frontend for demonstration.
Multi-agent workflows with Llama3: A private on-device multi-agent framework
SiliconSoap is a cool new way to watch AI agents talk to each other. We make open-source AIs and the newest, most advanced AIs battle it out in exciting, planned conversations. Get ready for AIs with strong personalities, powerful speeches, and quick comebacks. We also add AI-made pictures and sounds to make it even more fun to watch. It's like ...
Add a description, image, and links to the open-source-llm topic page so that developers can more easily learn about it.
To associate your repository with the open-source-llm topic, visit your repo's landing page and select "manage topics."