Ollama
Ollama is a command-line tool for running large language models on a local computer. It allows users to download and run models like Llama 2, Code Llama, and others locally, and supports custom and custom model creation. This free and open-source project currently supports macOS and Linux operating systems, with future support for Windows.
In addition, Ollam provides an official Docker image, making it easier to deploy large language models using Docker containers and ensuring that all interactions with these models occur locally, without sending private data to third-party services. Ollam supports GPU acceleration on macOS and Linux and provides a simple command-line interface (CLI) as well as a REST API for interacting with applications.
This tool is particularly useful for developers or researchers who need to run and experiment with large language models on their local machines without relying on external cloud services.
Ollama now offers desktop versions for macOS and Windows , featuring new file handling and multimodal interaction capabilities, making it more intuitive and convenient for users to interact with the local large language model.