Local Storage Llm

The best 47 Local Storage Llm AI tools - Free & Paid

Free AI tools 💸 All categories 🎨 Deals ％ For you 👀

Explore 47 AI for Local Storage Llm

Free Only

liteLLM

LiteLLM is an open‑source gateway that unifies access to 100+ LLMs through a single OpenAI‑compatible API, enabling provider fallback, cost tracking, tag‑based budgeting, guardrails, observability, and on‑prem or cloud deployment with a lightweight SDK.

LLM

Freemium

Lmstudio.ai

14 11

LM Studio runs open‑source large language models locally on Mac (M‑series), Windows, and Linux, enabling private, offline inference. It offers command‑line and headless deployment, server‑side API, SDKs, a model hub, and LM Link for remote model access.

Infrastructure tools

Free

LLMChat

4 2

LLMChat is an AI chat tool that offers a beta version experience with diverse AI models, personalized memory, custom assistant creation, and privacy-focused locally stored conversations. Explore features like plugin integration, tailored preferences, and prompt examples for various tasks.

Chat

Free

Arena AI

4 0

LLM Arena enables users to compare multiple large language models side-by-side, analyzing features like accuracy and capabilities. It supports up to 10 models, facilitating informed decision-making for researchers and developers in selecting the right LLM for their needs.

LLM

Free

LLMWare.ai

LLMWare AI installs a lightweight client on PCs, providing instant access to 100+ AI models optimized for Intel and Qualcomm hardware. It supports RAG, auto‑tunes weights, runs locally without Wi‑Fi, and offers an admin console for monitoring, scaling, and audit logs.

LLM

Freemium

LM Studio

LM Studio is a local platform for running various large language models like Llama 2 and Mistral. It offers an offline environment, user-friendly interface, and supports multiple operating systems, enhancing privacy and allowing for simultaneous model execution.

LLM

Freemium

Pieces for Developers

Pieces stores and organizes work‑related context—code, docs, chats—within familiar tools, creating OS‑level long‑term memory. It supports real‑time LLM context via local plugins, letting users keep data on‑device or sync to a chosen cloud, aiding continuity for teams.

Code assistant

Freemium

Related topics: 🔍 llm research assistant 🔍 opensource llm 🔍 llm builder 🔍 open-source llm model 🔍 next-generation llm 🔍 llm ops

Pocket LLM

Pocketllm is an AI-powered personal document search engine that allows you to easily search and retrieve information from thousands of pages of PDFs and documents. It offers semantic search capability, fine-tuning search results and summarizing results.

Document assistant

Free trial

MICRO LLM

Micro LLM is a personal AI assistant that enhances productivity by managing tasks, scheduling appointments, and answering questions. It operates on devices like iPads and iPhones, offering offline functionality and an intuitive interface for seamless organization.

LLM

Free

LLM Price Check

LLM Price Check aggregates LLM API models and provider details into sortable tables and a cost calculator, showing context windows, input/output cost metrics, and quality indicators to help developers and teams evaluate cost–performance tradeoffs.

LLM

Freemium - $1

Vllm

1 0 1

VLLM is a high-throughput, memory-efficient inference engine for Large Language Models, enabling faster responses and effective memory management. It supports multi-node configurations for scalability and offers robust documentation for seamless integration into workflows.

Infrastructure tools

Free

Ollama.ai

20 7

Llama is a local AI tool that enables users to create customizable and efficient language models without relying on cloud-based platforms, available for download on MacOS, Windows, and Linux.

Infrastructure tools

Free

LLMStack

3 1

LLMStack is an open‑source platform that lets developers build AI agents and workflows without coding, supports multiple model providers, imports data from web, PDFs, audio, cloud services, and offers a collaborative React UI with granular permissions.

LLM

Freemium

LLM Pulse

LLM Pulse tracks brand visibility and search presence across LLMs (ChatGPT, Perplexity, Google AI), offering prompt tracking and suggestions, citation analysis, visibility scoring and competitor benchmarking, sentiment and response inspection, plus API and reporting exports.

SEO

Free trial

Lmql

1 0

LMQL is a Python‑based language that enables modular, constraint‑driven prompts for large language models. It supports nested queries, type‑enforced outputs, and runtime distribution checks while switching between backends such as llama.cpp, OpenAI, and Hugging Face.

Code assistant

Freemium

RunLLM

RunLLM is an AI platform that automates incident investigations by querying observability tools, correlating telemetry, and delivering root-cause analyses. It generates live runbooks and remediation recommendations to accelerate MTTR and create an auditable history of incidents.

Automation

Freemium

Inceptionlabs - Mercury coder

Inception Labs' diffusion-based large language models (dLLMs) offer faster, more efficient, and cost-effective text generation than traditional autoregressive models. With built-in error correction, multimodal support, and structured output control, they excel in function calling and complex data ge

LLM

Freemium

Awan LLM

Awan LLM offers unlimited token generation with Meta Llama 3.1 8B and 70B models, no censorship or caps, supporting persistent AI assistance, autonomous agents, roleplay, data processing, and code completion, hosted on owned GPUs for continuous use.

LLM

Subscription

Remind ai

Remind AI is an open‑source memory system that records and indexes digital activity. It stores data from email, messaging, and document editors in a structured graph, enabling users to retrieve past actions and files with natural language via local LLMs.

Memory

Freemium

LLMOps.Space

LLMOps Space is a global community for LLM practitioners, offering curated content, discussion forums, event recordings, and resources on production deployment, fine‑tuning, observability, and search optimization, plus networking via Discord and newsletters.

LLM

Freemium

LLMule

llmule is a decentralized network that enables users to run AI models locally, ensuring data privacy. It offers a library of community-shared models, promoting flexibility and collaboration while eliminating reliance on cloud services.

LLM

Free

Langbase

1 0

Langbase offers a serverless platform for building, deploying, and scaling AI agents. It unifies access to 600+ LLMs, provides built‑in memory, vector, and file storage, and supports durable multi‑step workflows with monitoring and custom actions.

AI Assistant

Freemium

LLM Pricing

1 0

LLM Pricing Comparison lets developers and businesses compare token costs, context lengths, and modalities for major large‑language models. An interactive calculator estimates application expenses based on input/output token volumes, helping teams budget AI workloads accurately.

LLM

Freemium

local.ai

local.ai runs language models locally without GPUs. Its Rust backend keeps the binary under 10 MB and performs CPU inference with GGML quantization. A single‑click interface streams responses to a UI, while a model manager tracks, verifies, and resumes downloads.

Developer tools

Freemium

BenchLLM

BenchLLM evaluates language‑model applications via API or CLI, running JSON/YAML test suites with automated, interactive, or custom strategies. It supports OpenAI, LangChain, and any API, detecting regressions, generating reports, and visualizing results for continuous QA.

Developer tools

Freemium

LexWorkplace

LexWorkplace is a cloud-based document and email management solution for law firms, offering features like advanced search, secure sharing, Microsoft Office integration, and robust data security to enhance efficiency and organization in legal practices.

Legal

Free trial

Little-Coder

1 0 1

little-coder is a Pi-based coding agent for running 5–25 GB local LLMs via llama.cpp or Ollama, offering Python/Node CLIs and TypeScript extensions, reproducible benchmarks, build/serve guides, and tools for local code generation, on-device development, and evaluation.

Code assistant

Free

RLAMA

Rlama is a document question-answering tool that supports multiple formats and offers intelligent parsing and local processing. It enables efficient retrieval-augmented generation with features like document chunking and automatic updates, suitable for secure knowledge management.

AI Agents

Subscription

LLMWizard

LLMWizard offers access to multiple AI models like GPT-4o and DALL-E 3, enabling users to automate tasks across coding, legal work, and content creation. The platform supports real-time comparison of AI responses for diverse insights.

LLM

Free trial

Exllama

1 0

exllama is a memory-efficient tool for executing Hugging Face transformers with the LLaMA models using quantized weights, enabling high-performance NLP tasks on modern GPUs while minimizing memory usage and supporting various hardware configurations.

LLM

Free

Dot App

Dot runs the Mistral 7B LLM locally, letting users upload documents and chat offline. All data stays on the device, providing secure, low‑latency Q&A and contextual conversation for researchers, writers, and professionals.

Document management

Paid

Secret Llama

Secret Llama is a private browser-based chatbot that stores data locally, ensuring enhanced privacy. It supports offline use after initial model download and functions on Chrome and Edge with GPU support, encouraging community contributions for ongoing improvements.

Chatbot builder

Free

LLMSelector

LLM Selector filters open‑source large language models by use case—chatbots, content, code, summarization, research—while presenting benchmarks, training data, architecture, and deployment details. The interface updates regularly to aid researchers, developers, and product managers in data‑driven mo

LLM

Freemium

Unsloth Studio

4 0 2

Unsloth Studio is a no-code web UI enabling local training, running, and exporting of open AI models like Qwen3.5 and NVIDIA Nemotron 3, simplifying experimentation for users without extensive technical expertise.

Infrastructure tools

Free

web2llm

Web2llm converts web documents into structured Markdown files, extracting relevant content while omitting extraneous elements. Users can input multiple URLs, and the tool organizes individual files and provides summaries in a dedicated 'docs' folder.

Document assistant

Freemium

Osmantic ODS

ODS is an on-premises AI server for PC, Mac, and Linux offering local LLM inference, model management, agents, chat/voice UI, RAG with local document search, image pipelines, integrations (Ollama, ComfyUI, n8n), and operational tooling for private deployments.

LLM

Free

The Attic AI

Attic AI offers no‑code tools to create and deploy custom LLMs that convert internal documents into searchable knowledge bases, automate grant and proposal drafting for contractors and universities, and analyze congressional appropriations for compliance—all on secure on‑prem or private cloud.

AI Assistant

Freemium

LLM Answer Engine

LLM-answer-engine is an advanced answer engine leveraging Groq, Mixtral, Langchain.JS, Brave Search, Serper API, and OpenAI to provide sources, answers, images, videos, and follow-up questions efficiently. It offers an opensource Perplexity alternative.

LLM

Free

Llama.cpp

3 0 1

Llama.cpp is an open-source tool for efficient inference of large language models. Run open source LLM models locally everywhere.

Infrastructure tools

Free

Oobabooga

The text-generation-webui is a Gradio-based web UI for Large Language Models, supporting various backends and multiple interface modes. It allows quick model switching, extension integration, and dynamic LoRA loading for custom training.

LLM

Free

Llongterm

llongterm is a memory-focused AI tool that enhances chatbots and virtual assistants by enabling contextual memory retention. It supports personalized interactions across various applications, including education and customer support, and is compatible with multiple programming languages.

Memory

Subscription

Mistral.rs

1 0

Mistral.rs is an efficient, versatile tool for high-speed large language model (LLM) inference, offering multi-device support and extensive quantization options for seamless deployment on diverse hardware setups.

LLM

Free

Ollm

Ollm.com is a confidential AI gateway providing a single API to route across hundreds of LLM models and providers. It ensures enterprise security with zero data retention, confidential computing, and centralized key management for private, compliant AI workloads.

LLM

Freemium

ComicLLM

ComicLLM allows users to easily create and customize comics, offering diverse styles and formats for both storyboards and editorial cartoons. It supports multiple languages and provides options for custom art styles, enhancing creative possibilities.

Art Generation

Freemium

Ava PLS

Ava is an open‑source desktop app that runs language models locally using llama.cpp, offering a GUI or headless mode. Built with Zig/C++ and SQLite, it enables rapid prototyping, privacy‑focused experimentation, and straightforward local deployment.

AI Assistant

Freemium

TextGen - oobabooga

1 0

Open-source desktop app for running local LLMs on Windows/macOS/Linux, supporting text and multimodal inputs, file attachments, multiple model backends with hot-switching, chat/instruction modes, prompt-engineering tools, API/tool-calling, extensibility, and conversation branching.

Infrastructure tools

Free

ManagePrompt

ManagePrompt logs every LLM API request and response locally, visualizing calls and tracking input/output tokens. It supports streaming, auto‑creates SQLite databases, and integrates via npm middleware for any LLM provider.

AI Assistant

Paid

Local Storage Llm

The best 47 Local Storage Llm AI tools - Free & Paid

Explore 47 AI for Local Storage Llm

Related topics

Related Topics