AI Models & Agents

Best Free AI Open Source Models tools

Explore 21 AI Open Source Models tools. 95% include a free plan, free trial, or free credits. Top picks include llmproxy, Wolbarg, LibArgus. Last updated on 8/2/2026.

21 tools95% free access

21

21 tools

20

Free

1

Freemium

1

Free Trial

Top free AI tools in this category with free credits or a free plan.

llmproxy logo

A lightweight, high-performance LLM proxy for caching, failover, cost tracking, and seamless integration between local and cloud AI providers.

Wolbarg logo

Wolbarg is a local-first TypeScript SDK providing shared semantic memory for AI agents using SQLite or PostgreSQL.

LibArgus logo

Unified, zero-allocation native AI inference runtime for Java, consolidating LLM, vision, and speech pipelines via Project Panama.

Osaurus logo

Osaurus is a private, local AI assistant for Mac that runs open models on-device, with optional cloud AI access and autonomous agents.

Reame logo

A lean, fully-tested LLM inference server built for cheap CPU hardware, with persistent caching and an OpenAI-compatible API.

colibri logo

A dependency-free C engine that streams expert weights from disk to run the 744B-parameter GLM-5.2 MoE model on consumer hardware with as little as 25 GB of RAM.

Frugon logo

Frugon is a free, open-source LLM cost analyzer that identifies where your LLM bill leaks by analyzing your call logs locally.

Branchless NCCL Router logo

A branchless, zero-jitter ingress router for 32-GPU distributed mesh networks utilizing JAX/XLA and NCCL.

Skeights logo

Skeights serializes fitted scikit-learn models to safetensors and JSON, replacing insecure pickle with a safe, inspectable format.

mlx-serve logo

Run any LLM locally on your Mac faster than LM Studio or Ollama, with chat, coding agents, images, video, and voice, all fully offline and open source.

Top Free AI Open Source Models tools

All tools
llmproxy logo

A lightweight, high-performance LLM proxy for caching, failover, cost tracking, and seamless integration between local and cloud AI providers.

Wolbarg logo

Wolbarg is a local-first TypeScript SDK providing shared semantic memory for AI agents using SQLite or PostgreSQL.

LibArgus logo

Unified, zero-allocation native AI inference runtime for Java, consolidating LLM, vision, and speech pipelines via Project Panama.

Osaurus logo

Osaurus is a private, local AI assistant for Mac that runs open models on-device, with optional cloud AI access and autonomous agents.

Reame logo

A lean, fully-tested LLM inference server built for cheap CPU hardware, with persistent caching and an OpenAI-compatible API.

colibri logo

A dependency-free C engine that streams expert weights from disk to run the 744B-parameter GLM-5.2 MoE model on consumer hardware with as little as 25 GB of RAM.

Frugon logo

Frugon is a free, open-source LLM cost analyzer that identifies where your LLM bill leaks by analyzing your call logs locally.

Branchless NCCL Router logo

A branchless, zero-jitter ingress router for 32-GPU distributed mesh networks utilizing JAX/XLA and NCCL.

Skeights logo

Skeights serializes fitted scikit-learn models to safetensors and JSON, replacing insecure pickle with a safe, inspectable format.

mlx-serve logo

Run any LLM locally on your Mac faster than LM Studio or Ollama, with chat, coding agents, images, video, and voice, all fully offline and open source.

AXIOM logo

AXIOM is a bootable Rust kernel that optimizes transformer inference by replacing generic OS abstractions with inference-specific primitives.

NanoEuler logo

NanoEuler is an open-source GPT-2-style language model built entirely from scratch in C/CUDA with hand-written backprop, BPE tokenizer, FlashAttention, and training pipelines.

autofit2 logo

Automated end-to-end few-shot text classification pipeline supporting 50+ languages, built on SetFit and SBERT embeddings.

Oxlo.ai logo

Oxlo.ai is a privacy-first AI inference API offering request-based pricing for over 45 open-source models.

Free
PII GUI logo

PII GUI is a local-first desktop application for detecting and redacting personal information in documents using regex and on-device ONNX models.

galdor logo

A Go-native framework for building, orchestrating, and observing AI agents with native OpenTelemetry observability and a self-hosted dashboard.

Ollama logo

Ollama is a platform for running large language models locally and scaling to the cloud, offering access to faster, larger models with parallel requests and real-time web information.

Jan logo

Jan is an open-source, local-first AI chat app that works as a private ChatGPT replacement.

Fireworks AI logo

Fireworks AI is a generative AI platform for fast inference, model hosting, fine-tuning, and scalable deployment of open models.

PixaryAI logo

PixaryAI is an online AI video generator for creating videos from text prompts or images with browser-based controls.

Free

Recently Added AI Open Source Models tools

The newest AI Open Source Models tools added to our directory.

llmproxy logo

A lightweight, high-performance LLM proxy for caching, failover, cost tracking, and seamless integration between local and cloud AI providers.

Wolbarg logo

Wolbarg is a local-first TypeScript SDK providing shared semantic memory for AI agents using SQLite or PostgreSQL.

LibArgus logo

Unified, zero-allocation native AI inference runtime for Java, consolidating LLM, vision, and speech pipelines via Project Panama.

Popular Use Cases for AI Open Source Models tools

The most common use cases across 21 AI Open Source Models tools.

Integrate local LLM tools with NVIDIA cloud-hosted models without client modificationsMonitor and track API costs and usage via stats dashboardReduce latency and API calls with response caching for non-streaming requestsMulti-agent collaboration with shared memoryPersistent knowledge storage for AI assistantsSemantic search over agent learningsBuilding low-latency AI chatbots and virtual assistants in JavaDeveloping multimodal applications that process text, images, audio, and video

Common Features in AI Open Source Models tools

The most frequently featured capabilities across AI Open Source Models tools.

Emulates Ollama, OpenAI, and llama.cpp APIsTransparent forwarding to NVIDIA's OpenAI-compatible APIOptional response caching with configurable TTL and sizeAutomatic failover and retry on transient upstream errorsLive /stats dashboard for metrics and process monitoringInbound authentication supportMulti-model discoveryStreaming support for chat and completionsLocal-first storage with no hosted vector SaaS requiredSemantic recall via natural language queriesMulti-agent safe concurrent writesSupport for SQLite and PostgreSQL backends

FAQ about free AI Open Source Models tools

Compare whether each tool is fully free, freemium, or trial-based. Then check the tool page for core features, use cases, pricing, and alternatives.

Related AI Categories

Discover more free AI software in similar categories.

Categories