Skip to main content
AdOpenFree logoPromote your productReach more potential users and drive product growth and revenue.Advertise

Models & Inference

Open models, embeddings, runtimes, and inference servers.

Favicon of ollama

ollama

Free ListingStars: 181.7K

Start building with open models.

Local Model Runners

Ollama lets you get up and running with open models such as Kimi, GLM, MiniMax, DeepSeek, gpt-oss, Qwen, and Gemma. It offers a CLI, REST API, and libraries for Python and JavaScript to run and manage models locally.

Favicon of llama.cpp

llama.cpp

Free ListingStars: 129.5K

LLM inference in C/C++

Language Models

LLM inference in C/C++ with minimal setup and state-of-the-art performance on a wide range of hardware, locally and in the cloud.

Favicon of DeepSeek-V3

DeepSeek-V3

Free ListingStars: 104.5K

Mixture-of-Experts language model with 671B parameters

Language Models

DeepSeek-V3 is a Mixture-of-Experts language model with 671B total parameters and 37B activated per token. It supports a 128K context and is available in Base and Chat variants with open weights for local deployment and API use.

Favicon of Qwen3

Qwen3

Free ListingStars: 27.7K

Large language model series by the Qwen team, Alibaba Cloud

Language Models

Qwen3 is the large language model series from Alibaba Cloud's Qwen team. It offers dense and Mixture-of-Experts models with thinking and non-thinking modes, multilingual, long-context, reasoning, coding, and agent capabilities.

Favicon of Qwen3-VL

Qwen3-VL

Free ListingStars: 20K

Multimodal large language model series by Qwen team

Language Models

Qwen3-VL is the multimodal large language model series developed by Qwen team, Alibaba Cloud, offering dense and MoE architectures, advanced visual reasoning, long context, and expanded OCR.

Ad
OpenFree logoPromote your product

Reach more potential users and drive product growth and revenue.

Advertise