20 entries · Updated 2026-10-01

Companies

Labs, hardware makers and toolmakers betting on AI that runs on your own machine.

hubFR · US

Hugging Face

Hosts the models, datasets and quantizations the whole scene depends on.

huggingface.co
labFR

Mistral AI

Paris lab that releases most of its models under Apache 2.0, from Ministral 3B to Mistral Large 3.

mistral.ai
labFR

Kyutai

Non-profit lab in Paris. Open speech models like Moshi that talk back in real time.

kyutai.org
labDE

Black Forest Labs

The FLUX image models, the default pick for local image generation.

huggingface.co/black-forest-labs
labUS · GB

Google Gemma

Gemma 4 covers phones (E2B) up to a single 24 GB card (31B), all Apache 2.0.

huggingface.co/google
labCN

Qwen (Alibaba)

Every size from phone to server, mostly Apache 2.0, and usually the first to get quantized.

huggingface.co/Qwen
labCN

DeepSeek

MIT-licensed MoE models that pulled open weights close to the frontier.

huggingface.co/deepseek-ai
labCN

Moonshot AI

The Kimi line. Trillion-parameter MoE with open weights.

huggingface.co/moonshotai
labCN

Z.ai

GLM models, strong at coding and agent work.

huggingface.co/zai-org
labUS

Nous Research

Hermes fine-tunes and open research on distributed training.

nousresearch.com
runtimeUS

Ollama

One command to pull and run a model. The easiest way in on any OS.

ollama.com
runtimeUS

LM Studio

Desktop app to find, download and chat with local models, with an OpenAI-compatible server built in.

lmstudio.ai
toolingUS

Comfy Org

The company behind ComfyUI. Day-one workflows for new image and video models, all running locally.

comfy.org
toolingUS

Unsloth

Fine-tuning that fits on one consumer GPU, plus carefully tested GGUF quants.

unsloth.ai
hardwareUS

NVIDIA

CUDA is still the first target of every inference engine.

nvidia.com
hardwareUS

AMD

ROCm keeps improving, and Ryzen AI Max puts 128 GB of unified memory in a small desktop.

amd.com
hardwareUS

Intel

Arc cards offer a lot of VRAM per euro and an open software stack.

intel.com
hardwareUS

Apple MLX

Unified memory on M-series Macs plus MLX make large models practical on a desk.

github.com/ml-explore/mlx
hardwareUS

Framework

Framework Desktop, a small repairable PC built around Ryzen AI Max with up to 128 GB.

frame.work
hardwareUS

tiny corp

tinygrad and the tinybox, multi-GPU machines that ship ready for local training and inference.

tinygrad.org