Curated by

avatar

Stacklist Team

stacklist.com/stacklist-team

More in Local AI & GPUs

See all 24 →

More from Stacklist Team

See all stacks →

Mesh-LLM: Distributed AI/LLM for Everyone

Mesh LLM is a framework that pools GPUs and memory across machines to expose a unified OpenAI-compatible API for running large language models. It supports distributed inference with automatic routing, stage splits for oversized models, and both public and private mesh configurations.

View card
Built for AI agentsACO · 2056 tokens

Summary

Mesh LLM is a framework that pools GPUs and memory across machines to expose a unified OpenAI-compatible API for running large language models. It supports distributed inference with automatic routing, stage splits for oversized models, and both public and private mesh configurations.

Tags

mesh-llm · distributed-inference · gpu-pooling · openai-api · large-language-models · skippy-splits · framework

Key entities

Mesh LLM (technology, 0.99) · OpenAI API (technology, 0.95) · Skippy stage splits (technology, 0.9) · GGUF (technology, 0.85) · Nostr discovery (technology, 0.8) · Mixture-of-Agents (concept, 0.88) · Qwen3-8B (technology, 0.85) · GLM-4.7-Flash (technology, 0.85) · Flash-MoE (technology, 0.8) · Mesh-LLM (organization, 0.9)

Classification

framework · language en · status final

Provenance

claude-haiku-4-5 via @stacklist/be@0.1.0, confidence 0.85, 12 Jul 2026