Apps & Tools
Velocity Local Services (VLS)
A local LLM server with 1B token context capacity and multi-machine mesh resources.
byVelocity
PlayableOpen source
Description
VLS is a local AI workspace for running LLMs via llama.cpp with a proprietary memory technology called Velocity Context that supports up to 1B tokens. It features an OpenAI-compatible API, a workflow engine for repeatable tasks, and VLS Mesh which allows distributing model layers across multiple local machines to pool hardware resources.
Descriptions, tags, and model credits may be AI-generated or inferred from public sources and can be incomplete or wrong. Learn more.