Chapter 4

Features

LocalAI provides a comprehensive set of features for running AI models locally. The pages in this section are grouped by capability, and the left navigation is ordered to match these groups.

Text

Agents

Audio

Vision

Image and Video

  • Image Generation - Create images with Stable Diffusion and other diffusion models.
  • Video Generation - Generate videos from text, image, or audio conditioning, including LongCat and avatar workflows.

Retrieval

  • Embeddings - Generate vector embeddings for semantic search and RAG applications.
  • Reranker - Improve retrieval accuracy with cross-encoder models.
  • Stores - Vector similarity search for embeddings.

Distributed and acceleration

Platform and model management

For operator-facing runtime, proxy, and monitoring concerns (middleware, cloud and MITM proxies, backend monitor), see the Operations section.

Getting Started

To start using these features, make sure you have LocalAI installed and have downloaded some models. Then explore the feature pages above to learn how to use each capability.