Servidor de IA local com API compatível com a OpenAI para texto, voz, imagem e vídeo

por Ettore Di Giacinto e equipe LocalAI

Em portuguêsUsa inteligência artificial
Nota dos leitores
Sem avaliaçõesAvaliar
No ranking
5ºIA local
Downloads pelo Flipters
0
Estrelas no GitHub
49,4 mil

Última versão: v4.11.0, de 2 de outubro de 2026

Sobre o LocalAI

O LocalAI é um runtime de código aberto, criado por Ettore Di Giacinto e mantido pela equipe LocalAI, para rodar modelos de texto, voz, visão, imagem e vídeo no próprio hardware. Atende quem quer trocar APIs de nuvem por modelos locais sem reescrever o código.

Expõe APIs compatíveis com OpenAI, Anthropic, Ollama e ElevenLabs, tem interface web e galeria de modelos e funciona sem GPU. Instala via Docker ou Podman, binário para Linux ou app launcher para macOS; no Windows, o caminho é o container.

Principais recursos

  • API compatível com OpenAI, Anthropic, Ollama e ElevenLabs
  • Texto, transcrição, síntese de voz, imagem e vídeo no mesmo servidor
  • Roda em CPU; a GPU é opcional
  • Galeria com mais de mil modelos e instalação em um clique
  • Agentes, MCP e RAG embutidos

Arquivos para baixar

Instaladores publicados pelo fabricante, separados por sistema. Na dúvida, use o botão principal: ele leva à página oficial de download.

macOS
ArquivoVersãoTamanhoBaixar
macOS (arm64)local-ai-v4.11.0-darwin-arm64v4.11.0182 MBBaixar macOS (arm64)
macOS .dmgLocalAI.dmgv4.11.019,5 MBBaixar macOS .dmg
Linux
ArquivoVersãoTamanhoBaixar
Linux (arm64)local-ai-v4.11.0-linux-arm64v4.11.0175 MBBaixar Linux (arm64)
Linux (x64)local-ai-v4.11.0-linux-amd64v4.11.0186 MBBaixar Linux (x64)
Linux .tar.xzlocal-ai-launcher-linux.tar.xzv4.11.015,9 MBBaixar Linux .tar.xz

Novidades

Novidades da versão v4.11.0Publicada em 2 de outubro de 2026. Notas do GitHub, em inglês.Mostrar

🎉 LocalAI 4.11.0 Release! 🚀

LocalAI 4.11.0 is out!

This release makes LocalAI more useful for audio understanding, resilient model serving, structured decisions, and day-to-day operation. Audio scenes can now combine transcription, diarization, sound detection, and remembered speaker names. Ordered failover chains keep a public model available across local and remote targets, while new Studio and operations pages expose these capabilities without requiring distributed mode.

The release also adds first-class decision models through /v1/systemone, signed OCI model galleries, and Kimodo text-to-animation. It includes 243 merged pull requests, 353 commits, 79 new gallery entries, focused fixes across APIs and backends, and broad backend-source updates.

Highlights:

  • 🎙️ Audio scenes and remembered speakers - combine speech-to-text, speaker diarization, sound-event detection, and speaker identification. Studio can diarize a recording, preview clean intervals, and register selected speakers by name.
  • 🔀 Model failover chains - place local and remote targets behind one model name, retry before response commitment, observe health and target switches, and pin a target from the API, MCP tools, or the React UI.
  • 🧭 Decision models - advertise the decisions use case and answer structured choice, score, and noul questions through /v1/systemone, with validation and dedicated gallery models.
  • 📦 Signed OCI galleries - distribute complete model galleries as OCI artifacts, verify them with Sigstore policies, and use digest-bound last-known-good caches.
  • 🕺 Kimodo text-to-animation - generate skeletal motion from text through POST /3d/animate, preview it in Studio, and export binary glTF animation.
  • 🖥️ Operate this machine - inspect resources and locally loaded models, view logs, and stop models on a single LocalAI host without enabling distributed mode.
  • 🛡️ Safer gallery and API behavior - path-confinement fixes, stricter credential checks, correct pre-stream errors, preserved streamed Responses items, and more reliable model installation.

Plus PDF attachment extraction in chat, deeper Hugging Face repository discovery, improved hardware detection, new Italian Piper voices, NeMo diarization and ASR models, and large model-gallery batches.

📸 [ screenshot: diarize a recording and remember speakers by name ] Studio turns anonymous speaker segments into reusable named speaker profiles.

---

📊 This release in numbers

| | | |---|---| | Pull requests merged | 243 | | Commits | 353 | | Files changed | 623 (+51,035 / -2,625) | | Development window | 15 days (2026-09-18 to 2026-10-02) | | Human contributors | 12, of whom 4 first-time | | Gallery entries | 1,847 to 1,926 (+79) |

Where the work landed:

| Area | Change | |---|---| | core/ | +23,883 / -807 across 350 files | | backend/ | +12,437 / -186 across 119 files | | gallery/ | +4,998 / -110 | | pkg/ | +3,310 / -170 across 61 files | | docs/ | +2,142 / -1,115 across 50 files | | swagger/ | +2,254 / -100 |

---

📌 TL;DR

| Area | Summary | |------|---------| | 🎙️ Audio scenes | parakeet-cpp can combine ASR, diarization, and sound detection in one model. POST /v1/audio/diarization can optionally include text and versioned speaker profiles. Transcription segments and words can carry speaker labels, and realtime sessions emit transcription-segment and sound-detection events. | | 🗣️ Remembered speakers | Configure a compatible speaker_model to identify voices from t

Ver a versão no GitHub

Avaliações de quem usa

Ainda sem avaliações

Usou o LocalAI? Conte para quem está escolhendo.

Avalie o LocalAI

Sua nota (obrigatória)

De 20 a 3.000 caracteres.0/3.000

Alternativas ao LocalAI

  • Ollama

    Ollama Inc. · Rodar IA no seu computador

    Baixa e roda modelos de linguagem abertos no computador, por app, terminal ou API local

    • 182,4 mil no GitHub
  • llama.cpp

    ggml-org (ggml.ai / Hugging Face) · Rodar IA no seu computador

    Motor em C/C++ para rodar LLMs localmente, com CLI e servidor compatível com a OpenAI

    • 130,5 mil no GitHub
  • LM Studio

    Element Labs, Inc. · Rodar IA no seu computador

    App de desktop para baixar e rodar modelos de linguagem no PC, com chat e API local

Ver todos em Rodar IA no seu computador

Fontes

Ficha escrita pela redação do Flipters a partir das fontes acima, revisada em 6 de outubro de 2026. Preços e versões mudam: confira no site oficial antes de comprar. Como trabalhamos.