SearchSavior/OpenArc
An inference engine for Intel devices that serves LLMs, VLMs, Whisper, TTS, and embedding models via OpenAI-compatible API endpoints.

Not currently ranked — collecting fresh signals.
star history
OpenArc is an inference engine for Intel hardware (CPU/GPU/NPU) that serves a variety of AI models including LLMs, VLMs, Whisper, Kokoro-TTS, Qwen-TTS, Qwen-ASR, and embedding/reranker models over OpenAI-compatible endpoints. It is powered by OpenVINO and supports features like speculative decoding, multi-GPU pipeline parallelism, CPU offload, and hybrid device deployment. The project targets local, private AI inference.
Frequently asked
- What is SearchSavior/OpenArc?
- An inference engine for Intel devices that serves LLMs, VLMs, Whisper, TTS, and embedding models via OpenAI-compatible API endpoints.
- Is OpenArc open source?
- Yes — SearchSavior/OpenArc is open source, released under the Apache-2.0 license.
- What language is OpenArc written in?
- SearchSavior/OpenArc is primarily written in Python.
- How popular is OpenArc?
- SearchSavior/OpenArc has 477 stars on GitHub.
- Where can I find OpenArc?
- SearchSavior/OpenArc is on GitHub at https://github.com/SearchSavior/OpenArc.