workflowstacks

The marketplace for AI skills that launch offers, rank in AI search, and automate operations. No coding required.

๐•โšก๐Ÿ’ฌ

Marketplace

  • Browse Skills
  • AI Agents
  • Claude Skills
  • MCP Servers
  • Prompts

Solutions

  • For Founders
  • For Agencies
  • For Ecommerce
  • Agent Builder
  • Starter Packs
  • Playbooks

Learn

  • How It Works
  • What Are Skills
  • What Are Agents
  • What Is MCP
  • For Creators
  • Submit a Tool
  • Security

Company

  • Become a Creator
  • About
  • Enterprise
  • API Docs
  • Terms
  • Privacy
  • Support
Compatible with
๐Ÿค–ChatGPT
โœจClaude
๐Ÿ’ŽGemini
๐Ÿ›๏ธShopify
๐Ÿ”Ahrefs
๐Ÿ“ŠSheets
๐Ÿ’ฌWhatsApp
๐Ÿ“ฑMeta Ads
+50 moreCreator program โ†’

ยฉ 2026 WorkflowStacks. All rights reserved.

TermsPrivacySupport
voice-ai

Xinference: Swap LLMs Easily

Run various LLMs with one API. For founders, by Xinference with 9.3k+ GitHub stars
intermediateโฑ 30 minutes๐Ÿ’ต Free (self-hosted)
9,418 stars845 forksPythonQuality 8/10Updated 7/10/2026100% free ยท open source
What it is

Run different AI models by changing a single line of code.

What you can make with it

Automations that integrate custom AI models with your apps or services, like using a different chatbot model with a single line of code.

How it helps

Swap out AI models without having to rewrite your entire app, and run them on any platform from cloud to your laptop.

Real use case example

"A founder wants to use the latest version of a chatbot model with their customer support app. They can switch out the model by updating a line of code, then test and deploy the changes without affecting their users. With Xinference, they can run the new model on their cloud server instantly."

If you're new

Picking this up first can give you a good foundation to build on when exploring AI models and APIs.

If you're senior

Senior engineers and professionals will want to reach for this to simplify the process of integrating and switching between AI models in their projects.

Common confusion cleared up

Xinference is not a specific AI model itself, but rather an API that allows you to run different models with a single line of code.

Best inside these AI tools
Claude DesktopClaude CodeCursor
Pairs with
Claude APIStripe webhookNotion database
Why we list it on WorkflowStacks: Xinference provides a unified API for running different AI models, which can save development time and increase flexibility for users.
What it does

Run various large language models (LLMs) with a single unified API, allowing you to easily switch between models by changing just one line of code.

Install / run
git clone https://github.com/xorbitsai/inference.git
When to use it
  • โ€ขWhen you need to compare the performance of different LLMs for your application
  • โ€ขWhen you want to deploy a single API that can handle multiple speech, open-source, and multimodal models
  • โ€ขWhen you need to test and integrate LLMs in different environments, such as cloud, on-prem, or local laptops
Quick start
  1. 1Clone the repository and navigate to the project directory: cd inference
  2. 2Review the README for specific configuration and model setup instructions
  3. 3Modify the model configuration as needed, for example by changing the model name in the code
  4. 4Run the inference API with a sample model, following the instructions in the README
  5. 5Use the API to test and integrate different LLMs, swapping them out by changing a single line of code
Ready-to-paste prompt
Modify the example code to run a specific LLM, such as swapping GPT for another model, and test the API with a prompt like 'What is the capital of France?'
Heads up: Make sure you have the necessary dependencies and runtime environment set up, as the inference API requires specific versions of Python and other libraries to function correctly
Saves to your device

Topics

artificial-intelligence
chatglm
deployment
flan-t5
gemma
ggml
glm4
inference
llama
llama3
llamacpp
llm
machine-learning
mistral
openai-api
pytorch
qwen
vllm
whisper
wizardlm
What's inside โ€” free to inspect
No purchase needed

Read the entire source before you build โ€” unlike paid marketplaces that hide it behind a buy button.

15
top-level files
8
folders
75.6M
repo size
Apache-2.0
license
Key files
.pre-commit-config.yaml
AGENTS.md
README.md
File tree
.github/
assets/
benchmark/
doc/
frontend/
monitor/
READMES/
xinference/
.dockerignore
.gitattributes
.gitignore
.pre-commit-config.yaml
.readthedocs.yaml
AGENTS.md
CLAUDE.md
LICENSE
MANIFEST.in
pyproject.toml
README.md
SECURITY.md
setup.cfg
setup.py
versioneer.py
Quick Actions
Details
Creator
xorbitsai
Language
Python
Category
voice-ai
Published
6/14/2023

Are you the creator of this tool? Claim your listing โ†’ and earn 85% of every sale.

Related skills

More voice-ai tools founders pair with this one.

voice-aiโ˜… 68,859
Unsloth: Train AI Models
Train and run open AI models locally with Unsloth Studio, for founders. 68k+ GitHub stars
voice-aiโ˜… 24,989
Boost AI coding with cmux
Get a macOS terminal with vertical tabs and notifications for AI coding agents, ideal for founders using AI coding tools, backed by 21k+ GitHub stars
voice-aiโ˜… 19,758
screenpipe: Record Your Experience
Get AI to learn from your daily interactions with screenpipe, a local and private tool for founders, with 19k+ GitHub stars.
voice-aiโ˜… 17,787
Speech: Build AI Voices
Get scalable speech AI with Speech, a framework for researchers and developers, backed by 18k+ GitHub stars.
voice-aiโ˜… 6,052
FunClip: Auto Video Transcripts
Get automatic video transcription and subtitles with FunClip. For founders needing AI video editing tools.
voice-aiโ˜… 4,350
AI-Youtube-Shorts-Generator: Convert Videos
Get viral YouTube shorts from long-form videos. For founders, with 4.3k+ GitHub stars.