- Primary focus
- Run Qwen3.8-Flash-Next locally on supported consumer GPUs
- Deployment
- Self-hosted / Web
- Source model
- Open source
- Pricing model
- open source
- API available
- Yes
- Windows desktop
- Yes
- Mobile app
- Not specified
- Overview
- Open-source local inference engine for running a 125B mixture-of-experts model on supported NVIDIA or AMD PCs, with OpenAI and Anthropic-compatible APIs.

| Attribute | ||
|---|---|---|
| Overview | ||
| Primary focus | Run Qwen3.8-Flash-Next locally on supported consumer GPUs | Free, open-source local and self-hosted LLM workspace |
| Deployment | Self-hosted / Web | Self-hosted |
| Source model | Open source | Open source |
| Pricing model | open source | open source |
| API available | Yes | Yes |
| Windows desktop | Yes | Not specified |
| Mobile app | Not specified | Not specified |
| Overview | Open-source local inference engine for running a 125B mixture-of-experts model on supported NVIDIA or AMD PCs, with OpenAI and Anthropic-compatible APIs. | AnythingLLM is a free, MIT-licensed desktop and self-hostable app (by Mintplex Labs) for chatting with your own documents locally, with a separate paid team/cloud option. |
| Explore further | ||
| Vendor website | Visit vendor ↗ | Visit vendor ↗ |
| ITHub profile | View full profile → | View full profile → |
- Primary focus
- Free, open-source local and self-hosted LLM workspace
- Deployment
- Self-hosted
- Source model
- Open source
- Pricing model
- open source
- API available
- Yes
- Windows desktop
- Not specified
- Mobile app
- Not specified
- Overview
- AnythingLLM is a free, MIT-licensed desktop and self-hostable app (by Mintplex Labs) for chatting with your own documents locally, with a separate paid team/cloud option.
ⓘ Confirm technical details and prices with each vendor before making a decision.