Last refreshed 2026-06-29. Next refresh: weekly.
Why use MOSS-Audio 8B Thinking on Hugging Face Inference Endpoints?
Hugging Face Inference Endpoints offers MOSS-Audio 8B Thinking with competitive pricing. Hugging Face is a leading AI community and platform dedicated to democratizing artificial intelligence.
Input / 1M
-
Output / 1M
-
Cache
Not sourced
Batch
Not sourced
Setup recipe
Docs fallbackInstall
Use the provider REST API or SDKAuth
Create a provider API keyCall
model: OpenMOSS-Team/MOSS-Audio-8B-ThinkingModel ID
OpenMOSS-Team/MOSS-Audio-8B-ThinkingRequest example
Curated snippets for this provider are not sourced yet. Use Hugging Face Inference Endpoints documentation with model ID
OpenMOSS-Team/MOSS-Audio-8B-Thinking.Gotchas
- Use provider model ID "OpenMOSS-Team/MOSS-Audio-8B-Thinking", not the LLMReference slug "moss-audio-8b-thinking".
Capabilities
MultimodalReasoningAudio
About MOSS-Audio 8B Thinking
MOSS-Audio 8B Thinking is the reasoning-tuned 8.6B variant of MOSI AI and OpenMOSS Team's open-weight audio understanding model. It uses the MOSS-Audio encoder and Qwen3-8B backbone, with Thinking post-training for complex audio reasoning over speech, environmental sound, music, timestamps, captions, and question answering.
Get Started
Model Specs
Released2026-04-13
Parameters8.6B
ArchitectureAudio / Speech