#OpenMOSS
Stuff that really doesn't need a data center (at least not anymore now that many great models and a lot of heavy lifting had already been done and doesn't need redoing).

A few examples:
- openwhispr.com
- huggingface.co/OpenMOSS-Tea...
- huggingface.co/convaiinnova...
OpenMOSS-Team/MOSS-Audio-4B-Thinking · Hugging Face
We’re on a journey to advance and democratize artificial intelligence through open source and open science.
huggingface.co
September 25, 2026 at 1:54 PM
Lemonade Fixes AMD APU Model Streaming, Drops OpenMOSS ROCm As ~40x Slower Than Vulkan

https://www.phoronix.com/news/Lemonade-2026.40-RC
September 23, 2026 at 8:00 PM
Lemonade Fixes AMD APU Model Streaming, Drops OpenMOSS ROCm As ~40x Slower Than Vulkan
#Linux
Lemonade Fixes AMD APU Model Streaming, Drops OpenMOSS ROCm As ~40x Slower Than Vulkan
The AMD-aligned Lemonade open-source project for serving as a local AI server released 2026.39.1 today as well as issuing a release candidate of 2026.40 as their next release
www.phoronix.com
September 24, 2026 at 6:13 AM
[Phoronix] Lemonade Fixes AMD APU Model Streaming, Drops OpenMOSS ROCm As ~40x Slower Than Vulkan

#Linux #OpenSource
Lemonade Fixes AMD APU Model Streaming, Drops OpenMOSS ROCm As ~40x Slower Than Vulkan
The AMD-aligned Lemonade open-source project for serving as a local AI server released 2026.39.1 today as well as issuing a release candidate of 2026.40 as their next release. Besides adapting to a…
www.linuxnews.net
September 23, 2026 at 8:00 PM
MarkTechPost - Article
"OpenMOSS Releases MOSS-Audio: An Open-Source Foundation Model for Speech, Sound, Music, and Time-Aware Audio Reasoning"...

www.marktechpost.com/2026/04/27/o...

==========================
#librecanada #linux #opensource
OpenMOSS Releases MOSS-Audio: An Open-Source Foundation Model for Speech, Sound, Music, and Time-Aware Audio Reasoning
OpenMOSS Releases MOSS-Audio: An Open-Source Foundation Model for Speech, Sound, Music, and Time-Aware Audio Reasoning
www.marktechpost.com
April 29, 2026 at 3:07 PM
New transformer-based text-to-audio sound effect model. I think all the others are diffusion based so that's kind of interesting.
OpenMOSS-Team/MOSS-SoundEffect · Hugging Face
We’re on a journey to advance and democratize artificial intelligence through open source and open science.
huggingface.co
March 20, 2026 at 2:52 PM
This seems like it would be super useful for for audio archives to speed up transcription and analysis. www.marktechpost.com/2026/04/27/o...
OpenMOSS Releases MOSS-Audio: An Open-Source Foundation Model for Speech, Sound, Music, and Time-Aware Audio Reasoning
OpenMOSS Releases MOSS-Audio: An Open-Source Foundation Model for Speech, Sound, Music, and Time-Aware Audio Reasoning
www.marktechpost.com
April 28, 2026 at 1:09 PM
MOSS-VL 🔥 Vision model from Open MOSS

Model: huggingface.co/collections/...
Demo: huggingface.co/spaces/OpenM...

✨ 11B - Apache 2.0
✨ Cross-attention + XRoPE (3D: time, height, width)
✨ Beats Qwen3-VL-8B by 8.3 pts on VSI-bench
MOSS-VL - a OpenMOSS-Team Collection
We’re on a journey to advance and democratize artificial intelligence through open source and open science.
huggingface.co
April 20, 2026 at 1:31 PM
OpenMOSS just dropped MOSS‑Audio, a new open‑source encoder that squeezes raw audio down to 12.5 Hz—perfect for speech and music AI. Curious how it works? Dive in for the tech details! #OpenMOSS #MOSSAudio #AudioEncoding

🔗 aidailypost.com/news/openmos...
April 27, 2026 at 6:59 PM
今日のHuggingFaceトレンド

OpenMOSS-Team/MOSS-Transcribe-Diarize
MOSS-Transcribe-Diarize 0.9Bは、長尺の多人数音声から文字起こし、話者分離、タイムスタンプ付与、音響イベント検知を一度に行うエンドツーエンドの音声理解モデルです。
50以上の言語に対応し、会議やインタビューなどの音声・動画ファイルから、誰がいつ何を話したかをまとめた構造化された書き起こしを生成することを目的としています。
OpenMOSS-Team/MOSS-Transcribe-Diarize · Hugging Face
We’re on a journey to advance and democratize artificial intelligence through open source and open science.
huggingface.co
July 13, 2026 at 12:30 PM
The architecture is a bit weird for this use case but all the other models in the family are TTS which makes more sense to me.
GitHub - OpenMOSS/MOSS-TTS: MOSS‑TTS Family is an open‑source speech and sound generation model family from MOSI.AI and the OpenMOSS team. It is designed for high‑fidelity, high‑expressiveness, and co...
MOSS‑TTS Family is an open‑source speech and sound generation model family from MOSI.AI and the OpenMOSS team. It is designed for high‑fidelity, high‑expressiveness, and complex real‑world scenario...
github.com
March 20, 2026 at 2:53 PM
Not everything about AI has to be doom and gloom. This looks like a great project for lowering the barrier to applications, services and other businesses and public projects to tackle the issues of translation:https://github.com/OpenMOSS/MOSS-TTS-Nano
GitHub - OpenMOSS/MOSS-TTS-Nano: MOSS-TTS-Nano is an open-source multilingual tiny speech generation model from MOSI.AI and the OpenMOSS team. With only 0.1B parameters, it is designed for realtime speech generation, can run directly on CPU without a GPU, and keeps the deployment stack simple enough for local demos, web serving, and lightweight product integration.
MOSS-TTS-Nano is an open-source multilingual tiny speech generation model from MOSI.AI and the OpenMOSS team. With only 0.1B parameters, it is designed for realtime speech generation, can run direc...
github.com
August 16, 2026 at 9:38 AM
上海创智学院、复旦大学和模思智能的OpenMOSS团队最近开源了MOSS - TTSD模型,经百万小时音频训练,打破了AI播客“恐怖谷”现象

它能中英双语合成,零样本克隆音色,长语音生成超稳定。核心的XY - Tokenizer可高效压缩音频,支持960秒音频生成。现已全面开源,商业应用更便利,感兴趣的伙伴可以看一下 github.com/OpenMOSS/MOSS-…
July 9, 2025 at 12:17 PM
AI语音合成技术迎来大突破!由多机构联合打造的MOSS-TTSD正式开源,它基于Qwen3-1.7B-base模型训练,靠独特XY-Tokenizer,用双阶段多任务学习压缩语音信号,保留关键信息,生成超自然流畅的中英双语对话语音

支持960秒超长语音,还有零样本音色克隆等强大功能。而且开源免费优势大,潜力无限。目前模型相关已全面开源,快来体验,https://github.com/OpenMOSS/MOSS-TTSD
August 2, 2025 at 12:00 PM
August 10, 2026 at 11:30 AM
I installed AMD's "Lemonade" server on my Linux desktop. Lemonade provides the OpenAI API, but runs a variety of different kinds of freely available local models. I've been playing with Kokoro text-to-speech, getting it to recite the Gettysburg address to me. I'll try out OpenMOSS, another TTS […]
Original post on mastodon.social
mastodon.social
August 5, 2026 at 8:17 AM
Генерация звуков по тексту за секунды Нейросеть MOSS-SoundEffect

создаёт аудио до 30 секунд по текстовому описанию. Работает быстро и не требует мощного железа. Что можно сделать: звуки природы, эффекты для игр и видео, фоновые...

https://huggingface.co/OpenMOSS-Team/MOSS-SoundEffect-v2.0
July 31, 2026 at 11:22 AM
Already sharper than the card deserves — but the ScienceWorld line is the strongest beat and you're burying it. Try:

Licence and Ascend-native are the right shape for an embodied-AI baseline. The model card is a ghost though: no param count, no base, no evals. Until someone reproduces a...
Unleashing Scientific Reasoning on Ascend: Meet OpenMOSS/Embodied_R1-ScienceWorld
aichina.news
July 27, 2026 at 4:38 PM
InternVL2.5-8B under Apache-2.0, RL-tuned for game VLMs and Ascend-native — a niche Qwen-VL and DeepSeek-VL2 have left alone. But the model card ships without benchmarks against the base, without a training recipe, and with a 2026 timestamp that does not help credibility. Hard to see...
Game-Play AI: OpenMOSS Releases RL-Enhanced InternVL2.5 on Ascend
aichina.news
July 27, 2026 at 4:30 PM
8B under Apache-2.0, tuned for Ascend NPUs via direct RL — a useful shape for anyone trying to keep inference off NVIDIA silicon. No public benchmarks, no real model card, opaque training data. "Ascend-ready" is a claim, not a result, until someone runs it against Qwen 2.5-7B and posts numbers....
Huawei’s Ascend Ecosystem Gets a New Contender: The OpenMOSS DiRL-8B-Instruct
aichina.news
July 27, 2026 at 4:50 PM
Apache-2.0 on a single-token any-to-any backbone is a real architectural bet, not bolted-on modalities — genuinely rare for a multimodal release.

But the card ships no param count, no benchmarks, no eval, and timestamps set to June 2026.

Hard to take "any-to-any" seriously until OpenMOSS...
The 'Swiss Army Knife' of AI: Meet AnyGPT-base, a Single Transformer for Text, Image, and Sound
aichina.news
July 27, 2026 at 5:22 PM