Skip to content

Navigation Menu

Sign in
Appearance settings

Search code, repositories, users, issues, pull requests...

Provide feedback

We read every piece of feedback, and take your input very seriously.

Saved searches

Use saved searches to filter your results more quickly

Appearance settings

Latest commit

 

History

History
History
59 lines (42 loc) · 4.65 KB

File metadata and controls

59 lines (42 loc) · 4.65 KB
Copy raw file
Download raw file
Outline
Edit and raw actions
OpenMOSS

Shanghai Innovation Institute (SII) · Fudan University · MOSI.AI

Open, collaborative research on Large Language Models and Multimodal Foundation Models.

Website Hugging Face GitHub Email


👋 About Us

OpenMOSS is a research group led by Prof. Xipeng Qiu, hosted at the Shanghai Innovation Institute (SII) and working in close collaboration with Fudan University and MOSI.AI. We conduct cutting-edge research across the full LLM stack — from model architecture and training to evaluation, interpretability, and real-world applications — with a strong commitment to open and reproducible science.

🔬 Research Directions

Direction Flagship Repositories
🧠 Foundation LLMs MOSS
👁️ Vision & Video MOSS-VL · MOSS-Video-Preview · MOVA
🌐 Omni LLMs AnyGPT
🗣️ Speech / TTS MOSS-TTS · MOSS-TTS-Nano · MOSS-TTSD · MOSS-Audio-Tokenizer
🗣️ Speech / Transcribe MOSS-Transcribe-Diarize
🗣️ Speech / Interaction MOSS-Speech · SpeechGPT-2.0-preview
🎵 Audio & Music MOSS-Audio · MOSS-Music
🤖 Embodied AI & Robotics RoboOmni · FRoM-W1
🔍 Interpretability Llamascopium (formerly Language-Model-SAEs) · Lorsa
📊 Benchmarks & Evaluation VLABench · GAOKAO-MM · Say-I-Dont-Know
Efficiency & Long Context CoLLiE · LongLLaDA · rope_pp
📚 Survey Awesome-WAM · Thus-Spake-Long-Context-LLM

✨ Recent Highlights

  • MOSS-TTS-Nano — 0.1B-param multilingual TTS that runs directly on CPU, no GPU required · ~3.9k★
  • MOSS-TTS — Expressive speech & sound-generation family: long-form, multi-speaker, voice/character design, and streaming TTS · ~3.7k★
  • MOVA — Scalable and synchronized video–audio generation · ~1.1k★
  • Awesome-WAM — Continuously updated reading list & resources for World Action Models in embodied AI · ~1.1k★
  • MOSS-Audio — Unified audio-understanding foundation model spanning speech, sound, music, captioning, QA & reasoning
  • MOSS-VL — Core multimodal vision-understanding series with the full training stack open-sourced
  • MOSS-Transcribe-Diarize — Speech transcription with speaker diarization (who spoke when) for multi-speaker, long-form audio

See the pinned repositories for quick access, or browse all 50+ repositories.

🤝 Join Us

We welcome researchers, students, and collaborators who share our vision. For PhD/intern openings, research collaborations, or general inquiries, please reach us at openmoss@sii.edu.cn.


The Shanghai Innovation Institute (SII) is dedicated to fostering innovation in education and research in the field of artificial intelligence.

Morty Proxy This is a proxified and sanitized view of the page, visit original site.