Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up
facebook 's Collections
MobileMoE
MoEViE
Sapiens2
EUPE
LagerNVS
MetaDepth
perception-encoder-audio-visual
sam-audio
Pixio
MobileLLM-R1.5
Meta CLIP 2
SAM 3D Body
MobileLLM-Pro
SAM3
MobileLLM-R1
cwm
DINOv3
Physics of Language Models: Part 4.2
Meta CLIP 1
V-JEPA 2
Web-SSL
blt
Perception LM
Perception Encoder
FAIR Chemistry
DRAMA
Meta Motivo
MobileLLM
Sparsh
LayerSkip
MelodyFlow
Seamless Communication
MAGNeT
Wav2Vec 2.0
SeamlessM4T
XLSR
XLS-R
Robust Wav2Vec 2.0
VoxPopuli
VoxPopuli v2
HuBERT
Fairseq S^2 TTS
DINOv2
MusicGen Stereo
LLM Compiler
Chameleon
Sapiens
OPT
FAIR's LayerSkip Llama models

MobileMoE

updated about 16 hours ago

a family of on-device MoE language models with sub-billion active parameters (0.3-0.9B active and 1.3-5.3B total) that establish a new Pareto frontier

Upvote
5

  • facebook/MobileMoE-S-Base

    Text Generation • 1B • Updated about 2 hours ago • 2 • 5

  • facebook/MobileMoE-M-Base

    Text Generation • 3B • Updated about 2 hours ago • 2

  • facebook/MobileMoE-L-Base

    Text Generation • 5B • Updated 12 minutes ago • 5

  • MobileMoE: Scaling On-Device Mixture of Experts

    Paper • 2605.27358 • Published May 26 • 17
Upvote
5
  • Collection guide
  • Browse collections
Company
TOS Privacy About Careers
Website
Models Datasets Spaces Pricing Docs