Skip to content

DeepSeek MoE: Efficient Scaling

  • Released: 2024
  • Architecture: Mixture of Experts
  • Strength: Efficient inference
  • Performance: Competitive with dense models
  • Speed: Fast token generation
  • Use: High-throughput inference
  • Status: Innovative approach