RADAR ARCHIVE
Week 45, 2024
1 tool added to the directory, Nov 4 – 10, 2024.
Added to the directory (1)
Tested and written up.
Tencent's 389B-parameter open-source MoE language model
Open-source Transformer-based Mixture-of-Experts language model with 389 billion total parameters and 52 billion active parameters. Supports instruction-tuned and long-context pretraining checkpoints up to 256K tokens, distributed via Hugging Face and GitHub.
Why: Largest open-source Transformer-based MoE model from Tencent, ideal for researchers and builders who want to self-host a capable long-context LLM.
Free
Best for Open-Source LLM Workloads
Visit