RADAR ARCHIVE

Week 45, 2024

1 tool added to the directory, Nov 4 – 10, 2024.

Added to the directory (1)

Tested and written up.

Tencent's 389B-parameter open-source MoE language model
Added Nov 4, 2024
Open-source Transformer-based Mixture-of-Experts language model with 389 billion total parameters and 52 billion active parameters. Supports instruction-tuned and long-context pretraining checkpoints up to 256K tokens, distributed via Hugging Face and GitHub.
Why: Largest open-source Transformer-based MoE model from Tencent, ideal for researchers and builders who want to self-host a capable long-context LLM.
Free Best for Open-Source LLM Workloads Visit