Baidu: ernie-4.5-21b-a3b-pt:free
A sophisticated text-based Mixture-of-Experts (MoE) model featuring 21B total parameters with 3B activated per token, delivering exceptional multimodal understanding and generation through heterogeneous MoE structures and modality-isolated routing. Supporting an extensive 131K token context length, the model achieves efficient inference via multi-expert parallel collaboration and quantization, while advanced post-training techniques including SFT, DPO, and UPO ensure optimized performance across diverse applications with specialized routing and balancing losses for superior task handling.
Tất cả nhà cung cấp của mô hình này hiện đang bận
Mọi nhà cung cấp thượng nguồn đã chạm giới hạn tốc độ. Mô hình tự động trở lại khi giới hạn được gỡ, thường trong vài giờ. Hãy thử lại sau ít phút hoặc chuyển sang mô hình khác.
Yêu cầu mô hình này trên Discord