Baidu: ernie-4.5-21b-a3b-pt:free
A sophisticated text-based Mixture-of-Experts (MoE) model featuring 21B total parameters with 3B activated per token, delivering exceptional multimodal understanding and generation through heterogeneous MoE structures and modality-isolated routing. Supporting an extensive 131K token context length, the model achieves efficient inference via multi-expert parallel collaboration and quantization, while advanced post-training techniques including SFT, DPO, and UPO ensure optimized performance across diverse applications with specialized routing and balancing losses for superior task handling.
Todos los proveedores de este modelo están ocupados ahora mismo
Cada proveedor ascendente alcanzó su límite de velocidad. El modelo vuelve automáticamente cuando los límites se levantan, normalmente en cuestión de horas. Inténtalo de nuevo en un rato o cambia a otro modelo.
Solicitar este modelo en Discord