
AMD
3 models in Artifacts Logshuggingface.co/amd
Instella-MoE-16B-A3B-Think24 Jul 2026
16B (2.8B active)ResearchRAIL2K · 30d
A 16B-A3B MoE trained by AMD on Instinct cards. AMD also provides all the different stages, from the base to the SFT checkpoints, as well as MidTrain and DPO.
Instella-3B05 Mar 2025
3BResearch-Only (RAIL-MS)854 · 30d
A fully open 3B model trained on the OLMo recipe and tools. The model was trained on the MI300X, AMD's equivalent to the H100. Their evaluations place it as competitive with some other small instruct models (Gemma 2 and Llama 3.2), which we'll follow in community adoption.
AMD-OLMo-1B31 Oct 2024
1.2BApache-2.0193 · 30d
Using our OLMo code with the accompanying dataset and our Tülu data for SFT, AMD has released a 1B model to showcase the training capabilities of their GPUs.