{"rewrite":{"id":"r_551e5ed920d6db50ebf5087e","clusterId":"c_ae58f9a9a9eece670e538356","slug":"amd-releases-instella-moe-ai-model-trained-on-its-own-gpus","model":"deepseek-v4-flash:free","headline":"AMD Releases Instella-MoE AI Model Trained on Its Own GPUs","summary":"AMD released Instella-MoE, a Mixture-of-Experts language model trained end-to-end on its Instinct MI300X and MI325X GPUs. The model has 16 billion total parameters and 2.8 billion active parameters. Six versions are available for free, and the training code and framework have been made public.","whyItMatters":"AMD's release of Instella-MoE, trained entirely on its own hardware and software, shows the company's push to establish a fully open AI development stack that competes with comparable small models.","webCardHtml":"\u003cp\u003eAMD has published Instella-MoE, a Mixture-of-Experts language model built and trained using its own Instinct MI300X and MI325X GPUs. The model activates only a subset of its 16 billion total parameters during inference, with 2.8 billion active.\u003c/p\u003e\u003cp\u003eSix versions are available, including a pretrained model on 7.1 trillion tokens, a mid-trained variant with enhanced math and coding skills, a base model with a 64k token context window, and fine-tuned versions using supervised fine-tuning, direct preference optimization, and reinforcement learning.\u003c/p\u003e\u003cp\u003eThe training framework Primus and the MoE architecture FarSkip-Collective were used, and the codebase is public on GitHub. AMD\u0026#39;s performance charts show the Think version scoring higher with fewer active parameters than Gemma-4-E4B-it.\u003c/p\u003e\u003cul\u003e\u003cli\u003e\u003cstrong\u003eInstella-MoE-16B-A3B-Pretrain\u003c/strong\u003e: pretrained on 7.1 trillion tokens\u003c/li\u003e\u003cli\u003e\u003cstrong\u003eInstella-MoE-16B-A3B-Midtrain\u003c/strong\u003e: enhanced math, coding, and reasoning\u003c/li\u003e\u003cli\u003e\u003cstrong\u003eInstella-MoE-16B-A3B-Base\u003c/strong\u003e: context window extended to 64k tokens\u003c/li\u003e\u003cli\u003e\u003cstrong\u003eInstella-MoE-16B-A3B-SFT\u003c/strong\u003e: supervised fine-tuning applied\u003c/li\u003e\u003cli\u003e\u003cstrong\u003eInstella-MoE-16B-A3B-DPO\u003c/strong\u003e: direct preference optimization applied\u003c/li\u003e\u003cli\u003e\u003cstrong\u003eInstella-MoE-16B-A3B-Think\u003c/strong\u003e: reinforcement learning applied\u003c/li\u003e\u003c/ul\u003e","blueskyPost":"AMD released Instella-MoE, a 16B-parameter MoE model trained end-to-end on its Instinct GPUs. Six versions are free, and the training code is public. The Think variant scores higher with fewer active parameters than Gemma-4-E4B-it.","twitterPost":"AMD released Instella-MoE, an MoE language model trained on its own Instinct GPUs. 16B total params, 2.8B active. Six free versions, open training code. Think variant beats Gemma-4-E4B-it on fewer active params.","threadsPost":null,"newsletterBlurb":"AMD released Instella-MoE, a Mixture-of-Experts language model trained end-to-end on its Instinct MI300X and MI325X GPUs. Six versions are free, and the training code is public. The Think variant scores higher with fewer active parameters than Gemma-4-E4B-it.","attributionJson":"[{\"source\":\"GIGAZINE\",\"url\":\"https://gigazine.net/news/20260728-amd-instella-moe/\",\"title\":\"AMD Releases Proprietary AI Model 'Instella-MoE' Trained on Its Own GPUs, Outperforming Gemma-4-E4B in Small Model Class\"}]","lintFlagsJson":null,"lintHits":0,"costUsd":0,"inputTokens":4805,"outputTokens":823,"status":"published","repairAttempts":0,"nextRepairAt":null,"factsAttemptedAt":1786247384,"createdAt":"2026-08-09T03:31:54.000Z","publishedAt":"2026-08-09T03:36:45.000Z","updatedAt":"2026-08-09T03:31:54.000Z"},"cluster":{"id":"c_ae58f9a9a9eece670e538356","canonicalTitle":"AMDが自社製GPUで学習した独自開発AIモデル「Instella-MoE」を公開、Gemma-4-E4Bより高性能な小型モデル","representativeArticleId":"a_d3eceb75cbc7f493dd93a9ab","sourceCount":1,"writtenSourceCount":1,"writeAttempts":0,"isSolo":true,"entitiesJson":"{\"anime_titles\":[],\"manga_titles\":[],\"work_titles\":[],\"studios\":[],\"people\":[],\"type\":\"news\",\"domain\":\"other\",\"is_roundup\":false}","contentType":"news","status":"published","firstSeenAt":"2026-07-28T03:15:00.000Z","lastSeenAt":"2026-07-28T03:15:00.000Z","updatedAt":"2026-08-09T03:36:45.000Z"},"attribution":[{"source":"GIGAZINE","url":"https://gigazine.net/news/20260728-amd-instella-moe/","title":"AMDが自社製GPUで学習した独自開発AIモデル「Instella-MoE」を公開、Gemma-4-E4Bより高性能な小型モデル"}],"entities":{"anime_titles":[],"manga_titles":[],"work_titles":[],"studios":[],"people":[],"type":"news","domain":"other","is_roundup":false},"keyFacts":null}
