AMD
1 update on AMD.
AMD trains its first small language model, AMD-Llama-135M, on MI250 accelerators
AMD trained a 135M model from scratch on Instinct MI250 accelerators and reports up to 3.88x faster CodeLlama-7b inference when it drafts tokens.
Cutting edge on edge devices.
1 update on AMD.
AMD trained a 135M model from scratch on Instinct MI250 accelerators and reports up to 3.88x faster CodeLlama-7b inference when it drafts tokens.