Aligned, AMD, and USC ISI Collaborate on MEGALODON Language Model
Aligned has announced a strategic collaboration with AMD and the University of Southern California's Information Sciences Institute (USC ISI) to accelerate the development of the MEGALODON language model, announced in a press release. This partnership focuses on utilizing AMD's Instinct MI300 GPUs and Aligned's AI expertise to train the model efficiently on non-NVIDIA architectures.
MEGALODON is USC ISI's flagship large language model, featuring a novel Moving Average Equipped Gated Attention (MEGA) architecture designed to improve long-context retention while reducing computational complexity. The collaboration aims to push performance boundaries beyond traditional CUDA-based environments by leveraging AMD's ROCm platform as a CUDA alternative.
Aligned will optimize MEGALODON's training on AMD's Instinct MI325X GPUs, providing the necessary tools and engineering support to USC ISI's researchers. This effort is part of a broader initiative to enable powerful AI on diverse hardware platforms, showcasing the potential of AMD's GPU infrastructure in large-scale AI training.
We hope you enjoyed this article
Consider subscribing to one of our newsletters like Silicon Brief or Daily AI Brief.
Also, consider following us on social media:
More from Data Centers
Sep 30 Sophia Space and Redwire Explore Orbital Data Centers Sep 30 SKF Recreates Greta Garbo With AI for Magnetic Bearing Campaign Sep 30 LG Innotek Targets $5.94 Billion in Semiconductor and Physical AI Businesses Sep 29 Efficient Computer Raises $97 Million to Scale Its Processors Sep 29 Hikvision Adds HIKO AI Engine to Hik-Connect 7Silicon Brief
Weekly coverage of AI hardware developments including chips, GPUs, cloud platforms, and data center technology.
Whitepaper
Tensordyne Napier: What If One Rack Could Do the Work of Nine?
Tensordyne
This Tensordyne whitepaper presents Napier, an inference-focused AI processor and rack-scale system based on the company’s TDN Math logarithmic number system. It examines infrastructure requirements for large mixture-of-experts and agentic models, compares major inference architecture approaches, and details the TDN AIP processor, TDN72 pod, TDN Link fabric, and Napier Ultra configuration. The paper reports simulation-based performance, cost, and accuracy-validation results, including Tensordyne’s projected comparison of one Napier rack with a nine-rack Nvidia Rubin plus Groq deployment; the chip is reported as taped out and in fabrication.
Read moreYou may also like
Morgan State and Google Public Sector Build AI Research Campus
SDMC Showcases Deployable AI Home Architecture at IBC 2026
LG Forum Examines Physical AI in US Manufacturing
Ultima Genomics and NVIDIA Collaborate on Pangenome Sequencing Analysis
Ohio State and Google Partner on AI Research Hub
Daily AI Brief: the AI news that matters, in your inbox.