Archive

Tag: AI Chips

Artificial Intelligence

AWS Makes Trainium3 and Trainium3 UltraServers Generally Available, Previews Trainium4 Custom AI Chip

Last year, Amazon Web Services introduced Trainum3, its third-generation chip for training large language models (LLMs). At the time, these were state-of-the-art and twice as fast as their predecessors. However, as AI models become larger and more complex, a better way is needed to reduce latency and training times. Today, the company is introducing Amazon […]

Artificial Intelligence

Arm Lumex Brings Cloud-Level AI Power to Smartphones

Arm is betting the future of mobile AI lies in the device and not exclusively in the cloud. Its new Lumex platform, first introduced in May, is built to prove it. Part of the company’s Compute Subsystems (CSS), it combines Arm’s high-performance CPUs with its second-generation Scalable Matrix Extension (SME2), GPUs, system IP, and an […]

Artificial Intelligence

ServiceNow and Nvidia Debut Apriel Nemotron 15B, an Open-Source Reasoning Model Built for Faster, Cheaper Agentic AI

ServiceNow and Nvidia have had a long-standing partnership building generative AI solutions for the enterprise. This week, at ServiceNow’s Knowledge customer conference, the two are introducing the latest fruits of their labor, a new large language model called Apriel Nemotron 15B with reasoning capabilities. The companies believe it performs as well as OpenAI’s o1-mini, Alibaba’s […]

Artificial Intelligence

AMD Unveils OLMo, Its First Fully Open 1B-Parameter LLM Series

AMD has introduced OLMo, a new series of large language models it trained in-house using trillions of tokens on a cluster of its Instinct MI250 GPUs. Though its specific purpose isn’t explicitly stated, it’s believed the company created OLMo to highlight the capabilities of its Instinct GPUs in running “large-scale multi-node LM training jobs with […]