Databricks switches to AMD GPUs to boost LLM training

November 3, 2023

1 min.

Databricks is switching to AMD GPUs to boost its large language model (LLM) training capabilities.

In a collaboration last year, Databricks joined forces with AMD to employ their 3rd Gen EPYC Instance processors. Subsequently, their acquisition of MosaicML, a company utilizing AMD MI250 GPUs for AI model training, further solidified their commitment to AMD’s capabilities.

AMD GPUs have been gaining traction in the AI community, with startups like Lamini and Moreh adopting AMD MI210 and MI250 systems for custom LLMs. Lamini recently disclosed that it runs its LLMs on AMD’s Instinct GPUs, while Moreh trained a language model with a staggering 221 billion parameters using 1200 AMD MI250 GPUs, receiving a $22 million investment.

This move is a testament to AMD’s growing prowess in the GPU space and the potential of its MI250 and MI300X GPUs for accelerating AI workloads. Databricks has achieved performance gains with AMD GPUs, recording a 1.13x improvement in training performance when using ROCm 5.7 and FlashAttention-2 compared to previous results with ROCm 5.4 and FlashAttention.

Databricks also successfully trained MPT-1B and MPT-3B models from scratch on 64 x MI250 GPUs, demonstrating the stability and scalability of AMD’s hardware and software stack.

The sources for this piece include an article in AnalyticsIndiaMag.

Tags
Development

TND Newsdesk

SUBSCRIBE NOW

Become a member

New, Relevant Tech Stories. Our article selection is done by industry professionals. Our writers summarize them to give you the key takeaways

Subscribe Now

North Korean hacker infiltrates US security vendor, loads malware

CrowdStrike releases an update from initial Post Incident Review: Hashtag Trending Special Edition for Thursday July 25, 2024

Security vendor CrowdStrike issues an update from their initial Post Incident Review

CrowdStrike CEO summoned by Homeland Security committee over software disaster

Canadian schools sue social media giants over alleged harm to children

ChatGPT mobile mania: Why users are flocking to ChatGPT Plus

iOS update brings back photos users thought were permanently deleted

Microsoft reveals critical security flaw affecting Android apps

CrowdStrike faces backlash over $10 “apology” voucher

North Korean hacker infiltrates US security vendor, loads malware

Security company accidentally hires a North Korean state hacker: Cybersecurity Today for Friday, July 26, 2024

Security vendor CrowdStrike issues an update from their initial Post Incident Review

Databricks switches to AMD GPUs to boost LLM training

North Korean hacker infiltrates US security vendor, loads malware

Security company accidentally hires a North Korean state hacker: Cybersecurity Today for Friday, July 26, 2024

CrowdStrike releases an update from initial Post Incident Review: Hashtag Trending Special Edition for Thursday July 25, 2024

Security vendor CrowdStrike issues an update from their initial Post Incident Review

Homeland Security committee demands appearance by CrowdStrike CEO

SUBSCRIBE NOW

Related articles

Is Oracle killing off MySQL?

Research Raises Concerns Over AI Impact on Code Quality

Microsoft to train 100,000 Indian developers in AI

NIST issues cybersecurity guide for AI developers

Become a member