AMD Challenges Nvidia with Helios Rack-Scale System for AI Computing
Product launch

AMD Challenges Nvidia with Helios Rack-Scale System for AI Computing

The new Helios system supports the computing requirements of leading AI labs, backed by strategic partnerships and a projected $1.4 trillion market

3 min read
Based on original reporting byTechCrunchTranslated, summarized and given business context by our systemHow we work

Executive summary

Key Takeaways

  • AMD has launched the Helios system, a rack-scale server system expected to begin shipping later this year to compete with Nvidia's Vera Rubin and Grace Blackwell systems.

  • The system has already secured five major industry customers operating AI laboratories, including OpenAI, Meta, Oracle, Anthropic, and Microsoft.

  • Anthropic announced a strategic partnership to deploy up to 2 gigawatts of GPUs utilizing the new rack-scale system.

  • AMD also introduced its new Venice-X CPU for data centers, which is officially expected to launch in 2027.

  • AMD CEO Dr. Lisa Su projects that the AI accelerator market will reach a value of $1.4 trillion by 2030, driven primarily by GPUs.

AMD Challenges Nvidia with Helios Rack-Scale System for AI Computing

  • AMD has launched the Helios system, a rack-scale server system expected to begin shipping later...
  • The system has already secured five major industry customers operating AI laboratories, including OpenAI, Meta,...
  • Anthropic announced a strategic partnership to deploy up to 2 gigawatts of GPUs utilizing the...
  • AMD also introduced its new Venice-X CPU for data centers, which is officially expected to...
  • AMD CEO Dr. Lisa Su projects that the AI accelerator market will reach a value...

According to a report by technology publication TechCrunch, global chipmaker AMD is taking direct aim at its chief rival, Nvidia, with the launch of its latest hardware innovation: a complete rack-scale server system known as Helios. This highly advanced system has been engineered specifically to satisfy the massive computational demands of the world's largest and most prominent artificial intelligence laboratories. During the company's sold-out "Advancing AI" conference, held on Thursday in San Francisco, AMD Chair and CEO Dr. Lisa Su showcased and promoted the new Helios system. Su highlighted the company’s expanding roster of customers for this system, which includes technology giant Microsoft, as AMD prepares to begin shipping the hardware later this year. Alongside this announcement, Su also utilized the conference to promote the company’s newest chips, designed to power an AI industry that consumes vast amounts of processing power—an industry she colorfully described as a compute-hungry dragon.

What is a Rack-Scale System and How Does It Work in Data Centers?

Rack-scale systems integrate and unify numerous diverse processors into a single, cohesive, and exceptionally high-performance unit. These complex hardware systems are designed and built specifically for data centers, where they are utilized to train and run complex artificial intelligence models, as well as to handle other highly intensive, demanding workloads that require high-intensity computing.

During her keynote address at the conference, Dr. Lisa Su defined the Helios system as the "highest performance AI rack" in the entire technology industry. Su went on to explain that the innovative system was purpose-built to train and run the most demanding and advanced "frontier models" currently in existence on a massive and extraordinarily large scale. According to AMD's official statement, the Helios system is set to be deployed and operated by leading AI companies at a gigawatt-scale.

Competition Against Market Leaders and Helios Performance Metrics

Historically, Nvidia has dominated and continues to heavily dominate this market for data center server systems, primarily through its own rack-scale offerings known as Vera Rubin and Grace Blackwell. Now, AMD is making it explicitly clear that it intends to actively participate and directly compete in this market segment.

According to a report published by the professional technology publication The Register, performance metrics and benchmarks for the Helios system indicate that AMD has a genuine and tangible opportunity to challenge the market leader. The report notes that the Helios system actually surpasses Nvidia’s Vera Rubin system across several significant performance metrics.

The Helios system, which was first revealed to the public in 2025 and physically demonstrated on stage this past January at the global technology exhibition CES 2026, already boasts a number of prominent and well-recognized industry customers. This impressive list includes major giants such as OpenAI, Meta, Oracle, Anthropic, and Microsoft, all of which have defined plans to deploy and integrate the new system into their technological infrastructures.

Strategic Partnerships: Microsoft, Anthropic, and Other Giants

The announcements surrounding the Helios system were accompanied by several significant business moves from AMD’s key partners. Microsoft CEO Satya Nadella declared on Monday that the company intends to expand its Azure cloud platform infrastructure by integrating and deploying AMD's new Helios system.

In parallel, Anthropic and AMD announced a broad, large-scale strategic partnership last Wednesday. Under this collaboration, the companies plan to deploy graphics processing units (GPUs) with a massive capacity of up to two gigawatts, powered by the newly developed AMD rack system.

Unveiling the Venice-X Processor for Heavy Data Center Workloads

In addition to presenting the Helios system, AMD leveraged its stage on Thursday to introduce its new Venice-X processor. This CPU is specifically designed and engineered for use in large data centers to manage and process ultra-high-intensity computing workloads. According to the company’s announcement, the Venice-X processor is scheduled to officially launch and hit the market during the year 2027.

The Future of the Chip Market up to 2030 and the Impact of Agentic AI

During her address and commentary at the conference, Dr. Lisa Su addressed the growth trajectory and future of the global chip industry. Su asserted and estimated that by the year 2030, chips powering AI applications will become a massive and highly dominant segment of the overall computing market. According to her, this trend is occurring because the technology industry is currently witnessing a "step change in compute demand," which is heavily driven by the rise and breakthrough of agentic AI.

Su explained the operational nature of these AI agents in detail, highlighting the growing need for processing power: "When you ask the agent to do something, it actually has dozens of steps, and it has to reason, and it has to call tools, and it has to access data, and it has to keep doing it over and over until it solves the problem, and so you need lots of GPUs to do all that."

Following these remarks, Su noted: "We’re now expecting that by 2030, the AI accelerator market is going to reach about $1.4 trillion." The CEO emphasized the deep significance of this figure, pointing out that by the end of the decade, the AI accelerator market will approach the size and total volume of the entire global semiconductor market as it is valued today.

Finally, Dr. Lisa Su explained why graphics processors will continue to lead this trend: "We do expect that GPUs are going to make up the vast majority of that market because the algorithms are still very much in their infancy, and we’re still continuing to see the workloads change, and that favors programmability in the overall silicon ecosystem."

Questions & Answers

FAQ

This article was produced by our AI-assisted system: translation, summarization and business context based on original reporting by TechCrunch. Read about our editorial process. Link to the original source.

Enjoyed the article?

Subscribe to our newsletter for the latest AI updates straight to your inbox

More from TechCrunch

All articles from TechCrunch
כיצד מגבלות הבטיחות של הבינה המלאכותית מקשות על חוקרי סייבר התקפי
חדשות
5 דקות
מ־TechCrunch

כיצד מגבלות הבטיחות של הבינה המלאכותית מקשות על חוקרי סייבר התקפי

דיווח חדש מאתר TechCrunch חושף כי מגבלות הבטיחות (guardrails) המחמירות שהטילו ענקיות הבינה המלאכותית כדי למנוע שימוש לרעה במודלים שלהן, פוגעות כעת דווקא בחוקרי אבטחת מידע לגיטימיים ובמגני רשתות. חוקרים בתחום הסייבר ההתקפי, המשתמשים בכלים אלו לאיתור חולשות, מדווחים על חסימות תכופות וחוסר עקביות במענה של המערכות, אפילו במסגרת תוכניות סינון מיוחדות של OpenAI ו-Anthropic. בעקבות המגבלות והחשש מדליפת מידע לענן, חוקרים רבים נאלצים לנטוש את המערכות האמריקאיות המפוקחות ולעבור לשימוש במודלים מקומיים של קוד פתוח, לעיתים של חברות זרות כגון מודל GLM הסיני. מומחים מזהירים כי המגבלות הנוכחיות עלולות לפגוע ביכולת ההגנה של המערב אל מול גל מתקפות סייבר עתידי.

קרא עוד
Runway משיקה מנתב מודלים של בינה מלאכותית למדיה גנרטיבית
מוצר חדש
4 דקות
מ־TechCrunch

Runway משיקה מנתב מודלים של בינה מלאכותית למדיה גנרטיבית

חברת הסטארט-אפ Runway משיקה את Runway Media Router, כלי חדש המאפשר למפתחים לנתב באופן אוטומטי בקשות יצירת תמונה, וידאו ואודיו למודל המתאים ביותר. הכלי, שהושק באמצעות פלטפורמת המפתחים Runway Dev, מנתב את הבקשות על פי העדפות המפתח לגבי איכות, מהירות או עלות. המהלך מסמן את שאיפתה של Runway להפוך לשכבת תשתית ואורקסטרציה עבור תעשיית המדיה הגנרטיבית, בתקופה שבה השוק נעשה צפוף ותחרותי במיוחד והחברה מתמודדת עם תחרות עזה מצד ענקיות טכנולוגיה כמו גוגל, בייטדאנס ועליבאבא.

קרא עוד
אנבידיה שולחת מעבדים גרפיים לירח: שבבי Jetson של החברה בחלל
חדשות
3 דקות
מ־TechCrunch

אנבידיה שולחת מעבדים גרפיים לירח: שבבי Jetson של החברה בחלל

לפי דיווח באתר TechCrunch, חברת הסטארט-אפ Lunar Outpost תשתמש בשבבי Nvidia Jetson כדי לשלוט במערכת הלידאר של רכב השטח הירחי הבא שלה, מה שצפוי להפוך אותו למעבד הגרפי (GPU) הראשון על פני השטח של הירח. השימוש בפלטפורמת השבבים הזו נועד לסייע לרכבי השטח לעבד נתונים מחיישנים באופן מקומי ולקבל החלטות מהירות בסביבות הקיצוניות של הירח. בנוסף, אנבידיה משתפת פעולה עם Firefly Aerospace להפעלת לוויין עיבוד תמונות במסלול סביב הירח. המשימות מתוכננות לשיגור על גבי טילי Falcon 9 של SpaceX לפני סוף השנה הנוכחית, כחלק מהמאמצים הרחבים של נאס"א וחברות פרטיות לבסס נוכחות מתמשכת על הירח לקראת חזרת אסטרונאוטים המתוכננת לשנת 2028.

קרא עוד
סטארטאפ שבבי ה-AI של Etched גייס לפי שווי של 10.3 מיליארד דולר
חדשות
4 דקות
מ־TechCrunch

סטארטאפ שבבי ה-AI של Etched גייס לפי שווי של 10.3 מיליארד דולר

סטארטאפ שבבי הבינה המלאכותית Etched, שהוקם בשנת 2022 על ידי שלושה נושרי הרווארד, השלים סבב גיוס הון של 300 מיליון דולר לפי שווי חברה מרשים של 10.3 מיליארד דולר. סבב הגיוס הובל על ידי קרן Sequoia המפורסמת ובהשתתפות משקיעים בולטים כמו Andreessen Horowitz, SK Hynix, פיטר תיל ואנדריי קארפאטי. בכך הכפילה החברה את שוויה המוערך בתוך שבעה חודשים בלבד, לאחר שהוערכה ב-5 מיליארד דולר בדצמבר האחרון. החברה מפתחת מערכות שבבים ייעודיות לבינה מלאכותית, הכוללות פתרונות חומרה מתקדמים לייעול שלבי ההסקה (prefill ו-decode) במהירות גבוהה ובעלויות נמוכות.

קרא עוד

More articles you might like

All articles
Runway משיקה מנתב מודלים של בינה מלאכותית למדיה גנרטיבית
מוצר חדש
4 דקות
מ־TechCrunch

Runway משיקה מנתב מודלים של בינה מלאכותית למדיה גנרטיבית

חברת הסטארט-אפ Runway משיקה את Runway Media Router, כלי חדש המאפשר למפתחים לנתב באופן אוטומטי בקשות יצירת תמונה, וידאו ואודיו למודל המתאים ביותר. הכלי, שהושק באמצעות פלטפורמת המפתחים Runway Dev, מנתב את הבקשות על פי העדפות המפתח לגבי איכות, מהירות או עלות. המהלך מסמן את שאיפתה של Runway להפוך לשכבת תשתית ואורקסטרציה עבור תעשיית המדיה הגנרטיבית, בתקופה שבה השוק נעשה צפוף ותחרותי במיוחד והחברה מתמודדת עם תחרות עזה מצד ענקיות טכנולוגיה כמו גוגל, בייטדאנס ועליבאבא.

קרא עוד
גוגל מציגה את דגמי Gemini 3.6 Flash ו-3.5 Flash-Lite
מוצר חדש
5 דקות
מ־DeepMind

גוגל מציגה את דגמי Gemini 3.6 Flash ו-3.5 Flash-Lite

חברת גוגל הכריזה על השקת דגמי בינה מלאכותית חדשים בסדרת פלאש: Gemini 3.6 Flash, Gemini 3.5 Flash-Lite ו-Gemini 3.5 Flash Cyber. הדגמים מיועדים לפיתוח סוכני AI בייצור ומציעים שיפור ביעילות הטוקנים, מהירות וחיסכון בעלויות. דגם 3.6 Flash מפחית ב-17% את צריכת טוקני הפלט בהשוואה לגרסה הקודמת, ודגם 3.5 Flash-Lite פועל במהירות של 350 טוקני פלט לשנייה. כמו כן, דגם הסייבר הייעודי 3.5 Flash Cyber ישולב במערכת CodeMender ויופץ בשלב זה במתכונת פיילוט מוגבלת לממשלות ולשותפים מהימנים בלבד לצורך זיהוי ותיקון חולשות אבטחה.

קרא עוד
OpenAI משיקה מקלדת לניהול סוכני קוד ברקע מאבק משפטי
מוצר חדש
4 דקות
מ־TechCrunch

OpenAI משיקה מקלדת לניהול סוכני קוד ברקע מאבק משפטי

לפי דיווח ב-TechCrunch, חברת OpenAI נכנסת לשוק החומרה עם השקת מקלדת ה-Codex Micro במחיר של 230 דולר, המיועדת לניהול סוכני תכנות חצי-אוטונומיים. המקלדת פותחה בשיתוף עם Work Louder וכוללת מקשים מוארים להצגת סטטוס, ג'ויסטיק וחוגה לכיוונון רמת החשיבה של הסוכן. במקביל, דיווח של Bloomberg חשף כי החברה מפתחת רמקול חכם נייד ונטול מסך בעל חלקים נעים, המעוצב על ידי מהנדסי Apple לשעבר. פיתוח זה עומד במרכז תביעה שהגישה Apple נגד OpenAI בשבוע שעבר בטענה לגניבת סודות מסחריים, טענות ש-OpenAI מכחישה לחלוטין.

קרא עוד
אפליקציית Reelful משתמשת בבינה מלאכותית ליצירת סרטונים קצרים
מוצר חדש
4 דקות
מ־TechCrunch

אפליקציית Reelful משתמשת בבינה מלאכותית ליצירת סרטונים קצרים

אפליקציית iOS חדשה בשם Reelful, המשתתפת בתוכנית Speedrun של קרן a16z, משתמשת בבינה מלאכותית כדי להפוך אוטומטית תמונות וסרטונים מגלריית המכשיר לסרטוני וידאו מלוטשים לרשתות החברתיות, בסגנון טיקטוק ואינסטגרם רילס. האפליקציה פותחה על ידי קייט דיינקה, מהנדסת למידת מכונה לשעבר בסנאפצ'אט, במטרה לסייע למייסדים ולבעלי עסקים קטנים לייצר תוכן ולבנות מותג אישי בקלות ובמהירות ללא צורך בעריכה ידנית מורכבת. המשתמשים מזינים הנחיה, מקליטים דגימת קול לשיבוט קול ובוחרים חומרים מגלריית המצלמה. המערכת מתכננת את הסרטון, כותבת תסריט, מוסיפה קריינות ומוזיקה ויוצרת את הסרטון הסופי, כולל הנפשת תמונות סטטיות. האפליקציה זמינה כעת ב-iOS ומציעה רכישות חד-פעמיות ומסלולי מנוי.

קרא עוד