Anthropic in Talks with Samsung to Develop Custom AI Chips
News

Anthropic in Talks with Samsung to Develop Custom AI Chips

Aiming to reduce its reliance on Nvidia, AI giant Anthropic is exploring custom hardware development with Samsung

4 min read
Based on original reporting byTechCrunch ↗Translated and summarized by our AI-assisted news systemHow we work

✨Executive summary

Key Takeaways

  • Reports reveal that Anthropic is in talks with Samsung to develop its own custom AI chip.

  • The move follows OpenAI's announcement of its own inference chip, 'Jalapeño,' developed in partnership with Broadcom.

  • Transitioning to independent and custom hardware could reduce model running (inference) costs by up to 40%.

  • Samsung already serves as a major manufacturing partner for Nvidia and is building a dedicated chip factory with them in South Korea.

Anthropic in Talks with Samsung to Develop Custom AI Chips

  • Reports reveal that Anthropic is in talks with Samsung to develop its own custom AI...
  • The move follows OpenAI's announcement of its own inference chip, 'Jalapeño,' developed in partnership with...
  • Transitioning to independent and custom hardware could reduce model running (inference) costs by up to...
  • Samsung already serves as a major manufacturing partner for Nvidia and is building a dedicated...

Developing Custom AI Chips: Anthropic’s New Move with Samsung

According to recent reports, the AI laboratory Anthropic is in advanced talks with Samsung to develop a dedicated, custom AI chip. This strategic initiative is part of a broader, industry-wide push among tech giants to reduce their heavy reliance on Nvidia, the dominant chipmaker, and to optimize the execution of their massive models. The negotiations highlight an intensifying struggle for hardware independence and infrastructure autonomy across the artificial intelligence sector in 2026.

What is a Custom AI Chip?

A custom AI chip (Custom AI Chip) is a specialized hardware processor designed from the ground up to execute specific neural network and artificial intelligence tasks with maximum efficiency. In a business context, these custom application-specific integrated circuits (ASICs) run Large Language Models (LLMs) at far greater speeds and with significantly lower power consumption compared to general-purpose graphics processing units (GPUs).

For instance, technology giants Google and Amazon already offer dedicated Tensor Processing Units (TPUs) in their cloud ecosystems, successfully lowering computing overheads. According to recently published industry statistics, utilizing custom silicon can reduce the inference costs of running AI models by up to 40% compared to standard, off-the-shelf hardware. This economic shift is vital for enterprises deploying high-throughput applications.

The Talks Between Anthropic and Samsung and the Race for Hardware Independence

A report published by The Information indicates that Anthropic, an American AI safety and research company, has initiated discussions with South Korean electronics giant Samsung to explore a partnership to design and manufacture its new custom AI chip. This development follows reports from Reuters back in April, which noted that Anthropic was starting to consider in-house chip production to mitigate global hardware shortages.

However, at this stage, Anthropic has not made a final decision regarding the exact purpose of the chip, how it will be integrated into server racks, or its total computational power.

When contacted by TechCrunch for comment, Anthropic emphasized that a diversified hardware stack—incorporating processors manufactured by Google, Amazon, and Nvidia—will remain a fundamental pillar of its long-term computing strategy.

Industry experts suggest that Anthropic’s current exploration is a direct response to a major announcement by its chief rival, OpenAI (an AI research and deployment company). OpenAI recently partnered with American technology and semiconductor firm Broadcom to design its own dedicated inference processor, named "Jalapeño." The Jalapeño chip boasts vastly improved performance-per-watt and energy efficiency compared to standard chips. These industry movements underscore how crucial custom-fit infrastructure is becoming—in much the same way that modern enterprises deploy tailored business automation systems to maximize operational workflows and reduce friction.

The Broader Context: Why is Everyone Fleeing Nvidia?

While Nvidia remains the undisputed leader in the global AI chip market, its near-monopoly has created severe supply chain bottlenecks and exorbitant infrastructure costs for technology companies. Building custom hardware allows companies like Anthropic and OpenAI to not only break free from these supply chain constraints but also to design silicon that is perfectly optimized for their proprietary algorithms.

Samsung represents a critical player in this ecosystem. It already functions as a key manufacturing partner for Nvidia, supplying essential components and collaborating on a dedicated AI chip factory in South Korea. Furthermore, Samsung has held similar discussions with Google in the past regarding chip manufacturing, positioning itself as the ideal foundry partner for Anthropic's hardware ambitions.

Implications for Businesses in Israel

Although this semiconductor battle is taking place overseas, the development of custom AI chips will have direct, tangible implications for businesses and organizations in Israel. The local high-tech ecosystem, particularly companies developing advanced systems based on AI agents for business, is highly sensitive to fluctuating cloud computing costs. As giants like Anthropic develop more efficient, cheaper hardware, the API costs for utilizing market-leading models like Claude are expected to drop significantly in the long run.

Additionally, lower execution and inference costs will enable Israeli enterprises in the financial, healthcare, and retail sectors to host and run localized, secure models internally. This makes it far more viable to comply with strict Israeli Privacy Protection regulations without having to pay astronomical fees for expensive GPU servers.

What to Do Now

  1. Map Current Compute Expenses: Audit your current spending on AI APIs and cloud computing resources to understand how future infrastructure cost reductions will impact your overall budget.
  2. Implement a Multi-LLM Strategy: Avoid vendor lock-in. Designing a flexible software architecture that can seamlessly transition between models from OpenAI, Anthropic, and Google will allow you to capitalize on the price drops driven by their hardware breakthroughs.
  3. Evaluate Open-Source Alternatives: Begin testing the integration of open-source models like Meta's Llama on local servers or specialized cloud instances. This approach is becoming highly cost-effective as custom chips enter the market.

Looking Ahead

The global race for custom silicon proves that the AI revolution is not just a battle of software and algorithms—it is a physical war over energy, infrastructure, and silicon. For companies aiming to maintain their competitive edge, closely monitoring these developments and building flexible architecture is essential. Strategic investments in optimizing your technology infrastructure today will ensure a smoother, more profitable transition to the technologies of tomorrow.

Questions & Answers

FAQ

This article was produced by our AI-assisted system through translation, summarization, and automated quality controls based on original reporting by TechCrunch. Read about our editorial process. Link to the original source.

Get useful AI updates by email

A concise digest from our news desk.

More from TechCrunch

All articles from TechCrunch
OpenAI מציגה יכולות אפליקציה ב-ChatGPT וסוכני AI בשם Dots
חדשות
4 דקות
מ־TechCrunch

OpenAI מציגה יכולות אפליקציה ב-ChatGPT וסוכני AI בשם Dots

באירוע Dev Day הציגה OpenAI שורת עדכונים שמטרתם להפוך את ChatGPT לפלטפורמה לגילוי, להפעלה ולשימוש באפליקציות ובסוכני AI. החברה הודיעה על שילוב אפליקציות ישירות בשיחה, השקת "Sign in with ChatGPT" עם 16 שותפות ראשוניות, ופתיחת זירת מסחר לאפליקציות ארגוניות עם יותר מ-30 שותפות. בנוסף הושקו סוכני AI אוטונומיים בשם Dots, הפועלים בענן ומסוגלים להתחבר למעל 4,000 אפליקציות.

קרא עוד
כתבת TechCrunch יצרה אווטאר AI אינטראקטיבי באמצעות Synthesia
חדשות
4 דקות
מ־TechCrunch

כתבת TechCrunch יצרה אווטאר AI אינטראקטיבי באמצעות Synthesia

כתבת TechCrunch, דומיניק-מדורי דייוויס, יצרה אווטאר דיגיטלי אינטראקטיבי של עצמה בשיתוף סטארטאפ האווטארים Synthesia. החברה, שהגיעה להערכת שווי של 4 מיליארד דולר ופיתחה פלטפורמות הדרכה ותרגול מבוססות AI, בנתה עבור דייוויס אווטאר אישי המקריא תסריטים ואווטאר אינטראקטיבי המשיב לשאלות על מאמר שפרסמה. המערכת משלבת מודלי המרת דיבור לטקסט, מודל שפה סוכנותי, מודל המרת טקסט לקול ומודל וידאו להנפשה. דייוויס בחנה את המערכת עם בני משפחה וחברים והעלתה שאלות לגבי מקומם של אווטארים בעיתונות ובסביבה התאגידית.

קרא עוד
מודל Jev של TypeSafe AI: קבלת החלטות מהירה לאוטומציה ללא הזיות
מוצר חדש
4 דקות
מ־TechCrunch

מודל Jev של TypeSafe AI: קבלת החלטות מהירה לאוטומציה ללא הזיות

חברת TypeSafe AI, שהוקמה על ידי חוקר OpenAI לשעבר דיוגו אלמיידה, השיקה את Jev — מודל טרנספורמר חדש שאינו מפיק טקסט אלא הסתברויות והחלטות מכוילות. המודל מאפשר קבלת החלטות מהירה וזולה לאוטומציית תוכנה ללא סכנת הזיות, הודות להגדרת הפלטים מראש על ידי המשתמש ואימונו הבלעדי על נתונים סינתטיים. מפתחים מדווחים על שיפורי מהירות משמעותיים ועלויות נמוכות בהשוואה למודלי שפה מסורתיים.

קרא עוד
מילון מונחי AI מקיף: המושגים המרכזיים שצריך להכיר
ניתוח
4 דקות
מ־TechCrunch

מילון מונחי AI מקיף: המושגים המרכזיים שצריך להכיר

במדריך מושגים מקיף שפורסם ב-TechCrunch, מציגים כתבי האתר מילון מונחים מרכזי בעולם הבינה המלאכותית. המילון כולל הגדרות ברורות למונחים כמו AGI, סוכני AI, סוכני תכנות, ארכיטקטורת תערובת מומחים (MoE), פרוטוקול MCP לחיבור מקורות מידע, וטכניקת הישנות עמומה (Opaque recurrence) המייעלת עיבוד אך מעלה שאלות בטיחות ומעקב. בנוסף מפורטים תהליכי אימון, זיקוק, הסקה, מטמון זיכרון והשפעות המחסור בחומרת זיכרון המכונה RAMageddon.

קרא עוד

More articles you might like

All articles
OpenAI מציגה יכולות אפליקציה ב-ChatGPT וסוכני AI בשם Dots
חדשות
4 דקות
מ־TechCrunch

OpenAI מציגה יכולות אפליקציה ב-ChatGPT וסוכני AI בשם Dots

באירוע Dev Day הציגה OpenAI שורת עדכונים שמטרתם להפוך את ChatGPT לפלטפורמה לגילוי, להפעלה ולשימוש באפליקציות ובסוכני AI. החברה הודיעה על שילוב אפליקציות ישירות בשיחה, השקת "Sign in with ChatGPT" עם 16 שותפות ראשוניות, ופתיחת זירת מסחר לאפליקציות ארגוניות עם יותר מ-30 שותפות. בנוסף הושקו סוכני AI אוטונומיים בשם Dots, הפועלים בענן ומסוגלים להתחבר למעל 4,000 אפליקציות.

קרא עוד
שנה למעבדת המחקר של מיקרוסופט בסינגפור: קידום מחקר וטאלנטים ב-AI
חדשות
4 דקות
מ־Microsoft Research

שנה למעבדת המחקר של מיקרוסופט בסינגפור: קידום מחקר וטאלנטים ב-AI

לפי פרסום של Microsoft Research, מעבדת המחקר Microsoft Research Asia – Singapore השלימה שנה להקמתה כמעבדת המחקר הראשונה של מיקרוסופט בדרום-מזרח אסיה. במהלך השנה התמקדה המעבדה בארבעה תחומים מרכזיים: מודלי AI מתקדמים ומערכות סוכנים, בינה מלאכותית לתחומים ספציפיים, שיטות מחקר מבוססות AI ופיתוח טאלנטים. המעבדה יזמה תשעה פרויקטי מחקר חדשים עם האוניברסיטאות NUS ו-NTU, הכשירה מעל 300 סטודנטים בבית ספר לקיץ, והרחיבה שיתופי פעולה עם סוכנויות ממשלתיות בסינגפור כגון EDB ו-IMDA. בשנתה השנייה מתכננת המעבדה להרחיב את שיתופי הפעולה ולתרגם מחקר בסיסי ליישומים מעשיים.

קרא עוד
כתבת TechCrunch יצרה אווטאר AI אינטראקטיבי באמצעות Synthesia
חדשות
4 דקות
מ־TechCrunch

כתבת TechCrunch יצרה אווטאר AI אינטראקטיבי באמצעות Synthesia

כתבת TechCrunch, דומיניק-מדורי דייוויס, יצרה אווטאר דיגיטלי אינטראקטיבי של עצמה בשיתוף סטארטאפ האווטארים Synthesia. החברה, שהגיעה להערכת שווי של 4 מיליארד דולר ופיתחה פלטפורמות הדרכה ותרגול מבוססות AI, בנתה עבור דייוויס אווטאר אישי המקריא תסריטים ואווטאר אינטראקטיבי המשיב לשאלות על מאמר שפרסמה. המערכת משלבת מודלי המרת דיבור לטקסט, מודל שפה סוכנותי, מודל המרת טקסט לקול ומודל וידאו להנפשה. דייוויס בחנה את המערכת עם בני משפחה וחברים והעלתה שאלות לגבי מקומם של אווטארים בעיתונות ובסביבה התאגידית.

קרא עוד
סוכן ה-AI החדש של מטא Muse: אוטומציה אישית וסוגיות פרטיות
חדשות
4 דקות
מ־Wired

סוכן ה-AI החדש של מטא Muse: אוטומציה אישית וסוגיות פרטיות

סוכן הבינה המלאכותית החדש של מטא, Muse, הושק לביצוע משימות יומיומיות ואוטומציה אישית, כמו איתור מוצרים והזמנות באינטרנט. לפי דיווח ב-WIRED ונתוני Sensor Tower, האפליקציה נרשמה עם למעלה מ-900,000 הורדות בשבוע הראשון. הסוכן פועל באמצעות מכונה וירטואלית לגלישה באתרים, משתלב עם פייסבוק מרקטפלייס, אינסטגרם ו-WhatsApp, ומאחסן נתונים במסמך זיכרון. השימוש בכלי מעורר ביקורת מצד מומחי פרטיות בשל צירוף אוטומטי של אינטראקציות לאימון מודלים של מטא ובקשות חוזרות לחיבור מקורות מידע רגישים כגון חשבונות בנק ודואר אלקטרוני. מטא מצידה מבהירה כי המידע מנוקה מפרטים מזהים ומציעה הגדרות שליטה ידניות.

קרא עוד