[{"data":1,"prerenderedAt":-1},["ShallowReactive",2],{"article-anthropic-custom-ai-inference-chips-claude-en":3,"article-related-anthropic-custom-ai-inference-chips-claude-en":29,"series-industry-bf672b43-7312-47e4-aadf-9d09e153446f":76},{"id":4,"slug":5,"title":6,"content":7,"summary":8,"source":9,"source_url":10,"author":11,"image_url":12,"cover_image":12,"category":13,"language":14,"translated_content":11,"related_article_id":15,"keywords":16,"key_takeaways":22,"views":26,"created_at":27,"published_at":28,"topic_cluster_id":11},"bf672b43-7312-47e4-aadf-9d09e153446f","anthropic-custom-ai-inference-chips-claude-en","Anthropic is building custom AI chips for Claude","\u003Cp data-speakable=\"summary\">\u003Ca href=\"\u002Ftag\u002Fanthropic\">Anthropic\u003C\u002Fa> is building custom \u003Ca href=\"\u002Ftag\u002Finference\">inference\u003C\u002Fa> chips for \u003Ca href=\"\u002Ftag\u002Fclaude\">Claude\u003C\u002Fa> to cut dependence on Nvidia GPUs.\u003C\u002Fp>\u003Cp>Anthropic is moving from software to silicon, and the timing says a lot about where AI economics are headed. The company is reportedly building an in-house \u003Ca href=\"\u002Fnews\u002Fanthropic-builds-in-house-chip-team-claude-en\">chip team\u003C\u002Fa> for inference workloads, while \u003Ca href=\"https:\u002F\u002Fwww.samsung.com\u002Fsemiconductor\u002F\" target=\"_blank\" rel=\"noopener\">Samsung\u003C\u002Fa> has been mentioned as a possible manufacturing partner.\u003C\u002Fp>\u003Cp>This is not a \u003Ca href=\"\u002Fnews\u002Frust-serious-gpu-programming-language-en\">side project\u003C\u002Fa>. Anthropic, the company behind \u003Ca href=\"https:\u002F\u002Fwww.anthropic.com\u002Fclaude\" target=\"_blank\" rel=\"noopener\">Claude\u003C\u002Fa>, is joining a growing club of AI firms that want more control over the hardware under their models. The goal is simple: make inference faster, cheaper, and easier to scale than renting piles of expensive \u003Ca href=\"https:\u002F\u002Fwww.nvidia.com\u002Fen-us\u002Fdata-center\u002Fgpu-cloud-computing\u002F\" target=\"_blank\" rel=\"noopener\">Nvidia\u003C\u002Fa> GPUs.\u003C\u002Fp>\u003Ctable>\u003Cthead>\u003Ctr>\u003Cth>Fact\u003C\u002Fth>\u003Cth>Number\u003C\u002Fth>\u003Cth>Why it matters\u003C\u002Fth>\u003C\u002Ftr>\u003C\u002Fthead>\u003Ctbody>\u003Ctr>\u003Ctd>AI chip design market share held by Broadcom and Marvell\u003C\u002Ftd>\u003Ctd>~95%\u003C\u002Ftd>\u003Ctd>Shows how concentrated the co-design business is\u003C\u002Ftd>\u003C\u002Ftr>\u003Ctr>\u003Ctd>Broadcom backlog\u003C\u002Ftd>\u003Ctd>$73 billion\u003C\u002Ftd>\u003Ctd>Signals heavy demand for custom silicon work\u003C\u002Ftd>\u003C\u002Ftr>\u003Ctr>\u003Ctd>Broadcom expected annual AI chip revenue by end of 2027\u003C\u002Ftd>\u003Ctd>Over $100 billion\u003C\u002Ftd>\u003Ctd>Shows how large this market could get\u003C\u002Ftd>\u003C\u002Ftr>\u003Ctr>\u003Ctd>Marvell expected co-design revenue in 2026\u003C\u002Ftd>\u003Ctd>Upwards of $11 billion\u003C\u002Ftd>\u003Ctd>Shows how much money lives in inference chip contracts\u003C\u002Ftd>\u003C\u002Ftr>\u003C\u002Ftbody>\u003C\u002Ftable>\u003Ch2>Why Anthropic wants its own silicon\u003C\u002Fh2>\u003Cp>The immediate target is inference, the part of AI that answers prompts, powers agents, and keeps chatbots alive in production. Training still leans heavily on Nvidia hardware, but inference is different. It can be optimized around specific model behavior, specific memory patterns, and specific latency targets.\u003C\u002Fp>\n\u003Cfigure class=\"my-6\">\u003Cimg src=\"https:\u002F\u002Fxxdpdyhzhpamafnrdkyq.supabase.co\u002Fstorage\u002Fv1\u002Fobject\u002Fpublic\u002Fcovers\u002Finline-1786645985881-sbve.png\" alt=\"Anthropic is building custom AI chips for Claude\" class=\"rounded-xl w-full\" loading=\"lazy\" \u002F>\u003C\u002Ffigure>\n\u003Cp>That matters because Anthropic’s customers are not sending a few prompts a day anymore. They are running Claude in products, internal tools, and agent workflows that can burn through tokens at a pace that makes general-purpose GPUs look expensive. If you are serving that kind of load, every watt and every millisecond starts to matter.\u003C\u002Fp>\u003Cp>Anthropic told \u003Ca href=\"https:\u002F\u002Fwww.businessinsider.com\u002F\" target=\"_blank\" rel=\"noopener\">Business Insider\u003C\u002Fa> that the chips are meant to help Claude run faster and more efficiently at the scale customers need. That is the real story here: the company is trying to turn usage growth into better unit economics instead of letting hardware costs eat the margin.\u003C\u002Fp>\u003Cul>\u003Cli>Inference workloads can be tuned for specific model shapes and traffic patterns.\u003C\u002Fli>\u003Cli>Custom silicon can lower total cost of ownership by up to 65%, according to the reporting cited in the source article.\u003C\u002Fli>\u003Cli>Anthropic’s growth in consumer use and government contracts raises the pressure to control compute spend.\u003C\u002Fli>\u003Cli>Agentic AI systems consume more tokens than a single-shot chatbot prompt.\u003C\u002Fli>\u003C\u002Ful>\u003Ch2>The chip race is already crowded\u003C\u002Fh2>\u003Cp>Anthropic is late to a race that is already full of very serious players. \u003Ca href=\"https:\u002F\u002Fcloud.google.com\u002Ftpu\" target=\"_blank\" rel=\"noopener\">Google\u003C\u002Fa> has spent more than a decade on its Tensor Processing Unit line. \u003Ca href=\"https:\u002F\u002Fwww.amazon.com\u002Fawslabs\u002F\" target=\"_blank\" rel=\"noopener\">Amazon\u003C\u002Fa> built Trainium and Inferentia for its own cloud business. \u003Ca href=\"https:\u002F\u002Fwww.meta.com\u002F\" target=\"_blank\" rel=\"noopener\">Meta\u003C\u002Fa> has MTIA. \u003Ca href=\"https:\u002F\u002Fwww.microsoft.com\u002Fen-us\u002Fai\" target=\"_blank\" rel=\"noopener\">Microsoft\u003C\u002Fa> has Maia. Even \u003Ca href=\"https:\u002F\u002Fopenai.com\u002F\" target=\"_blank\" rel=\"noopener\">OpenAI\u003C\u002Fa> has been exploring custom inference hardware.\u003C\u002Fp>\u003Cp>The reason is easy to understand: if you are spending billions on AI, paying retail prices for GPUs forever starts to look irrational. The companies with the biggest workloads want chips designed around their own models, their own serving stacks, and their own \u003Ca href=\"\u002Fnews\u002Fanthropic-macquarie-gic-data-centers-en\">data centers\u003C\u002Fa>.\u003C\u002Fp>\u003Cblockquote>“The chips will allow Claude to run faster and more efficiently at the scale its customers need,” Anthropic told Business Insider.\u003C\u002Fblockquote>\u003Cp>That quote matters because it points to the real business case. This is not about bragging rights or owning every layer of the stack. It is about serving more requests with less electricity and less hardware waste.\u003C\u002Fp>\u003Ch2>Samsung, Broadcom, and Marvell each bring something different\u003C\u002Fh2>\u003Cp>Anthropic has not named its design partner, but the manufacturing side is where the industry gets interesting. The source article says \u003Ca href=\"https:\u002F\u002Fwww.samsung.com\u002Fsemiconductor\u002Ffoundry\u002F\" target=\"_blank\" rel=\"noopener\">Samsung Foundry\u003C\u002Fa> has been reported as the production partner. That would make sense if Anthropic wants a giant foundry with advanced process capacity and a willingness to support custom AI silicon.\u003C\u002Fp>\n\u003Cfigure class=\"my-6\">\u003Cimg src=\"https:\u002F\u002Fxxdpdyhzhpamafnrdkyq.supabase.co\u002Fstorage\u002Fv1\u002Fobject\u002Fpublic\u002Fcovers\u002Finline-1786645989977-x3e8.png\" alt=\"Anthropic is building custom AI chips for Claude\" class=\"rounded-xl w-full\" loading=\"lazy\" \u002F>\u003C\u002Ffigure>\n\u003Cp>On the design side, the market is dominated by \u003Ca href=\"https:\u002F\u002Fwww.broadcom.com\u002F\" target=\"_blank\" rel=\"noopener\">Broadcom\u003C\u002Fa> and \u003Ca href=\"https:\u002F\u002Fwww.marvell.com\u002F\" target=\"_blank\" rel=\"noopener\">Marvell\u003C\u002Fa>. The Tom’s Hardware report says those two firms account for roughly 95% of the ASIC co-design market. Broadcom has already worked with Google for years and has also been tied to \u003Ca href=\"\u002Ftag\u002Fopenai\">OpenAI\u003C\u002Fa>’s inference chip plans. Marvell has major contracts with Amazon and \u003Ca href=\"\u002Ftag\u002Fmicrosoft\">Microsoft\u003C\u002Fa>.\u003C\u002Fp>\u003Cp>These numbers explain why the chip business is so attractive to suppliers. Broadcom says it has a $73 billion backlog and expects more than $100 billion in annual AI chip revenue by the end of 2027. Marvell is expected to make upwards of $11 billion from these co-design jobs in 2026 alone. That is a lot of money flowing to the people who sell the shovels.\u003C\u002Fp>\u003Cul>\u003Cli>Broadcom and Marvell dominate custom ASIC co-design, with about 95% of the market.\u003C\u002Fli>\u003Cli>Broadcom’s backlog is already at $73 billion.\u003C\u002Fli>\u003Cli>Broadcom expects over $100 billion in annual AI chip revenue by the end of 2027.\u003C\u002Fli>\u003Cli>Marvell’s co-design revenue could exceed $11 billion in 2026.\u003C\u002Fli>\u003C\u002Ful>\u003Ch2>What this means for Nvidia and for Claude users\u003C\u002Fh2>\u003Cp>Nvidia is not going away. Training frontier models still depends heavily on its GPUs, and its software ecosystem remains the default for a lot of AI work. But inference is where the pressure is building, because that is where the bills pile up after the model is already trained.\u003C\u002Fp>\u003Cp>If Anthropic succeeds, Claude could become cheaper to run at scale, which would help the company compete harder in enterprise and government deployments. It could also give Anthropic more room to price aggressively without taking the same hit on inference margins.\u003C\u002Fp>\u003Cp>There is a catch, though. Designing chips is expensive, packaging is expensive, and software support is expensive. A custom ASIC only pays off if the workload is big enough and stable enough to justify the investment. Anthropic appears to believe Claude has crossed that threshold.\u003C\u002Fp>\u003Cp>That is probably the right bet. The AI companies with the largest token bills are the ones most likely to build their own hardware, and Anthropic now looks ready to join them. The open question is whether Samsung, Broadcom, or another partner ends up turning Claude’s workload into silicon that can actually beat the economics of off-the-shelf GPUs.\u003C\u002Fp>","Anthropic is building custom inference chips for Claude, with Samsung reported as a manufacturing partner and Nvidia costs in its crosshairs.","www.tomshardware.com","https:\u002F\u002Fwww.tomshardware.com\u002Ftech-industry\u002Fanthropic-to-build-its-own-co-designed-custom-ai-accelerator-for-inferencing-workloads-samsung-reported-to-be-partnering-with-the-claude-ai-maker-for-manufacturing",null,"https:\u002F\u002Fxxdpdyhzhpamafnrdkyq.supabase.co\u002Fstorage\u002Fv1\u002Fobject\u002Fpublic\u002Fcovers\u002Finline-1786645985881-sbve.png","industry","en","0ffd4803-74bb-43f8-a5be-0a5681ecc049",[17,18,19,20,21],"Anthropic","Claude","custom AI chips","inference ASIC","Samsung",[23,24,25],"Anthropic is hiring for a custom ASIC effort aimed at inference, not training.","The move mirrors Google, Amazon, Meta, Microsoft, and OpenAI.","Samsung has been reported as the manufacturing partner, but Anthropic has not confirmed it.",1,"2026-08-13T18:32:30.072113+00:00","2026-08-13T18:32:30.059+00:00",{"tags":30,"relatedLang":35,"relatedPosts":39},[31,33],{"name":17,"slug":32},"anthropic",{"name":18,"slug":34},"claude",{"id":15,"slug":36,"title":37,"language":38},"anthropic-custom-ai-inference-chips-claude-zh","Anthropic 也要自己做 AI 晶片","zh",[40,46,52,58,64,70],{"id":41,"slug":42,"title":43,"cover_image":44,"image_url":44,"created_at":45,"category":13},"f667de0f-bddf-4971-a0d3-23f1441adc22","openais-astra-pause-safety-model-releases-en","OpenAI’s Astra pause shows safety now shapes model releases","https:\u002F\u002Fxxdpdyhzhpamafnrdkyq.supabase.co\u002Fstorage\u002Fv1\u002Fobject\u002Fpublic\u002Fcovers\u002Finline-1786647770223-7lnz.png","2026-08-13T19:02:30.491753+00:00",{"id":47,"slug":48,"title":49,"cover_image":50,"image_url":50,"created_at":51,"category":13},"004596a9-1596-4845-99ee-ba713e42aa9d","moka-ai-hrms-choice-growing-teams-en","Moka AI Is a Stronger HRMS Choice for Growing Teams","https:\u002F\u002Fxxdpdyhzhpamafnrdkyq.supabase.co\u002Fstorage\u002Fv1\u002Fobject\u002Fpublic\u002Fcovers\u002Finline-1786611783121-ji6j.png","2026-08-13T09:02:28.700702+00:00",{"id":53,"slug":54,"title":55,"cover_image":56,"image_url":56,"created_at":57,"category":13},"c71ed61a-317a-440b-8ebd-6b8092334a4b","august-2026-market-review-ai-stocks-flipped-fast-en","August 2026 Market Review: AI Stocks Flipped Fast","https:\u002F\u002Fxxdpdyhzhpamafnrdkyq.supabase.co\u002Fstorage\u002Fv1\u002Fobject\u002Fpublic\u002Fcovers\u002Finline-1786584782010-3fqk.png","2026-08-13T01:32:37.446894+00:00",{"id":59,"slug":60,"title":61,"cover_image":62,"image_url":62,"created_at":63,"category":13},"512e86eb-fbce-4ce8-8b37-6bbe9d51ddd5","pixel-11-launch-live-google-reveals-en","Pixel 11 launch live: every Google reveal","https:\u002F\u002Fxxdpdyhzhpamafnrdkyq.supabase.co\u002Fstorage\u002Fv1\u002Fobject\u002Fpublic\u002Fcovers\u002Finline-1786581161788-lwbv.png","2026-08-13T00:32:20.660314+00:00",{"id":65,"slug":66,"title":67,"cover_image":68,"image_url":68,"created_at":69,"category":13},"e118a315-279d-496d-9efc-2b4e896b9feb","anthropic-invisible-watermark-right-move-en","Anthropic’s invisible watermark is the right move","https:\u002F\u002Fxxdpdyhzhpamafnrdkyq.supabase.co\u002Fstorage\u002Fv1\u002Fobject\u002Fpublic\u002Fcovers\u002Finline-1786579366798-6w22.png","2026-08-13T00:02:25.597044+00:00",{"id":71,"slug":72,"title":73,"cover_image":74,"image_url":74,"created_at":75,"category":13},"bb70741d-b446-4f24-8828-94d571538e3b","dockers-latest-releases-security-compose-en","Docker’s latest releases now center security and Compose","https:\u002F\u002Fxxdpdyhzhpamafnrdkyq.supabase.co\u002Fstorage\u002Fv1\u002Fobject\u002Fpublic\u002Fcovers\u002Finline-1786557770638-01iq.png","2026-08-12T18:02:27.492805+00:00",[77,82,87,92,97,102,107,112,117,122],{"id":78,"slug":79,"title":80,"created_at":81},"d35a1bd9-e709-412e-a2df-392df1dc572a","ai-impact-2026-developments-market-en","AI's Impact in 2026: Key Developments and Market Shifts","2026-03-25T16:20:33.205823+00:00",{"id":83,"slug":84,"title":85,"created_at":86},"5ed27921-5fd6-492e-8c59-78393bf37710","trumps-ai-legislative-framework-en","Trump's AI Legislative Framework: What's Inside?","2026-03-25T16:22:20.005325+00:00",{"id":88,"slug":89,"title":90,"created_at":91},"e454a642-f03c-4794-b185-5f651aebbaca","nvidia-gtc-2026-key-highlights-innovations-en","NVIDIA GTC 2026: Key Highlights and Innovations","2026-03-25T16:22:47.882615+00:00",{"id":93,"slug":94,"title":95,"created_at":96},"0ebb5b16-774a-4922-945d-5f2ce1df5a6d","claude-usage-diversifies-learning-curves-en","Claude Usage Diversifies, Learning Curves Emerge","2026-03-25T16:25:50.770376+00:00",{"id":98,"slug":99,"title":100,"created_at":101},"69934e86-2fc5-4280-8223-7b917a48ace8","openclaw-ai-commoditization-concerns-en","OpenClaw's Rise Raises Concerns of AI Model Commoditization","2026-03-25T16:26:30.582047+00:00",{"id":103,"slug":104,"title":105,"created_at":106},"b4b2575b-2ac8-46b2-b90e-ab1d7c060797","google-gemini-ai-rollout-2026-en","Google's Gemini AI Rollout Extended to 2026","2026-03-25T16:28:14.808842+00:00",{"id":108,"slug":109,"title":110,"created_at":111},"6e18bc65-42ae-4ad0-b564-67d7f66b979e","meta-llama4-fabricated-results-scandal-en","Meta's Llama 4 Scandal: Fabricated AI Test Results Unveiled","2026-03-25T16:29:15.482836+00:00",{"id":113,"slug":114,"title":115,"created_at":116},"bf888e9d-08be-4f47-996c-7b24b5ab3500","accenture-mistral-ai-deployment-en","Accenture and Mistral AI Team Up for AI Deployment","2026-03-25T16:31:01.894655+00:00",{"id":118,"slug":119,"title":120,"created_at":121},"5382b536-fad2-49c6-ac85-9eb2bae49f35","mistral-ai-high-stakes-2026-en","Mistral AI: Facing High Stakes in 2026","2026-03-25T16:31:39.941974+00:00",{"id":123,"slug":124,"title":125,"created_at":126},"9da3d2d6-b669-4971-ba1d-17fdb3548ed5","cursors-meteoric-rise-pressures-en","Cursor's Meteoric Rise Faces Industry Pressures","2026-03-25T16:32:21.899217+00:00"]