{"id":984285,"date":"2026-07-23T17:41:45","date_gmt":"2026-07-23T21:41:45","guid":{"rendered":"https:\/\/www.marketnewsdesk.com\/index.php\/amd-and-cerebras-announce-industry-leading-ultra-low-latency-and-high-throughput-ai-inference-solution\/"},"modified":"2026-07-23T17:41:45","modified_gmt":"2026-07-23T21:41:45","slug":"amd-and-cerebras-announce-industry-leading-ultra-low-latency-and-high-throughput-ai-inference-solution","status":"publish","type":"post","link":"https:\/\/www.marketnewsdesk.com\/index.php\/amd-and-cerebras-announce-industry-leading-ultra-low-latency-and-high-throughput-ai-inference-solution\/","title":{"rendered":"AMD and Cerebras Announce Industry-Leading Ultra-Low-Latency and High Throughput AI Inference Solution"},"content":{"rendered":"<div class=\"mw_release\">\n<p>\n        <strong>News Highlights<\/strong>\u00a0<\/p>\n<ul type=\"disc\">\n<li>\n          <em>AMD and Cerebras are collaborating to advance a workload-optimized approach to\u00a0ultra-low-latency AI inference infrastructure.\u00a0<\/em>\u00a0<\/li>\n<li>\n          <em><br \/>\n            <i>AMD Helios\u2122 and the Cerebras Wafer-Scale Engine will\u00a0operate\u00a0as a single disaggregated inference workflow, combining ultra-high-throughput from AMD Instinct\u2122 GPUs, with ultra-fast token generation of Cerebras Wafer-Scale Engine.<\/i><br \/>\n          <\/em>\n        <\/li>\n<li>\n          <em>Cerebras plans to deploy AMD Helios in its data centers, with the joint solution expected to be available first through Cerebras Cloud in the second half of 2026.<\/em>\u00a0<\/li>\n<\/ul>\n<p>SAN FRANCISCO and SUNNYVALE, Calif., July  23, 2026  (GLOBE NEWSWIRE) &#8212; AMD (NASDAQ: AMD) and Cerebras Systems (NASDAQ: CBRS) announced a technical partnership to deliver a new disaggregated AI inference solution that combines AMD Helios\u2122 rackscale solutions with the Cerebras Wafer-Scale Engine. Unveiled at Advancing AI 2026, the solution is designed to deliver the ultra-low latency required for the most advanced AI applications while dramatically increasing the throughput and efficiency.\u00a0<\/p>\n<p>The joint AMD and Cerebras solution will deploy AMD Helios alongside Cerebras Wafer-Scale Engine technology integrated in a single inference workflow for maximum performance and efficiency. AMD Helios\u00a0will provide a high-performance, scalable throughput engine. Cerebras Wafer-Scale Engine technology will provide ultra-fast, ultra-low latency decode and token generation. Together, the two compute engines are expected to deliver up to 5x higher\u00a0tokens\u00a0per\u00a0second\u00a0per\u00a0watt\u00a0(T\/s\/W)\u00a0<sup>i<\/sup>.\u00a0<\/p>\n<p>AI inference workloads increasingly have different requirements across latency, throughput, token capacity,\u00a0cost\u00a0and scale. High-volume workloads prioritize maximizing token generation, while coding, real-time copilots, live\u00a0agents\u00a0and agentic workflows demand faster response times. These differences are driving demand for heterogeneous infrastructure that matches compute technologies to specific workload requirements.\u00a0\u00a0<\/p>\n<p>The AMD and Cerebras solution addresses this challenge through disaggregated inference,\u00a0optimizing\u00a0the two primary stages of the workflow independently. AMD Helios provides ultra-high throughput, processing prompts and large context windows. The Cerebras Wafer-Scale Engine accelerates the memory-bandwidth-intensive token generation, with ultra-low latency. By connecting these best-in-class engines through one integrated workflow, the companies are creating a differentiated platform for ultra-low-latency inference without sacrificing throughput or scale.\u00a0<\/p>\n<p>\u201cAI inference is becoming one of the largest infrastructure opportunities in AI, and its growing diversity requires a more flexible approach,\u201d said\u00a0Dr. Lisa Su,\u00a0chair and CEO, AMD. \u201cAMD Helios\u00a0delivers leadership performance\u00a0and scale for the broadest range of\u00a0inference workloads.\u00a0Together\u00a0with Cerebras,\u00a0we are\u00a0extending\u00a0that leadership into the most latency-sensitive applications and creating\u00a0a powerful new platform for\u00a0real-time agentic AI.\u201d\u00a0\u00a0\u00a0<\/p>\n<p>\u00a0\u201cThe demand for ultra-fast inference is growing at an unprecedented pace. Cerebras delivers the world\u2019s fastest, ultra-low-latency inference,\u201d said Andrew Feldman, CEO and co-founder, Cerebras.\u00a0\u201cPartnering with AMD gives us an incredible opportunity to bring that performance to even more customers.\u201d\u00a0<\/p>\n<p>Fast token generation is becoming increasingly important as AI moves into software development, autonomous agents, robotics, scientific\u00a0discovery\u00a0and other applications where response time directly shapes the user experience and the usefulness of the system. The joint solution brings together complementary architectures purpose-built for these demands.\u00a0\u00a0<\/p>\n<p>AMD Helios provides the high-throughput prompt engine, rack-scale\u00a0efficiency\u00a0and deployment scale\u00a0required\u00a0to process large numbers of complex requests. Cerebras Wafer-Scale Engine technology provides the ultra-low-latency and decode performance needed to return tokens in real time. The result is a solution designed specifically for the ultra-low-latency segment of the inference market,\u00a0with AMD\u00a0Helios as the foundation for high-throughput and balanced inference workloads across the data center.\u00a0<\/p>\n<p>Cerebras plans to deploy AMD Helios systems in its data centers, with the joint\u00a0solution\u00a0expected to\u00a0become\u00a0available\u00a0initially through Cerebras Cloud\u00a0in the second half of 2026.\u00a0\u00a0<\/p>\n<p>\n        <strong>Supporting Resources<\/strong>\n      <\/p>\n<ul>\n<li>Follow AMD at <a href=\"https:\/\/newsroom.amd.com\/press-kits\/advancing-ai-2026-all-news\" rel=\"nofollow\" target=\"_blank\">Advancing AI 2026<\/a> (Press Kit)<\/li>\n<li>Learn more about <a href=\"https:\/\/www.amd.com\/en\/products\/rackscale-solutions\/helios.html\" rel=\"nofollow\" target=\"_blank\">AMD Helios\u2122 rackscale solution<\/a><\/li>\n<li>Learn more about <a href=\"https:\/\/www.amd.com\/en\/products\/accelerators\/instinct.html\" rel=\"nofollow\" target=\"_blank\">AMD Instinct\u2122 Accelerators<\/a><\/li>\n<li>Connect with AMD on <a href=\"https:\/\/www.amd.com\/en\/products\/accelerators\/instinct.html\" rel=\"nofollow\" target=\"_blank\">LinkedIn<\/a><\/li>\n<li>Follow AMD on <a href=\"https:\/\/x.com\/AMD?lang=en\" rel=\"nofollow\" target=\"_blank\">X<\/a><\/li>\n<li>Learn more about the <a href=\"https:\/\/www.cerebras.ai\/chip\" rel=\"nofollow\" target=\"_blank\">Cerebras Wafer-Scale Engine<\/a><\/li>\n<li>Connect with Cerebras on <a href=\"https:\/\/www.linkedin.com\/company\/cerebras-systems\/\" rel=\"nofollow\" target=\"_blank\">LinkedIn<\/a> | Follow Cerebras on <a href=\"https:\/\/x.com\/cerebras?lang=en\" rel=\"nofollow\" target=\"_blank\">X<\/a><\/li>\n<\/ul>\n<p>\n        <strong>About AMD<\/strong><br \/>\n        <br \/>AMD (NASDAQ: AMD) drives innovation in high-performance and AI computing to solve the world\u2019s most important challenges. Today, AMD technology powers billions of experiences across cloud and AI infrastructure, embedded systems, AI PCs and gaming. With a broad portfolio of AI-optimized CPUs, GPUs, networking and software, AMD delivers full-stack AI solutions that provide the performance and scalability needed for a new era of intelligent computing. Learn more at <a href=\"http:\/\/www.amd.com\" rel=\"nofollow\" target=\"_blank\">www.amd.com<\/a>.<\/p>\n<p>\n        <strong>About Cerebras Systems<\/strong>\n      <\/p>\n<p>Cerebras Systems (NASDAQ: CBRS) builds the world\u2019s fastest AI infrastructure. The Cerebras team of pioneering computer architects, computer scientists, AI researchers, and engineers of all types came together to make AI blisteringly fast through innovation and invention. We believe that when AI is fast, it will change the world. Leading global corporations, research institutes, and governments choose Cerebras to run their AI workloads. Cerebras solutions are available on premises and in the cloud. Visit\u00a0<a href=\"http:\/\/cerebras.ai\/\" rel=\"nofollow\" target=\"_blank\">cerebras.ai<\/a><a href=\"https:\/\/cerebras.ai\/\" rel=\"nofollow\" target=\"_blank\">\u00a0for more<\/a>.<\/p>\n<p>\n        <strong>AMD CAUTIONARY STATEMENT <\/strong>\n      <\/p>\n<p>This press release\u00a0contains\u00a0forward-looking statements concerning Advanced Micro Devices, Inc. (AMD) such as\u00a0the features, functionality, performance, availability,\u00a0scalability, deployment, timing\u00a0and expected benefits of AMD\u2019s collaboration and joint solution with Cerebras, which are made\u00a0pursuant to\u00a0the Safe Harbor provisions of the Private Securities Litigation Reform Act of 1995. Forward-looking statements are commonly\u00a0identified\u00a0by words such as &#8220;would,&#8221; &#8220;may,&#8221; &#8220;expects,&#8221; &#8220;believes,&#8221; &#8220;plans,&#8221; &#8220;intends,&#8221; &#8220;projects&#8221; and other terms with similar\u00a0meaning. Investors are cautioned that the forward-looking statements in this press release are based on current beliefs, assumptions and expectations, speak only as of the date of this press release and involve risks and uncertainties that could cause actual results to differ materially from current expectations. Such statements are subject to certain known and unknown risks and uncertainties, many of which are difficult to predict and are generally beyond AMD&#8217;s control, that could cause actual results and other future events to differ materially from those expressed in, or implied or projected by, the forward-looking information and statements. Material factors that could cause actual results to differ materially from current expectations include, without limitation, the following:\u00a0impact of government actions and regulations such as export regulations, import tariffs, trade protection measures, and licensing requirements;\u00a0competitive markets in which AMD\u2019s products are sold; the cyclical nature of the semiconductor\u00a0industry; market conditions of the\u00a0industries\u00a0in which AMD products are sold; AMD\u2019s ability to introduce products on a timely basis with expected features and performance levels; loss of a significant customer; economic and market\u00a0uncertainty; quarterly and seasonal sales patterns; AMD&#8217;s ability to adequately protect its technology or other intellectual property; unfavorable currency exchange rate fluctuations; ability of third party manufacturers to manufacture AMD&#8217;s products on a timely basis in sufficient quantities and using competitive technologies; availability of essential equipment, materials,\u00a0components (such as memory supply),\u00a0substrates or manufacturing processes; ability to achieve expected manufacturing yields for AMD\u2019s products; AMD&#8217;s ability to generate revenue from its semi-custom SoC products; potential security vulnerabilities; potential security incidents including IT outages, data loss, data breaches and cyberattacks; uncertainties involving the ordering and shipment of AMD\u2019s products; AMD\u2019s reliance on third-party intellectual property to design and introduce new products; AMD&#8217;s reliance on third-party companies for design, manufacture and supply of motherboards, software, memory and other computer platform components; AMD&#8217;s reliance on Microsoft and other software vendors&#8217; support to design and develop software to run on AMD\u2019s products; AMD\u2019s reliance on third-party distributors and add-in-board partners; impact of modification or interruption of AMD\u2019s internal business processes and information systems; compatibility of AMD\u2019s products with some or all\u00a0industry-standard software and hardware; costs related to defective products; failure to maintain an efficient supply chain\u00a0as customer demand changes; AMD&#8217;s ability to rely on third party supply-chain logistics functions; AMD\u2019s ability to effectively control sales of its products on the gray market; impact of climate change on AMD\u2019s business; AMD\u2019s ability to realize its deferred tax assets; potential tax liabilities; current and future claims and litigation; impact of environmental laws, conflict minerals related provisions and other laws or regulations; evolving expectations from governments, investors, customers and other stakeholders regarding corporate responsibility matters; issues related to the responsible use of AI; restrictions imposed by agreements governing AMD\u2019s notes, the guarantees of Xilinx\u2019s notes and the revolving credit agreement; AMD\u2019s ability to satisfy financial obligations under guarantees, leases\u00a0and other commercial commitments; impact of acquisitions, joint ventures and\/or\u00a0investments on AMD\u2019s business and AMD\u2019s ability to integrate acquired businesses; impact of any impairment of the combined company\u2019s assets; political, legal and economic risks and natural disasters; future impairments of technology license purchases; AMD\u2019s ability to attract and retain key employees; and AMD\u2019s stock price volatility. Investors are urged to review in detail the risks and uncertainties in AMD\u2019s Securities and Exchange Commission filings, including but not limited to AMD\u2019s most recent reports on Forms 10-K and 10-Q.\u00a0\u00a0\u00a0<\/p>\n<p>\n        <strong>CEREBRAS DISCLOSURE INFORMATION<\/strong>\n      <\/p>\n<p>Cerebras uses its investor relations page (investors.cerebras.ai), its X account (@cerebras), and its LinkedIn page (linkedin.com\/company\/cerebras-systems\/) to disclose material non-public information and for complying with its disclosure obligations under Regulation FD. Accordingly, investors should monitor these channels, in addition to following Cerebras\u2019 press releases, Securities and Exchange Commission (SEC) filings, public conference calls and public webcasts.<\/p>\n<p>\n        <strong>Forward-Looking Statements<\/strong>\n      <\/p>\n<p>This press release contains \u201cforward-looking statements\u201d within the meaning of applicable securities laws. All statements other than statements of historical fact could be deemed to be forward-looking, including, but not limited to, statements regarding the features, capacity, scalability, performance, timing, data center deployment and implementation, costs and expected benefits and opportunities associated with Cerebras&#8217; collaboration and joint solution with AMD, and any assumptions relating to the foregoing. The words \u201cmay,\u201d \u201cwill,\u201d \u201cshall,\u201d \u201cshould,\u201d \u201cexpects,\u201d \u201cplans,\u201d \u201canticipates,\u201d \u201ccould,\u201d \u201cintends,\u201d \u201ctarget,\u201d \u201cprojects,\u201d \u201ccontemplates,\u201d \u201cbelieves,\u201d \u201cestimates,\u201d \u201cpredicts,\u201d \u201cpotential,\u201d \u201cobjective,\u201d or \u201ccontinue,\u201d or the negative of these words or other similar terms or expressions that concern our expectations, strategy, plans, or intentions are intended to identify forward-looking statements, although not all forward-looking statements contain these identifying words. These forward-looking statements are subject to a number of risks and uncertainties, many of which involve factors or circumstances that are beyond Cerebras\u2019 control. These risks and uncertainties include, but are not limited to: Cerebras\u2019 ability to sustain and manage its growth, access borrowings and other sources of capital on acceptable terms, and deploy available capital to support growth; its history of net losses and ability to achieve and maintain profitability; its limited operating history at its current scale and ability to accurately forecast revenue and appropriately budget and manage expenses; its dependence on a limited number of significant customers, including OpenAI, Group 42 Holding Ltd, Mohamed bin Zayed University of Artificial Intelligence, and AWS, and the potential impact of any reduction in demand from, material adverse development in its relationships with, or failure to meet its obligations to, such customers, including under its Master Relationship Agreement with OpenAI; the timing, execution and expected benefits of its strategic customer, partner and financing arrangements; its historical reliance on sales of hardware systems and the early-stage, rapidly evolving market for its cloud-based offerings and AI infrastructure; its ability to secure sufficient data center capacity and capital to support its cloud-based offerings; its ability to launch new offerings and add new product capabilities; and its ability to compete effectively in the rapidly evolving and competitive market for AI computing solutions.<\/p>\n<p>Cerebras\u2019 actual results could differ materially from those stated or implied in forward-looking statements due to a number of factors. Accordingly, undue reliance should not be placed on such statements. These forward-looking statements are made as of the date they were first issued and are based on information available to Cerebras together with Cerebras\u2019 expectations, estimates, forecasts, projections, beliefs, and assumptions as of such date. These forward-looking statements should not be relied upon as representing Cerebras\u2019 views as of any date subsequent to the date of this press release. Past performance is not necessarily indicative of future results. Cerebras undertakes no intention or obligation to update or revise any forward-looking statements, whether as a result of new information, future events, or otherwise, except as required by law.<\/p>\n<p>Further information on potential risks that could affect actual results is included in Cerebras\u2019 most recent filings with the Securities and Exchange Commission (the \u201cSEC\u201d), including in Cerebras\u2019 most recent Quarterly Report on Form 10-Q, copies of which may be obtained by visiting Cerebras\u2019 Investor Relations website at investors.cerebras.ai or the SEC\u2019s website at www.sec.gov.<\/p>\n<p align=\"left\">\n        <strong>Contacts:<\/strong><br \/>\n        <br \/>\n        <strong>Aaron Grabein<\/strong><br \/>\n        <br \/>AMD Communications<br \/>737-256-9518<br \/><a href=\"mailto:aaron.grabein@amd.com\" rel=\"nofollow\" target=\"_blank\">aaron.grabein@amd.com<\/a><\/p>\n<p align=\"left\">Liz Stine<br \/>AMD Investor Relations<br \/>720-652-3965<br \/><a href=\"mailto:liz.stine@amd.com\" rel=\"nofollow\" target=\"_blank\"><u>liz.stine@amd.com<\/u><\/a><\/p>\n<p align=\"left\">Kriselle Laran<br \/>Cerebras<br \/><a href=\"mailto:pr@cerebras.ai\" rel=\"nofollow\" target=\"_blank\"><u>pr@cerebras.ai<\/u><\/a><\/p>\n<p align=\"left\">Sean Dorsey<br \/>Cerebras Investor Relations<br \/><a href=\"mailto:investors@cerebras.ai\" rel=\"nofollow\" target=\"_blank\">investors@cerebras.ai<\/a><\/p>\n<p align=\"left\">_______________<br \/><sup>i<\/sup> Based on modelling by AMD Performance Labs and Cerebras in July 2026 to determine tokens per second per kilowatt (TPS\/kW) at a comparable interactivity point with Kimi 2.6 1T Model comparing an AMD Helios rackscale solution with Cerebras WSE to a Cerebras WSE-only configuration. System manufacturers may vary configurations, yielding different results. MI400-021<\/p>\n<p>      <img decoding=\"async\" alt=\"\" class=\"__GNW8366DE3E__IMG\" src=\"https:\/\/www.globenewswire.com\/newsroom\/ti?nf=OTc2NzI3OSM3NzIwNzQ4IzIwMDcxMTY=\" \/><br \/>\n      <br \/>\n      <img decoding=\"async\" alt=\"\" src=\"https:\/\/ml.globenewswire.com\/media\/ZWY1NWFkNjMtOWEyYy00Y2RjLWIzNzktMzliNjc1NDUxNzdkLTEwMTg2ODktMjAyNi0wNy0yMy1lbg==\/tiny\/Advanced-Micro-Devices-Inc-.png\" \/>\n    <\/div>\n<div class=\"mw_contactinfo\"><\/div>\n","protected":false},"excerpt":{"rendered":"<p>News Highlights\u00a0 AMD and Cerebras are collaborating to advance a workload-optimized approach to\u00a0ultra-low-latency AI inference infrastructure.\u00a0\u00a0 AMD Helios\u2122 and the Cerebras Wafer-Scale Engine will\u00a0operate\u00a0as a single disaggregated inference workflow, combining ultra-high-throughput from AMD Instinct\u2122 GPUs, with ultra-fast token generation of Cerebras Wafer-Scale Engine. Cerebras plans to deploy AMD Helios in its data centers, with the joint solution expected to be available first through Cerebras Cloud in the second half of 2026.\u00a0 SAN FRANCISCO and SUNNYVALE, Calif., July 23, 2026 (GLOBE NEWSWIRE) &#8212; AMD (NASDAQ: AMD) and Cerebras Systems (NASDAQ: CBRS) announced a technical partnership to deliver a new disaggregated AI inference solution that combines AMD Helios\u2122 rackscale solutions with the Cerebras Wafer-Scale Engine. Unveiled at Advancing AI 2026, the solution &hellip; <\/p>\n<p class=\"link-more\"><a href=\"https:\/\/www.marketnewsdesk.com\/index.php\/amd-and-cerebras-announce-industry-leading-ultra-low-latency-and-high-throughput-ai-inference-solution\/\" class=\"more-link\">Continue reading<span class=\"screen-reader-text\"> &#8220;AMD and Cerebras Announce Industry-Leading Ultra-Low-Latency and High Throughput AI Inference Solution&#8221;<\/span><\/a><\/p>\n","protected":false},"author":2,"featured_media":0,"comment_status":"closed","ping_status":"closed","sticky":false,"template":"","format":"standard","meta":{"footnotes":""},"categories":[],"tags":[],"class_list":["post-984285","post","type-post","status-publish","format-standard","hentry"],"yoast_head":"<!-- This site is optimized with the Yoast SEO plugin v28.1 - https:\/\/yoast.com\/product\/yoast-seo-wordpress\/ -->\n<title>AMD and Cerebras Announce Industry-Leading Ultra-Low-Latency and High Throughput AI Inference Solution - Market Newsdesk<\/title>\n<meta name=\"robots\" content=\"index, follow, max-snippet:-1, max-image-preview:large, max-video-preview:-1\" \/>\n<link rel=\"canonical\" href=\"https:\/\/www.marketnewsdesk.com\/index.php\/amd-and-cerebras-announce-industry-leading-ultra-low-latency-and-high-throughput-ai-inference-solution\/\" \/>\n<meta property=\"og:locale\" content=\"en_US\" \/>\n<meta property=\"og:type\" content=\"article\" \/>\n<meta property=\"og:title\" content=\"AMD and Cerebras Announce Industry-Leading Ultra-Low-Latency and High Throughput AI Inference Solution - Market Newsdesk\" \/>\n<meta property=\"og:description\" content=\"News Highlights\u00a0 AMD and Cerebras are collaborating to advance a workload-optimized approach to\u00a0ultra-low-latency AI inference infrastructure.\u00a0\u00a0 AMD Helios\u2122 and the Cerebras Wafer-Scale Engine will\u00a0operate\u00a0as a single disaggregated inference workflow, combining ultra-high-throughput from AMD Instinct\u2122 GPUs, with ultra-fast token generation of Cerebras Wafer-Scale Engine. Cerebras plans to deploy AMD Helios in its data centers, with the joint solution expected to be available first through Cerebras Cloud in the second half of 2026.\u00a0 SAN FRANCISCO and SUNNYVALE, Calif., July 23, 2026 (GLOBE NEWSWIRE) &#8212; AMD (NASDAQ: AMD) and Cerebras Systems (NASDAQ: CBRS) announced a technical partnership to deliver a new disaggregated AI inference solution that combines AMD Helios\u2122 rackscale solutions with the Cerebras Wafer-Scale Engine. Unveiled at Advancing AI 2026, the solution &hellip; Continue reading &quot;AMD and Cerebras Announce Industry-Leading Ultra-Low-Latency and High Throughput AI Inference Solution&quot;\" \/>\n<meta property=\"og:url\" content=\"https:\/\/www.marketnewsdesk.com\/index.php\/amd-and-cerebras-announce-industry-leading-ultra-low-latency-and-high-throughput-ai-inference-solution\/\" \/>\n<meta property=\"og:site_name\" content=\"Market Newsdesk\" \/>\n<meta property=\"article:published_time\" content=\"2026-07-23T21:41:45+00:00\" \/>\n<meta property=\"og:image\" content=\"https:\/\/www.globenewswire.com\/newsroom\/ti?nf=OTc2NzI3OSM3NzIwNzQ4IzIwMDcxMTY=\" \/>\n<meta name=\"author\" content=\"Newsdesk\" \/>\n<meta name=\"twitter:card\" content=\"summary_large_image\" \/>\n<meta name=\"twitter:label1\" content=\"Written by\" \/>\n\t<meta name=\"twitter:data1\" content=\"Newsdesk\" \/>\n\t<meta name=\"twitter:label2\" content=\"Est. reading time\" \/>\n\t<meta name=\"twitter:data2\" content=\"11 minutes\" \/>\n<script type=\"application\/ld+json\" class=\"yoast-schema-graph\">{\"@context\":\"https:\\\/\\\/schema.org\",\"@graph\":[{\"@type\":\"Article\",\"@id\":\"https:\\\/\\\/www.marketnewsdesk.com\\\/index.php\\\/amd-and-cerebras-announce-industry-leading-ultra-low-latency-and-high-throughput-ai-inference-solution\\\/#article\",\"isPartOf\":{\"@id\":\"https:\\\/\\\/www.marketnewsdesk.com\\\/index.php\\\/amd-and-cerebras-announce-industry-leading-ultra-low-latency-and-high-throughput-ai-inference-solution\\\/\"},\"author\":{\"name\":\"Newsdesk\",\"@id\":\"https:\\\/\\\/www.marketnewsdesk.com\\\/#\\\/schema\\\/person\\\/482f27a394d4fda80ecb5499e519d979\"},\"headline\":\"AMD and Cerebras Announce Industry-Leading Ultra-Low-Latency and High Throughput AI Inference Solution\",\"datePublished\":\"2026-07-23T21:41:45+00:00\",\"mainEntityOfPage\":{\"@id\":\"https:\\\/\\\/www.marketnewsdesk.com\\\/index.php\\\/amd-and-cerebras-announce-industry-leading-ultra-low-latency-and-high-throughput-ai-inference-solution\\\/\"},\"wordCount\":2155,\"image\":{\"@id\":\"https:\\\/\\\/www.marketnewsdesk.com\\\/index.php\\\/amd-and-cerebras-announce-industry-leading-ultra-low-latency-and-high-throughput-ai-inference-solution\\\/#primaryimage\"},\"thumbnailUrl\":\"https:\\\/\\\/www.globenewswire.com\\\/newsroom\\\/ti?nf=OTc2NzI3OSM3NzIwNzQ4IzIwMDcxMTY=\",\"inLanguage\":\"en-US\"},{\"@type\":\"WebPage\",\"@id\":\"https:\\\/\\\/www.marketnewsdesk.com\\\/index.php\\\/amd-and-cerebras-announce-industry-leading-ultra-low-latency-and-high-throughput-ai-inference-solution\\\/\",\"url\":\"https:\\\/\\\/www.marketnewsdesk.com\\\/index.php\\\/amd-and-cerebras-announce-industry-leading-ultra-low-latency-and-high-throughput-ai-inference-solution\\\/\",\"name\":\"AMD and Cerebras Announce Industry-Leading Ultra-Low-Latency and High Throughput AI Inference Solution - Market Newsdesk\",\"isPartOf\":{\"@id\":\"https:\\\/\\\/www.marketnewsdesk.com\\\/#website\"},\"primaryImageOfPage\":{\"@id\":\"https:\\\/\\\/www.marketnewsdesk.com\\\/index.php\\\/amd-and-cerebras-announce-industry-leading-ultra-low-latency-and-high-throughput-ai-inference-solution\\\/#primaryimage\"},\"image\":{\"@id\":\"https:\\\/\\\/www.marketnewsdesk.com\\\/index.php\\\/amd-and-cerebras-announce-industry-leading-ultra-low-latency-and-high-throughput-ai-inference-solution\\\/#primaryimage\"},\"thumbnailUrl\":\"https:\\\/\\\/www.globenewswire.com\\\/newsroom\\\/ti?nf=OTc2NzI3OSM3NzIwNzQ4IzIwMDcxMTY=\",\"datePublished\":\"2026-07-23T21:41:45+00:00\",\"author\":{\"@id\":\"https:\\\/\\\/www.marketnewsdesk.com\\\/#\\\/schema\\\/person\\\/482f27a394d4fda80ecb5499e519d979\"},\"breadcrumb\":{\"@id\":\"https:\\\/\\\/www.marketnewsdesk.com\\\/index.php\\\/amd-and-cerebras-announce-industry-leading-ultra-low-latency-and-high-throughput-ai-inference-solution\\\/#breadcrumb\"},\"inLanguage\":\"en-US\",\"potentialAction\":[{\"@type\":\"ReadAction\",\"target\":[\"https:\\\/\\\/www.marketnewsdesk.com\\\/index.php\\\/amd-and-cerebras-announce-industry-leading-ultra-low-latency-and-high-throughput-ai-inference-solution\\\/\"]}]},{\"@type\":\"ImageObject\",\"inLanguage\":\"en-US\",\"@id\":\"https:\\\/\\\/www.marketnewsdesk.com\\\/index.php\\\/amd-and-cerebras-announce-industry-leading-ultra-low-latency-and-high-throughput-ai-inference-solution\\\/#primaryimage\",\"url\":\"https:\\\/\\\/www.globenewswire.com\\\/newsroom\\\/ti?nf=OTc2NzI3OSM3NzIwNzQ4IzIwMDcxMTY=\",\"contentUrl\":\"https:\\\/\\\/www.globenewswire.com\\\/newsroom\\\/ti?nf=OTc2NzI3OSM3NzIwNzQ4IzIwMDcxMTY=\"},{\"@type\":\"BreadcrumbList\",\"@id\":\"https:\\\/\\\/www.marketnewsdesk.com\\\/index.php\\\/amd-and-cerebras-announce-industry-leading-ultra-low-latency-and-high-throughput-ai-inference-solution\\\/#breadcrumb\",\"itemListElement\":[{\"@type\":\"ListItem\",\"position\":1,\"name\":\"Home\",\"item\":\"https:\\\/\\\/www.marketnewsdesk.com\\\/\"},{\"@type\":\"ListItem\",\"position\":2,\"name\":\"AMD and Cerebras Announce Industry-Leading Ultra-Low-Latency and High Throughput AI Inference Solution\"}]},{\"@type\":\"WebSite\",\"@id\":\"https:\\\/\\\/www.marketnewsdesk.com\\\/#website\",\"url\":\"https:\\\/\\\/www.marketnewsdesk.com\\\/\",\"name\":\"Market Newsdesk\",\"description\":\"Latest Business News in Real Time\",\"potentialAction\":[{\"@type\":\"SearchAction\",\"target\":{\"@type\":\"EntryPoint\",\"urlTemplate\":\"https:\\\/\\\/www.marketnewsdesk.com\\\/?s={search_term_string}\"},\"query-input\":{\"@type\":\"PropertyValueSpecification\",\"valueRequired\":true,\"valueName\":\"search_term_string\"}}],\"inLanguage\":\"en-US\"},{\"@type\":\"Person\",\"@id\":\"https:\\\/\\\/www.marketnewsdesk.com\\\/#\\\/schema\\\/person\\\/482f27a394d4fda80ecb5499e519d979\",\"name\":\"Newsdesk\",\"image\":{\"@type\":\"ImageObject\",\"inLanguage\":\"en-US\",\"@id\":\"https:\\\/\\\/secure.gravatar.com\\\/avatar\\\/a0d0bd5b0f0ca12a265a459b13169dac35f33776d8501eda5e68844a366f2f46?s=96&d=mm&r=g\",\"url\":\"https:\\\/\\\/secure.gravatar.com\\\/avatar\\\/a0d0bd5b0f0ca12a265a459b13169dac35f33776d8501eda5e68844a366f2f46?s=96&d=mm&r=g\",\"contentUrl\":\"https:\\\/\\\/secure.gravatar.com\\\/avatar\\\/a0d0bd5b0f0ca12a265a459b13169dac35f33776d8501eda5e68844a366f2f46?s=96&d=mm&r=g\",\"caption\":\"Newsdesk\"},\"url\":\"https:\\\/\\\/www.marketnewsdesk.com\\\/index.php\\\/author\\\/newsdesk\\\/\"}]}<\/script>\n<!-- \/ Yoast SEO plugin. -->","yoast_head_json":{"title":"AMD and Cerebras Announce Industry-Leading Ultra-Low-Latency and High Throughput AI Inference Solution - Market Newsdesk","robots":{"index":"index","follow":"follow","max-snippet":"max-snippet:-1","max-image-preview":"max-image-preview:large","max-video-preview":"max-video-preview:-1"},"canonical":"https:\/\/www.marketnewsdesk.com\/index.php\/amd-and-cerebras-announce-industry-leading-ultra-low-latency-and-high-throughput-ai-inference-solution\/","og_locale":"en_US","og_type":"article","og_title":"AMD and Cerebras Announce Industry-Leading Ultra-Low-Latency and High Throughput AI Inference Solution - Market Newsdesk","og_description":"News Highlights\u00a0 AMD and Cerebras are collaborating to advance a workload-optimized approach to\u00a0ultra-low-latency AI inference infrastructure.\u00a0\u00a0 AMD Helios\u2122 and the Cerebras Wafer-Scale Engine will\u00a0operate\u00a0as a single disaggregated inference workflow, combining ultra-high-throughput from AMD Instinct\u2122 GPUs, with ultra-fast token generation of Cerebras Wafer-Scale Engine. Cerebras plans to deploy AMD Helios in its data centers, with the joint solution expected to be available first through Cerebras Cloud in the second half of 2026.\u00a0 SAN FRANCISCO and SUNNYVALE, Calif., July 23, 2026 (GLOBE NEWSWIRE) &#8212; AMD (NASDAQ: AMD) and Cerebras Systems (NASDAQ: CBRS) announced a technical partnership to deliver a new disaggregated AI inference solution that combines AMD Helios\u2122 rackscale solutions with the Cerebras Wafer-Scale Engine. Unveiled at Advancing AI 2026, the solution &hellip; Continue reading \"AMD and Cerebras Announce Industry-Leading Ultra-Low-Latency and High Throughput AI Inference Solution\"","og_url":"https:\/\/www.marketnewsdesk.com\/index.php\/amd-and-cerebras-announce-industry-leading-ultra-low-latency-and-high-throughput-ai-inference-solution\/","og_site_name":"Market Newsdesk","article_published_time":"2026-07-23T21:41:45+00:00","og_image":[{"url":"https:\/\/www.globenewswire.com\/newsroom\/ti?nf=OTc2NzI3OSM3NzIwNzQ4IzIwMDcxMTY=","type":"","width":"","height":""}],"author":"Newsdesk","twitter_card":"summary_large_image","twitter_misc":{"Written by":"Newsdesk","Est. reading time":"11 minutes"},"schema":{"@context":"https:\/\/schema.org","@graph":[{"@type":"Article","@id":"https:\/\/www.marketnewsdesk.com\/index.php\/amd-and-cerebras-announce-industry-leading-ultra-low-latency-and-high-throughput-ai-inference-solution\/#article","isPartOf":{"@id":"https:\/\/www.marketnewsdesk.com\/index.php\/amd-and-cerebras-announce-industry-leading-ultra-low-latency-and-high-throughput-ai-inference-solution\/"},"author":{"name":"Newsdesk","@id":"https:\/\/www.marketnewsdesk.com\/#\/schema\/person\/482f27a394d4fda80ecb5499e519d979"},"headline":"AMD and Cerebras Announce Industry-Leading Ultra-Low-Latency and High Throughput AI Inference Solution","datePublished":"2026-07-23T21:41:45+00:00","mainEntityOfPage":{"@id":"https:\/\/www.marketnewsdesk.com\/index.php\/amd-and-cerebras-announce-industry-leading-ultra-low-latency-and-high-throughput-ai-inference-solution\/"},"wordCount":2155,"image":{"@id":"https:\/\/www.marketnewsdesk.com\/index.php\/amd-and-cerebras-announce-industry-leading-ultra-low-latency-and-high-throughput-ai-inference-solution\/#primaryimage"},"thumbnailUrl":"https:\/\/www.globenewswire.com\/newsroom\/ti?nf=OTc2NzI3OSM3NzIwNzQ4IzIwMDcxMTY=","inLanguage":"en-US"},{"@type":"WebPage","@id":"https:\/\/www.marketnewsdesk.com\/index.php\/amd-and-cerebras-announce-industry-leading-ultra-low-latency-and-high-throughput-ai-inference-solution\/","url":"https:\/\/www.marketnewsdesk.com\/index.php\/amd-and-cerebras-announce-industry-leading-ultra-low-latency-and-high-throughput-ai-inference-solution\/","name":"AMD and Cerebras Announce Industry-Leading Ultra-Low-Latency and High Throughput AI Inference Solution - Market Newsdesk","isPartOf":{"@id":"https:\/\/www.marketnewsdesk.com\/#website"},"primaryImageOfPage":{"@id":"https:\/\/www.marketnewsdesk.com\/index.php\/amd-and-cerebras-announce-industry-leading-ultra-low-latency-and-high-throughput-ai-inference-solution\/#primaryimage"},"image":{"@id":"https:\/\/www.marketnewsdesk.com\/index.php\/amd-and-cerebras-announce-industry-leading-ultra-low-latency-and-high-throughput-ai-inference-solution\/#primaryimage"},"thumbnailUrl":"https:\/\/www.globenewswire.com\/newsroom\/ti?nf=OTc2NzI3OSM3NzIwNzQ4IzIwMDcxMTY=","datePublished":"2026-07-23T21:41:45+00:00","author":{"@id":"https:\/\/www.marketnewsdesk.com\/#\/schema\/person\/482f27a394d4fda80ecb5499e519d979"},"breadcrumb":{"@id":"https:\/\/www.marketnewsdesk.com\/index.php\/amd-and-cerebras-announce-industry-leading-ultra-low-latency-and-high-throughput-ai-inference-solution\/#breadcrumb"},"inLanguage":"en-US","potentialAction":[{"@type":"ReadAction","target":["https:\/\/www.marketnewsdesk.com\/index.php\/amd-and-cerebras-announce-industry-leading-ultra-low-latency-and-high-throughput-ai-inference-solution\/"]}]},{"@type":"ImageObject","inLanguage":"en-US","@id":"https:\/\/www.marketnewsdesk.com\/index.php\/amd-and-cerebras-announce-industry-leading-ultra-low-latency-and-high-throughput-ai-inference-solution\/#primaryimage","url":"https:\/\/www.globenewswire.com\/newsroom\/ti?nf=OTc2NzI3OSM3NzIwNzQ4IzIwMDcxMTY=","contentUrl":"https:\/\/www.globenewswire.com\/newsroom\/ti?nf=OTc2NzI3OSM3NzIwNzQ4IzIwMDcxMTY="},{"@type":"BreadcrumbList","@id":"https:\/\/www.marketnewsdesk.com\/index.php\/amd-and-cerebras-announce-industry-leading-ultra-low-latency-and-high-throughput-ai-inference-solution\/#breadcrumb","itemListElement":[{"@type":"ListItem","position":1,"name":"Home","item":"https:\/\/www.marketnewsdesk.com\/"},{"@type":"ListItem","position":2,"name":"AMD and Cerebras Announce Industry-Leading Ultra-Low-Latency and High Throughput AI Inference Solution"}]},{"@type":"WebSite","@id":"https:\/\/www.marketnewsdesk.com\/#website","url":"https:\/\/www.marketnewsdesk.com\/","name":"Market Newsdesk","description":"Latest Business News in Real Time","potentialAction":[{"@type":"SearchAction","target":{"@type":"EntryPoint","urlTemplate":"https:\/\/www.marketnewsdesk.com\/?s={search_term_string}"},"query-input":{"@type":"PropertyValueSpecification","valueRequired":true,"valueName":"search_term_string"}}],"inLanguage":"en-US"},{"@type":"Person","@id":"https:\/\/www.marketnewsdesk.com\/#\/schema\/person\/482f27a394d4fda80ecb5499e519d979","name":"Newsdesk","image":{"@type":"ImageObject","inLanguage":"en-US","@id":"https:\/\/secure.gravatar.com\/avatar\/a0d0bd5b0f0ca12a265a459b13169dac35f33776d8501eda5e68844a366f2f46?s=96&d=mm&r=g","url":"https:\/\/secure.gravatar.com\/avatar\/a0d0bd5b0f0ca12a265a459b13169dac35f33776d8501eda5e68844a366f2f46?s=96&d=mm&r=g","contentUrl":"https:\/\/secure.gravatar.com\/avatar\/a0d0bd5b0f0ca12a265a459b13169dac35f33776d8501eda5e68844a366f2f46?s=96&d=mm&r=g","caption":"Newsdesk"},"url":"https:\/\/www.marketnewsdesk.com\/index.php\/author\/newsdesk\/"}]}},"_links":{"self":[{"href":"https:\/\/www.marketnewsdesk.com\/index.php\/wp-json\/wp\/v2\/posts\/984285","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/www.marketnewsdesk.com\/index.php\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/www.marketnewsdesk.com\/index.php\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/www.marketnewsdesk.com\/index.php\/wp-json\/wp\/v2\/users\/2"}],"replies":[{"embeddable":true,"href":"https:\/\/www.marketnewsdesk.com\/index.php\/wp-json\/wp\/v2\/comments?post=984285"}],"version-history":[{"count":0,"href":"https:\/\/www.marketnewsdesk.com\/index.php\/wp-json\/wp\/v2\/posts\/984285\/revisions"}],"wp:attachment":[{"href":"https:\/\/www.marketnewsdesk.com\/index.php\/wp-json\/wp\/v2\/media?parent=984285"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/www.marketnewsdesk.com\/index.php\/wp-json\/wp\/v2\/categories?post=984285"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/www.marketnewsdesk.com\/index.php\/wp-json\/wp\/v2\/tags?post=984285"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}