{"id":12572,"date":"2026-08-31T08:59:33","date_gmt":"2026-08-30T23:59:33","guid":{"rendered":"https:\/\/news.skhynix.com\/en\/?p=12572"},"modified":"2026-08-31T08:59:33","modified_gmt":"2026-08-30T23:59:33","slug":"ai-infrastructure-insight-ep2","status":"publish","type":"post","link":"https:\/\/news.skhynix.com\/en\/ai-infrastructure-insight-ep2\/","title":{"rendered":"[AI Infrastructure Insight] Why faster GPUs alone can\u2019t deliver AI performance"},"content":{"rendered":"<div class=\"post-intro\">\n<p>AI is no longer defined by a single model or chip. For AI to operate effectively in real-world services and industrial applications, it takes faster compute, greater memory bandwidth, higher-performance networking, more efficient storage, and stable power and cooling working together. This is why competition in AI is shifting beyond individual technologies to the design and operation of the entire infrastructure.<\/p>\n<div style=\"height: 16px; line-height: 4px;\"><\/div>\n<p>In the<strong> AI Infrastructure Insight<\/strong> series, we explore how the architecture behind AI infrastructure is evolving \u2014 from compute, memory, storage, networking, and power and cooling to the way they come together as a unified system \u2014 and what this means for the industry.<\/p>\n<div style=\"height: 16px; line-height: 4px;\"><\/div>\n<p><strong>[Series overview]<\/strong><br \/>\n\u2460 What is changing inside AI data centers?<br \/>\n<strong>\u2461 Why faster GPUs alone can\u2019t deliver AI performance<\/strong><br \/>\n\u2462 Why power and cooling have become the next challenge for AI infrastructure<br \/>\n\u2463 How AI infrastructure will be designed for the future<\/p>\n<\/div>\n<p>Does using a faster GPU<span style=\"color: #ff0000;\">*<\/span> result in faster AI operations? GPUs and AI accelerators<span style=\"color: #ff0000;\">*<\/span> are core technologies that have enabled the training of large-scale models and the processing of complex inference requests. In real-world service environments, however, AI performance is not determined by compute speed alone.<\/p>\n<div class=\"footnote\"><span style=\"color: red;\">* <\/span>Graphics processing unit(GPU): A parallel processing device originally developed for graphics computation. Its ability to process large volumes of data simultaneously has led to its widespread use in AI training and inference<br \/>\n<span style=\"color: red;\">* <\/span>AI accelerator: A semiconductor or computing device designed to rapidly perform the large-scale computations required for AI training and inference. Examples include GPUs, NPUs, and TPUs<\/div>\n<p>Even when a processor is ready to perform calculations, an expensive accelerator can be left waiting if data is not supplied in time, reducing overall system efficiency.<br \/>\nAn AI system can therefore deliver its full performance only when accelerator compute capability, data delivery speed, and data movement paths work together effectively.<\/p>\n<p>The research supports this point.<br \/>\nThe paper \u201cAI and Memory Wall,\u201d<span style=\"color: #ff0000;\">*<\/span> by researchers from UC Berkeley, ICSI, and LBNL, found that over the past 20 years, peak FLOPS<span style=\"color: #ff0000;\">*<\/span> in server hardware increased by approximately threefold every two years, while DRAM bandwidth<span style=\"color: #ff0000;\">*<\/span> increased by only 1.6 times and interconnect<span style=\"color: #ff0000;\">*<\/span> bandwidth by 1.4 times.<br \/>\nWhile compute performance has advanced rapidly, the ability to supply and move data has progressed at a comparatively slower pace. As this gap widens, bottlenecks<span style=\"color: #ff0000;\">*<\/span> in memory and data movement paths \u2014 not just compute performance \u2014 may become more pronounced in AI systems.<\/p>\n<div class=\"footnote\"><span style=\"color: red;\">* <\/span>AI and Memory Wall: A study by researchers from UC Berkeley, ICSI, and LBNL analyzing the gap between the rate of growth in compute performance and the growth of memory and interconnect bandwidth in AI systems<br \/>\n<span style=\"color: red;\">* <\/span>Peak floating point operations per second(FLOPS): The theoretical maximum floating-point compute performance of a computing system. FLOPS refers to the number of floating-point operations that can be performed per second<br \/>\n<span style=\"color: red;\">* <\/span>Bandwidth: A measure of the amount of data that can be transferred over a given period of time. It is commonly used to describe data transfer performance in memory and networks.<br \/>\n<span style=\"color: red;\">* <\/span>Interconnect: An architecture or technology that connects components such as servers, chips, memory, and accelerators, enabling data to move between them<br \/>\n<span style=\"color: red;\">* <\/span>Bottleneck: A point that limits overall system performance. In AI systems, bottlenecks can occur not only in processors but also in data delivery, memory, storage, and networks<\/div>\n<div><img loading=\"lazy\" decoding=\"async\" class=\"alignnone size-full wp-image-12668\" src=\"https:\/\/d18r0a86za96sg.cloudfront.net\/wp-content\/uploads\/2026\/08\/21180713\/AI-Infrastructure-Insight-Part-2-Why-faster-GPUs-alone-cant-deliver-AI-performance-1-info-2026.jpg\" alt=\"\" width=\"1600\" height=\"1150\" srcset=\"https:\/\/d18r0a86za96sg.cloudfront.net\/wp-content\/uploads\/2026\/08\/21180713\/AI-Infrastructure-Insight-Part-2-Why-faster-GPUs-alone-cant-deliver-AI-performance-1-info-2026.jpg 1600w, https:\/\/d18r0a86za96sg.cloudfront.net\/wp-content\/uploads\/2026\/08\/21180713\/AI-Infrastructure-Insight-Part-2-Why-faster-GPUs-alone-cant-deliver-AI-performance-1-info-2026-300x216.jpg 300w, https:\/\/d18r0a86za96sg.cloudfront.net\/wp-content\/uploads\/2026\/08\/21180713\/AI-Infrastructure-Insight-Part-2-Why-faster-GPUs-alone-cant-deliver-AI-performance-1-info-2026-1024x736.jpg 1024w, https:\/\/d18r0a86za96sg.cloudfront.net\/wp-content\/uploads\/2026\/08\/21180713\/AI-Infrastructure-Insight-Part-2-Why-faster-GPUs-alone-cant-deliver-AI-performance-1-info-2026-768x552.jpg 768w, https:\/\/d18r0a86za96sg.cloudfront.net\/wp-content\/uploads\/2026\/08\/21180713\/AI-Infrastructure-Insight-Part-2-Why-faster-GPUs-alone-cant-deliver-AI-performance-1-info-2026-1536x1104.jpg 1536w, https:\/\/d18r0a86za96sg.cloudfront.net\/wp-content\/uploads\/2026\/08\/21180713\/AI-Infrastructure-Insight-Part-2-Why-faster-GPUs-alone-cant-deliver-AI-performance-1-info-2026-1200x863.jpg 1200w\" sizes=\"auto, (max-width: 1600px) 100vw, 1600px\" \/><\/div>\n<div style=\"height: 16px; line-height: 16px;\"><\/div>\n<div style=\"text-align: center;\">\n<div style=\"display: inline-block; max-width: 748px; width: 100%; text-align: left;\">\n<div style=\"height: 2px; background: #666666; margin-bottom: 6px;\"><\/div>\n<h3 class=\"sub-title\" style=\"margin: 0; line-height: 1.4;\">Data never stays in one place<\/h3>\n<div style=\"height: 2px; background: #666666; margin-top: 6px;\"><\/div>\n<\/div>\n<\/div>\n<p>In an AI system, data does not remain stagnant in one place. Training data, model parameters<span style=\"color: #ff0000;\">*<\/span>, user requests, contextual information, search results, and compute results move between different locations depending on the stage of processing. Some data is retrieved from high-capacity storage, while other data passes through system memory before reaching memory closer to the accelerator. Once processing is complete, the results are stored again or transmitted over the network to other servers and services.<\/p>\n<div class=\"footnote\"><span style=\"color: red;\">* <\/span>Model parameter: An internal value adjusted as an AI model learns, determining how the model interprets inputs and generates outputs<\/div>\n<blockquote>\n<p style=\"text-align: left;\"><strong><span style=\"font-size: 28px;\">\u2611\ufe0e Where does data travel before AI answers?<\/span><\/strong><\/p>\n<div style=\"height: 16px; line-height: 4px;\"><\/div>\n<p>When a user enters a question into a chatbot or AI search engine, the data center processes more than just the user\u2019s request. Previous conversational context, along with search results, relevant documents, and other information the model may need, is gathered as required. This data moves through storage and memory before reaching layers closer to the accelerator. Once computation is complete, the result is stored again or returned to the user over the network.<\/p>\n<div style=\"height: 16px; line-height: 4px;\"><\/div>\n<p>As the use of AI search and chatbots grows, these inference requests are occurring more frequently and at greater scale. Ultimately, the response speed experienced by users depends not only on model performance but also on how quickly the necessary data can be gathered and moved.<\/p>\n<div style=\"height: 16px; line-height: 4px;\"><\/div>\n<p>As of July 2025, 18 billion messages were exchanged on ChatGPT each week, according to OpenAI, while Gemini surpassed 1 billion monthly users after August 2026.<\/p>\n<div style=\"height: 16px; line-height: 4px;\"><\/div>\n<p><img loading=\"lazy\" decoding=\"async\" class=\"size-full wp-image-12592\" src=\"https:\/\/d18r0a86za96sg.cloudfront.net\/wp-content\/uploads\/2026\/08\/19164228\/AI-Infrastructure-Insight-Part-2-Why-faster-GPUs-alone-cant-deliver-AI-performance-1-etc-2026.jpg\" alt=\"\" width=\"1600\" height=\"1066\" srcset=\"https:\/\/d18r0a86za96sg.cloudfront.net\/wp-content\/uploads\/2026\/08\/19164228\/AI-Infrastructure-Insight-Part-2-Why-faster-GPUs-alone-cant-deliver-AI-performance-1-etc-2026.jpg 1600w, https:\/\/d18r0a86za96sg.cloudfront.net\/wp-content\/uploads\/2026\/08\/19164228\/AI-Infrastructure-Insight-Part-2-Why-faster-GPUs-alone-cant-deliver-AI-performance-1-etc-2026-300x200.jpg 300w, https:\/\/d18r0a86za96sg.cloudfront.net\/wp-content\/uploads\/2026\/08\/19164228\/AI-Infrastructure-Insight-Part-2-Why-faster-GPUs-alone-cant-deliver-AI-performance-1-etc-2026-1024x682.jpg 1024w, https:\/\/d18r0a86za96sg.cloudfront.net\/wp-content\/uploads\/2026\/08\/19164228\/AI-Infrastructure-Insight-Part-2-Why-faster-GPUs-alone-cant-deliver-AI-performance-1-etc-2026-768x512.jpg 768w, https:\/\/d18r0a86za96sg.cloudfront.net\/wp-content\/uploads\/2026\/08\/19164228\/AI-Infrastructure-Insight-Part-2-Why-faster-GPUs-alone-cant-deliver-AI-performance-1-etc-2026-1536x1023.jpg 1536w, https:\/\/d18r0a86za96sg.cloudfront.net\/wp-content\/uploads\/2026\/08\/19164228\/AI-Infrastructure-Insight-Part-2-Why-faster-GPUs-alone-cant-deliver-AI-performance-1-etc-2026-1200x800.jpg 1200w\" sizes=\"auto, (max-width: 1600px) 100vw, 1600px\" \/><\/p>\n<p class=\"caption\">\u25b2 Stargate Data Center Source : OpenAI, \u300cStargate Advances with Partnership with Oracle\u300d(2025)<\/p>\n<\/blockquote>\n<div style=\"height: 16px; line-height: 4px;\"><\/div>\n<p style=\"text-align: left; word-break: keep-all; overflow-wrap: break-word;\">Within this flow, each layer plays a different role. Storage retains large volumes of information, while system memory provides a workspace for the entire server. Memory located close to the accelerator rapidly supplies the information needed by the processor, while networks connect multiple servers and accelerators into a single system.<\/p>\n<p>AI performance, therefore, is not guaranteed simply by having the necessary data available. Even when using the same GPU, systems can deliver different real-world performance depending on how efficiently data is supplied and moved.<\/p>\n<div style=\"height: 16px; line-height: 16px;\"><\/div>\n<div style=\"text-align: center;\">\n<div style=\"display: inline-block; max-width: 748px; width: 100%; text-align: left;\">\n<div style=\"height: 2px; background: #666666; margin-bottom: 6px;\"><\/div>\n<h3 class=\"sub-title\" style=\"margin: 0; line-height: 1.4;\">Bottlenecks beyond compute<\/h3>\n<div style=\"height: 2px; background: #666666; margin-top: 6px;\"><\/div>\n<\/div>\n<\/div>\n<p>Bottlenecks in AI systems do not occur in a single location. No matter how fast a processor is, insufficient memory bandwidth can prevent the system from supplying enough data when it is needed. In this case, an accelerator may be ready to compute but still spend time waiting for data, reducing actual utilization rates.<\/p>\n<p>Storage can also become a bottleneck<span style=\"color: #ff0000;\">* <\/span>point. When an AI service needs to retrieve large volumes of documents, images, logs, user histories, or search results, response speeds slow down if the required data cannot be read quickly enough.<br \/>\nAs the number of requests for inference services grows, such delays can directly degrade service quality.<\/p>\n<p>Similar challenges arise during large-scale AI training. Training an AI model requires checkpointing<span style=\"color: #ff0000;\">*<\/span>, which periodically saves the model\u2019s intermediate training state. If storage fails to save or retrieve these large volumes of data quickly enough, latency can occur across the entire training cluster. This is why MLCommons<span style=\"color: #ff0000;\">*<\/span> separately measures storage performance for AI and machine learning training workloads through its MLPerf Storage<span style=\"color: #ff0000;\">*<\/span> benchmark. Data management is a core element of AI training and inference, making storage performance an increasingly important part of AI infrastructure.<\/p>\n<div class=\"footnote\"><span style=\"color: red;\">* <\/span>Checkpoint: A process or set of saved data that periodically records the current state of AI model training, allowing training to resume from the saved point following an interruption or failure<br \/>\n<span style=\"color: red;\">* <\/span>MLCommons: A global AI engineering consortium that develops industry-standard benchmarks for measuring the performance and efficiency of AI systems<br \/>\n<span style=\"color: red;\">* <\/span>MLPerf Storage: A suite of MLCommons benchmarks that evaluates the performance of storage systems supporting AI and machine learning workloads. It measures how data reads and delivery, as well as checkpoint saving and recovery during AI training, affect system efficiency<\/div>\n<p>Networks are another potential source of bottlenecks. In large-scale AI training or high-performance inference, multiple servers and accelerators exchange data simultaneously to divide and process a single task. If data transmission between servers are delayed or network bandwidth is insufficient, slowdowns in one part of the system can reduce overall processing speed. It is not enough for an individual device to be fast. Multiple devices need to exchange data efficiently and operate as a unified system.<\/p>\n<p>Power and cooling can also constrain performance. High-performance equipment requires a stable power supply and effective thermal management to sustain peak performance. Part 3 of this series will explore these issues in greater detail, but the key point is clear: Bottlenecks and performance limitations can arise throughout the system \u2014 from data movement paths involving memory, storage, and networks to power and cooling infrastructure.<\/p>\n<p><img loading=\"lazy\" decoding=\"async\" class=\"alignnone size-full wp-image-12772\" src=\"https:\/\/d18r0a86za96sg.cloudfront.net\/wp-content\/uploads\/2026\/08\/26162421\/AI-Infrastructure-Insight-Why-faster-GPUs-alone-cant-deliver-AI-performance-4-infographic-2026.jpg\" alt=\"\" width=\"1600\" height=\"900\" srcset=\"https:\/\/d18r0a86za96sg.cloudfront.net\/wp-content\/uploads\/2026\/08\/26162421\/AI-Infrastructure-Insight-Why-faster-GPUs-alone-cant-deliver-AI-performance-4-infographic-2026.jpg 1600w, https:\/\/d18r0a86za96sg.cloudfront.net\/wp-content\/uploads\/2026\/08\/26162421\/AI-Infrastructure-Insight-Why-faster-GPUs-alone-cant-deliver-AI-performance-4-infographic-2026-300x169.jpg 300w, https:\/\/d18r0a86za96sg.cloudfront.net\/wp-content\/uploads\/2026\/08\/26162421\/AI-Infrastructure-Insight-Why-faster-GPUs-alone-cant-deliver-AI-performance-4-infographic-2026-1024x576.jpg 1024w, https:\/\/d18r0a86za96sg.cloudfront.net\/wp-content\/uploads\/2026\/08\/26162421\/AI-Infrastructure-Insight-Why-faster-GPUs-alone-cant-deliver-AI-performance-4-infographic-2026-768x432.jpg 768w, https:\/\/d18r0a86za96sg.cloudfront.net\/wp-content\/uploads\/2026\/08\/26162421\/AI-Infrastructure-Insight-Why-faster-GPUs-alone-cant-deliver-AI-performance-4-infographic-2026-1536x864.jpg 1536w, https:\/\/d18r0a86za96sg.cloudfront.net\/wp-content\/uploads\/2026\/08\/26162421\/AI-Infrastructure-Insight-Why-faster-GPUs-alone-cant-deliver-AI-performance-4-infographic-2026-1200x675.jpg 1200w\" sizes=\"auto, (max-width: 1600px) 100vw, 1600px\" \/><\/p>\n<div style=\"height: 16px; line-height: 16px;\"><\/div>\n<div style=\"text-align: center;\">\n<div style=\"display: inline-block; max-width: 748px; width: 100%; text-align: left;\">\n<div style=\"height: 2px; background: #666666; margin-bottom: 6px;\"><\/div>\n<h3 class=\"sub-title\" style=\"margin: 0; line-height: 1.4;\">Data architecture improves accelerator utilization<\/h3>\n<div style=\"height: 2px; background: #666666; margin-top: 6px;\"><\/div>\n<\/div>\n<\/div>\n<p>The key to AI infrastructure lies not only in securing more accelerators but also in designing systems that keep those accelerators continuously engaged in computation. When data is supplied too slowly, accelerator utilization<span style=\"color: #ff0000;\">*<\/span> falls, reducing performance relative to the investment made. This is why even systems equipped with high-performance hardware may fail to perform at their full potential.<\/p>\n<div class=\"footnote\"><span style=\"color: red;\">* <\/span>Accelerator utilization: The proportion of time a GPU or AI accelerator is actively used for computation. Delays in data delivery can leave accelerators waiting and reduce utilization<\/div>\n<p>To prevent this, data paths need to be designed so that data can be supplied where and when it is needed. Frequently used data should be placed close to processors, while large volumes of data should be retrieved efficiently through storage and networks. This is due to the fact that it is impossible to place all information in the fastest location. Ultimately, the location and movement paths of data must be meticulously designed with speed, capacity, cost, and power efficiency in mind.<\/p>\n<p>Software coordinates this architecture for real-world operations. Accelerator utilization and overall performance vary depending on which tasks are processed first, when data is retrieved, and how they are both distributed across multiple devices. Even with fast hardware, inefficient scheduling and placement can create new bottlenecks.<\/p>\n<p>Designing to minimize data movement can also improve power efficiency. Intel explains that while the process of moving data within a system consumes energy, it does not directly contribute to computation itself. This is why data should travel shorter distances and be placed closer to where it is needed.<\/p>\n<p>Within this architecture, memory serves as a critical layer that supplies data where and when it is needed, helping accelerators remain continuously engaged in computation. How efficiently compute and memory are connected can therefore determine real-world system performance.<\/p>\n<div><img loading=\"lazy\" decoding=\"async\" class=\"alignnone size-full wp-image-12596\" src=\"https:\/\/d18r0a86za96sg.cloudfront.net\/wp-content\/uploads\/2026\/08\/19164611\/AI-Infrastructure-Insight-Part-2-Why-faster-GPUs-alone-cant-deliver-AI-performance-5-info-2026.jpg\" alt=\"\" width=\"1600\" height=\"850\" srcset=\"https:\/\/d18r0a86za96sg.cloudfront.net\/wp-content\/uploads\/2026\/08\/19164611\/AI-Infrastructure-Insight-Part-2-Why-faster-GPUs-alone-cant-deliver-AI-performance-5-info-2026.jpg 1600w, https:\/\/d18r0a86za96sg.cloudfront.net\/wp-content\/uploads\/2026\/08\/19164611\/AI-Infrastructure-Insight-Part-2-Why-faster-GPUs-alone-cant-deliver-AI-performance-5-info-2026-300x159.jpg 300w, https:\/\/d18r0a86za96sg.cloudfront.net\/wp-content\/uploads\/2026\/08\/19164611\/AI-Infrastructure-Insight-Part-2-Why-faster-GPUs-alone-cant-deliver-AI-performance-5-info-2026-1024x544.jpg 1024w, https:\/\/d18r0a86za96sg.cloudfront.net\/wp-content\/uploads\/2026\/08\/19164611\/AI-Infrastructure-Insight-Part-2-Why-faster-GPUs-alone-cant-deliver-AI-performance-5-info-2026-768x408.jpg 768w, https:\/\/d18r0a86za96sg.cloudfront.net\/wp-content\/uploads\/2026\/08\/19164611\/AI-Infrastructure-Insight-Part-2-Why-faster-GPUs-alone-cant-deliver-AI-performance-5-info-2026-1536x816.jpg 1536w, https:\/\/d18r0a86za96sg.cloudfront.net\/wp-content\/uploads\/2026\/08\/19164611\/AI-Infrastructure-Insight-Part-2-Why-faster-GPUs-alone-cant-deliver-AI-performance-5-info-2026-1200x638.jpg 1200w\" sizes=\"auto, (max-width: 1600px) 100vw, 1600px\" \/><\/div>\n<div style=\"text-align: center;\">\n<div style=\"display: inline-block; max-width: 748px; width: 100%; text-align: left; word-break: keep-all; overflow-wrap: break-word;\">\n<div style=\"height: 2px; background: #666666; margin-bottom: 6px;\"><\/div>\n<h3 class=\"sub-title\" style=\"margin: 0; line-height: 1.4;\">AI infrastructure competitiveness depends on data flow<\/h3>\n<div style=\"height: 2px; background: #666666; margin-top: 6px;\"><\/div>\n<\/div>\n<\/div>\n<p>This approach of designing systems around data paths is bringing about changes in the competitive landscape of the AI infrastructure industry. While fast GPUs, high memory bandwidth, high-capacity storage, and high-speed networks are all important individually, they are insufficient if they operate in isolation. AI systems can deliver real-world performance only when data moves smoothly from storage to the processor, and results are delivered where they are needed.<\/p>\n<p>Consequently, collaboration methods across the entire industry ecosystem are also shifting. There are limitations to maximizing overall system performance if semiconductor, server, networking, cloud, and storage companies each optimize only their own components independently. As AI workloads grow larger and more complex, collaboration based on the entire data flow becomes crucial. This is due to the fact that required infrastructure architecture varies depending on which models customers use, what data they process, and what level of response speed is required.<\/p>\n<p>The role of memory companies is expanding as well. Beyond merely supplying products, they increasingly need the capabilities to jointly design data flows within customer systems and determine what memory architectures are required from a system-level perspective encompassing processors, storage, and networks.<\/p>\n<p>Compute remains essential. Faster GPUs and AI accelerators will continue to be core components of AI infrastructure. However, translating their capabilities into real-world performance requires optimizing the flow of data across the system as well.<\/p>\n<p>Looking ahead, understanding AI infrastructure requires examining not only the speed of processors but also the ways in which data moves through the system. Memory, storage, networks, and power and cooling should not be considered as independent components but as organically interconnected parts of a single system that keep data flowing without interruption.<\/p>\n<p>The need for faster GPUs remains as essential as ever. But the question has now evolved one step further: How quickly and efficiently does the data reach the GPU for processing?<\/p>\n<p>The next bottleneck determining AI performance begins with that question.<\/p>\n<div class=\"post-intro\">\n<p style=\"font-weight: 400;\"><strong>&lt;References<\/strong><strong>&gt;<\/strong><\/p>\n<ul style=\"font-weight: 400;\">\n<li>Amir Gholami et al., \u201c<a href=\"https:\/\/arxiv.org\/abs\/2403.14123\" target=\"_blank\" rel=\"noopener\">AI and Memory Wall<\/a>,\u201d IEEE Micro, Vol. 44, No. 3, 2024.<\/li>\n<li>OpenAI, \u201c<a href=\"https:\/\/cdn.openai.com\/pdf\/a253471f-8260-40c6-a2cc-aa93fe9f142e\/economic-research-chatgpt-usage-paper.pdf\" target=\"_blank\" rel=\"noopener\">How People Use ChatGPT<\/a>,\u201d 2025.<\/li>\n<li>Google, \u201c<a href=\"https:\/\/blog.google\/innovation-and-ai\/products\/gemini-app\/one-billion-monthly-users\/\" target=\"_blank\" rel=\"noopener\">Google\u2019s Gemini app hits 1 billion monthly active users<\/a>,\u201d<br \/>\nGoogle Blog, Aug. 11, 2026.<\/li>\n<\/ul>\n<\/div>\n","protected":false},"excerpt":{"rendered":"<p>AI is no longer defined by a single model or chip. For AI to operate effectively in real-world services and industrial applications, it takes faster compute, greater memory bandwidth, higher-performance networking, more efficient storage, and stable power and cooling working<\/p>\n","protected":false},"author":23,"featured_media":12601,"comment_status":"open","ping_status":"open","sticky":false,"template":"","format":"standard","meta":{"_acf_changed":false,"_migrated_source_id":0,"footnotes":"","_members_access_role":[],"_members_access_error":""},"categories":[5],"tags":[173,610,14,1597],"class_list":["post-12572","post","type-post","status-publish","format-standard","has-post-thumbnail","hentry","category-tech-and-ai","tag-ai-data-center","tag-ai-infra","tag-ai-memory","tag-gpu"],"acf":[],"_links":{"self":[{"href":"https:\/\/news.skhynix.com\/en\/wp-json\/wp\/v2\/posts\/12572","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/news.skhynix.com\/en\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/news.skhynix.com\/en\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/news.skhynix.com\/en\/wp-json\/wp\/v2\/users\/23"}],"replies":[{"embeddable":true,"href":"https:\/\/news.skhynix.com\/en\/wp-json\/wp\/v2\/comments?post=12572"}],"version-history":[{"count":33,"href":"https:\/\/news.skhynix.com\/en\/wp-json\/wp\/v2\/posts\/12572\/revisions"}],"predecessor-version":[{"id":12864,"href":"https:\/\/news.skhynix.com\/en\/wp-json\/wp\/v2\/posts\/12572\/revisions\/12864"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/news.skhynix.com\/en\/wp-json\/wp\/v2\/media\/12601"}],"wp:attachment":[{"href":"https:\/\/news.skhynix.com\/en\/wp-json\/wp\/v2\/media?parent=12572"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/news.skhynix.com\/en\/wp-json\/wp\/v2\/categories?post=12572"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/news.skhynix.com\/en\/wp-json\/wp\/v2\/tags?post=12572"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}