{"id":92667,"date":"2024-07-11T14:34:44","date_gmt":"2024-07-11T05:34:44","guid":{"rendered":"https:\/\/kr-dev.rebellions.ai\/?p=92667"},"modified":"2025-08-21T17:00:39","modified_gmt":"2025-08-21T08:00:39","slug":"atom-architecture-finding-the-sweet-spot-for-genai","status":"publish","type":"post","link":"https:\/\/dev-rbln-kr.locomotion.co.kr\/?p=92667","title":{"rendered":"ATOM\u2122 Architecture: Finding the Sweet Spot for GenAI"},"content":{"rendered":"\n<h2 class=\"wp-block-heading has-large-font-size\"><strong>Introduction<\/strong><\/h2>\n\n\n\n<p class=\".entry-content h1.wp-block-heading {    font-size: 2.5rem;   font-weight: 700;   line-height: 1.2; } wp-block-paragraph\">Generative AI (GenAI) is transforming industries, necessitating the development of specialized hardware to manage its computational demands. AI accelerators or AI chips crafted specifically for AI tasks are critical in this advancement, but designing effective AI chips presents significant challenges.<\/p>\n\n\n\n<p class=\"wp-block-paragraph\">Certain applications hinge on the absolute criticality of every millisecond\u2014these are known as latency-critical scenarios. For example, latency is crucial in high-frequency trading environments, where algorithms execute transactions in microseconds to leverage fleeting market opportunities. Similarly, cloud services are committed to maintaining 99-percentile latency guarantees to ensure stable and predictable performance, even during peak loads.<\/p>\n\n\n\n<p class=\"wp-block-paragraph\">Another critical measure is the throughput, which is directly related to the performance of the AI application. One traditional method to manage high computational throughput is batching, where large amounts of tasks are grouped and executed consecutively. However, this technique typically sacrifices latency for throughput, which is a critical trade-off.<\/p>\n\n\n\n<p class=\"wp-block-paragraph\">Equally critical, yet often overlooked, is flexibility: the balance between memory and compute operations at the system level. Text-based Large Language Model (LLM) operations are memory-intensive, relying heavily on frequent RAM access to manage the vast array of parameters that underpin their linguistic functions. In contrast, Text-to-Video applications demand robust compute capabilities to manage intensive graphical processing and real-time data handling effectively. To support such a diverse array of applications, an AI chip must adeptly navigate the demands of both memory and computational intensity without compromise.<\/p>\n\n\n\n<p class=\"wp-block-paragraph\">In sum, an effective AI chip must find the sweet spot between latency, throughput and flexibility.<\/p>\n\n\n\n<h2 class=\"wp-block-heading has-large-font-size\"><strong>Flexibility and High Utilization<\/strong><\/h2>\n\n\n\n<p class=\"wp-block-paragraph\">In designing our chip, we prioritized flexibility and high compute utilization to address these key challenges. By adopting a CGRA (Coarse-Grained Reconfigurable Array) architecture, the processing element tiles can be programmed and reprogrammed to carry out a variety of functions, maximizing its flexibility. We also kept the utilization rate high, so that tasks are processed continuously with minimal idle resources, directly enhancing efficiency and reducing latency.<br><\/p>\n\n\n\n<p class=\"wp-block-paragraph\">Built on a flexible architecture, ATOM\u2122 leverages synchronization mechanisms to activate resources precisely when needed, to support its powerful parallelism. The time and effort to reach operational readiness is minimized, leading to reduced latency. Moreover, its robust multi-layered memory hierarchy provides significant bandwidth, reducing data dependency, while the sophisticated synchronization lessens control dependency. These features together optimize resource utilization, significantly boosting overall performance and efficiency in a seamless integration.<\/p>\n\n\n\n<h2 class=\"wp-block-heading has-large-font-size\"><strong>ATOM\u2122: System-on-Chip for AI Inference<\/strong><\/h2>\n\n\n\n<p class=\"wp-block-paragraph\">Rebellions\u2019 ATOM\u2122 is an AI accelerator engineered specifically for AI inference tasks with formidable capacity, manufactured on Samsung\u2019s advanced 5nm process. It delivers 32 Tera Floating Point Operations per Second (TFLOPS) for FP16 and 128 Trillion Operations Per Second (TOPS) for INT8, enhanced by eight Neural Engines and 64 MB of on-chip SRAM. With an intricate memory architecture engineered with unparalleled technical mastery, ATOM\u2122 is designed for high performance and peak efficiency.<\/p>\n\n\n\n<figure class=\"wp-block-image size-large\"><img decoding=\"async\" src=\"https:\/\/rebellions.ai\/wp-content\/uploads\/2025\/08\/image-1024x632.png\" alt=\"\" class=\"wp-image-92386\"\/><\/figure>\n\n\n\n<h2 class=\"wp-block-heading has-large-font-size\"><strong>RBLN-CA12<\/strong><\/h2>\n\n\n\n<p class=\"wp-block-paragraph\">ATOM\u2122 comes in RBLN-CA12, a single slot, FHFL (Full Height, Full Length) PCIe Gen5 card with a TDP (Thermal Design Power) of 60-130 W. RBLN-CA12 features 16 GB of GDDR6 memory with a bandwidth of 256 GB\/s and host and card-to-card interfaces via PCIe Gen5 x16. It also has the Multi-Instance capability, partitioning ATOM\u2122 into 16 independent hardware-isolated instances, allowing a dynamic allocation of resources and powerful multitasking.<\/p>\n\n\n<div class=\"wp-block-image\">\n<figure class=\"aligncenter size-full is-resized\"><img decoding=\"async\" width=\"902\" height=\"848\" src=\"https:\/\/kr-dev.rebellions.ai\/wp-content\/uploads\/2025\/08\/image.png\" alt=\"\" class=\"wp-image-92837\" style=\"width:625px;height:auto\" srcset=\"https:\/\/dev-rbln-kr.locomotion.co.kr\/wp-content\/uploads\/2025\/08\/image.png 902w, https:\/\/dev-rbln-kr.locomotion.co.kr\/wp-content\/uploads\/2025\/08\/image-300x282.png 300w, https:\/\/dev-rbln-kr.locomotion.co.kr\/wp-content\/uploads\/2025\/08\/image-768x722.png 768w, https:\/\/dev-rbln-kr.locomotion.co.kr\/wp-content\/uploads\/2025\/08\/image-640x602.png 640w\" sizes=\"(max-width: 902px) 100vw, 902px\" \/><figcaption class=\"wp-element-caption\">[Table 1. RBLN-CA12 Specifications]<\/figcaption><\/figure>\n<\/div>\n\n\n<h2 class=\"wp-block-heading has-large-font-size\"><strong>ATOM\u2122 SoC<\/strong><\/h2>\n\n\n<div class=\"wp-block-image\">\n<figure class=\"aligncenter size-large\"><img decoding=\"async\" src=\"https:\/\/rebellions.ai\/wp-content\/uploads\/2025\/08\/image-2-1024x493.png\" alt=\"\" class=\"wp-image-92390\"\/><figcaption class=\"wp-element-caption\">[Figure 1. ATOM\u2122 Multi-layered SoC Architecture]<\/figcaption><\/figure>\n<\/div>\n\n\n<p class=\"wp-block-paragraph\">ATOM\u2122 is a multi-core System-on-Chip, consolidating all essential components onto a unified substrate. As shown in Figure 1, this architecture integrates Neural Engines, the Command Processor, shared on-chip memory (SRAM), and GDDR6 memory within one compact surface. The high degree of integration not only diminishes physical footprint but also optimizes power efficiency.<\/p>\n\n\n\n<p class=\"wp-block-paragraph\">While this configuration in itself ensures streamlined inter-component communication and significantly reduced latency, it is further supported by a sophisticated Network-on-Chip (NoC) that provides high bandwidth. The architecture is also designed to support synchronizations between multiple layers.<\/p>\n\n\n\n<h2 class=\"wp-block-heading has-large-font-size\"><strong>Neural Engine<\/strong><\/h2>\n\n\n\n<p class=\"wp-block-paragraph\">ATOM\u2122\u2019s Neural Engine is where the actual computations take place. The compute units within the Neural Engines incorporate a blend of heterogeneous Single Instruction, Multiple Data (SIMD) and Multiple Instruction, Multiple Data (MIMD) compute elements, harnessing their respective capabilities for parallel performance and dependency control across diverse computational scenarios at the instruction levels.<\/p>\n\n\n<div class=\"wp-block-image\">\n<figure class=\"aligncenter size-full is-resized\"><img decoding=\"async\" width=\"614\" height=\"564\" src=\"https:\/\/kr-dev.rebellions.ai\/wp-content\/uploads\/2025\/08\/image-1.png\" alt=\"\" class=\"wp-image-92841\" style=\"width:485px;height:auto\" srcset=\"https:\/\/dev-rbln-kr.locomotion.co.kr\/wp-content\/uploads\/2025\/08\/image-1.png 614w, https:\/\/dev-rbln-kr.locomotion.co.kr\/wp-content\/uploads\/2025\/08\/image-1-300x276.png 300w\" sizes=\"(max-width: 614px) 100vw, 614px\" \/><figcaption class=\"wp-element-caption\">[Figure 2. ATOM\u2122 Neural Engine]<\/figcaption><\/figure>\n<\/div>\n\n\n<p class=\"wp-block-paragraph\">The compute units are fortified with a 4 MB Scratch Pad memory, facilitating access to interim data in the SRAM at a speed up to 8 TB\/s. This design mitigates bandwidth limitations and reduces latency by minimizing reliance on off-chip memory sources, thereby optimizing performance and energy efficiency.<\/p>\n\n\n\n<p class=\"wp-block-paragraph\">Task Managers reside in each Neural Engine to accelerate synchronization on the local hardware level, effectively working alongside the Command Processor in bringing about maximum compute utilization.<\/p>\n\n\n\n<p class=\"wp-block-paragraph\">The compute units, Scratch Pad memory, and Task Managers within the Neural Engines collectively contribute to ATOM\u2122\u2019s high utilization and low latency performance.<\/p>\n\n\n\n<h2 class=\"wp-block-heading has-large-font-size\"><strong>Hierarchical Memory Subsystem<\/strong><\/h2>\n\n\n<div class=\"wp-block-image\">\n<figure class=\"aligncenter size-large\"><img decoding=\"async\" src=\"https:\/\/rebellions.ai\/wp-content\/uploads\/2025\/08\/image-4-1024x309.png\" alt=\"\" class=\"wp-image-92395\"\/><figcaption class=\"wp-element-caption\">[Figure 3. ATOM\u2122 Hierarchical Memory Subsystem]<\/figcaption><\/figure>\n<\/div>\n\n\n<p class=\"wp-block-paragraph\">ATOM\u2122\u2019s multi-layered memory architecture is designed to ensure peak performance efficiency, delivering ample bandwidth for the Neural Engines while preserving minimal latency.<\/p>\n\n\n\n<p class=\"wp-block-paragraph\">At the foundation, a dedicated 4 MB Scratch Pad (L0) within each Neural Engine facilitates immediate local data access. The L1 Neural Cache, located close to the Engines, provides faster access to data. The L2 Shared Memory, a 32 MB SRAM, employs multiple levels of interleaving to support parallelism, optimize bandwidth, and minimize latency. Finally, ATOM\u2122 integrates 16 GB of GDDR6, ensuring high throughput with lower power consumption.<\/p>\n\n\n\n<h2 class=\"wp-block-heading has-large-font-size\"><strong>Multi-level Synchronization and Parallelism<\/strong><\/h2>\n\n\n<div class=\"wp-block-image\">\n<figure class=\"aligncenter size-full\"><img decoding=\"async\" src=\"https:\/\/rebellions.ai\/wp-content\/uploads\/2025\/08\/image-12.png\" alt=\"\" class=\"wp-image-92439\"\/><figcaption class=\"wp-element-caption\">[Figure 4. ATOM\u2122 Synchronization Scheme]<\/figcaption><\/figure>\n<\/div>\n\n\n<p class=\"wp-block-paragraph\">ATOM\u2122\u2019s advanced synchronization mechanisms effectively support parallelism, allowing the chip to scale its performance. Synchronization takes place both at the instruction and task levels, enabled by Command Processors and Task Managers and dedicated local buses that ensure smooth flow through reliable bandwidth.<\/p>\n\n\n\n<p class=\"wp-block-paragraph\">Neural Engines communicate through Task Managers across the L1 Sync Bus for instruction-level inter Engine communication. Neural Engines Clusters, which consist of four Neural Engines, are connected to the Task Direct Memory Access (TDMA) through the L 2 Sync Bus. TDMA and the Host Direct Memory Access (HDMA) are linked to the Command Processor via L3 and L4 Sync Bus, respectively. This arrangement allows the system to globally check for dependencies, synchronizing different Engines and thereby allowing for the dense compute operations.<\/p>\n\n\n<div class=\"wp-block-image\">\n<figure class=\"aligncenter size-full\"><img decoding=\"async\" src=\"https:\/\/rebellions.ai\/wp-content\/uploads\/2025\/08\/image-6.png\" alt=\"\" class=\"wp-image-92399\"\/><figcaption class=\"wp-element-caption\">[Figure 5-1. Sequential Execution without Task Managers]<\/figcaption><\/figure>\n<\/div>\n\n\n<p class=\"wp-block-paragraph\">In a basic configuration where the Command Processor solely governs command execution, tasks are processed sequentially, leading to high latency. In Figure 5-1, each task can only be executed once its dependency is resolved by the Command Processor, resulting in a slow and inefficient process. There is communication overhead, critically impacting latency and necessitating further optimization.<\/p>\n\n\n<div class=\"wp-block-image\">\n<figure class=\"aligncenter size-full\"><img decoding=\"async\" src=\"https:\/\/rebellions.ai\/wp-content\/uploads\/2025\/08\/image-7.png\" alt=\"\" class=\"wp-image-92401\"\/><figcaption class=\"wp-element-caption\">[Figure 5-2. Parallel Execution with Task Managers]<\/figcaption><\/figure>\n<\/div>\n\n\n<p class=\"wp-block-paragraph\">To address this, we introduced Task Managers that autonomously resolve local dependencies directly at the hardware level. In Figure 5-2, the DMA\/COMP tasks, each belonging to a Neural Engine, can be executed at the same time, in parallel, without having to wait for the Command Processor to resolve their dependencies. The Task Manager resolves dependencies across all Neural Engines, allowing the tasks to be executed simultaneously. This process is made possible by the dedicated L1\/L2 data paths designed explicitly for this purpose, as shown in Figure 4. Consequently, tasks across all Neural Engines are coordinated efficiently, enabling smooth parallel execution and achieving minimal latency.<\/p>\n\n\n\n<h2 class=\"wp-block-heading has-large-font-size\"><strong>Benchmark Results<\/strong><\/h2>\n\n\n\n<p class=\"wp-block-paragraph\">To demonstrate ATOM\u2122\u2019s inference performance for GenAI use cases, we conducted performance measurements on the T5-3B and SDXL-Turbo models, which are renowned for their applications in Natural Language Processing and Text-to-Image Generation, respectively.<\/p>\n\n\n\n<p class=\"wp-block-paragraph\">These tests were carried out against NVIDIA\u2019s A100, which serves as an appropriate competitor for evaluating ATOM\u2122\u2019s capabilities in the market. By focusing on prominent GenAI use cases, we provide a clear and direct comparison of how ATOM\u2122 stands against existing solutions in handling cutting-edge AI tasks.<\/p>\n\n\n\n<h3 class=\"wp-block-heading\" style=\"font-size:26px\"><strong>Language Model Benchmark: T5-3B<\/strong><\/h3>\n\n\n\n<p class=\"wp-block-paragraph\">Introduced by Google, the T5, or Text-to-Text Transfer Transformer, is a groundbreaking Large Language Model that leverages the architecture of the widely-utilized Transformer. T5 models are offered in configurations ranging from 60 million to 11 billion parameters.<\/p>\n\n\n\n<p class=\"wp-block-paragraph\">For our demonstration, we deployed the 3 billion parameter model, which is versatile enough for tasks such as language translation, text summarization, answering questions, and text generation. The test was conducted on batch size 1.<\/p>\n\n\n\n<p class=\"wp-block-paragraph\">The resulting metrics\u2014performance, quantified by tokens generated per second; power consumption, measured in watts; and power efficiency, calculated as performance per watt\u2014reveal that ATOM\u2122 achieves up to 44% greater power efficiency compared to the A100. This not only underscores ATOM\u2122\u2019s robust capabilities but also its superior efficiency in harnessing computational power for complex language processing tasks.<\/p>\n\n\n\n<figure class=\"wp-block-image size-large is-resized\"><img decoding=\"async\" src=\"https:\/\/rebellions.ai\/wp-content\/uploads\/2025\/08\/image-8-1024x194.png\" alt=\"\" class=\"wp-image-92403\" style=\"width:855px;height:auto\"\/><figcaption class=\"wp-element-caption\">Both tests are conducted on FP16 precision. ATOM\u2122\u2019s result is based on projected data. A100\u2019s result is based on the Hugging Face transformers library<\/figcaption><\/figure>\n\n\n\n<h3 class=\"wp-block-heading\" style=\"font-size:26px\"><strong>Text-to-Image Model Benchmark: SDXL-Turbo<\/strong><\/h3>\n\n\n\n<p class=\"wp-block-paragraph\">SDXL-Turbo, a Text-to-Image model developed by Stability AI, excels in generating high-resolution images and offers significantly faster inference speeds compared to standard stable diffusion models. This advancement has catalyzed the adoption of diffusion-based image generation for practical<br>applications.<\/p>\n\n\n\n<p class=\"wp-block-paragraph\">Testing results indicate that ATOM\u2122 consumes considerably less power than the A100 while still delivering high-quality outputs. This efficiency demonstrates ATOM\u2122\u2019s capability to achieve superior results with fewer resources, markedly reducing operational costs and enhancing sustainability in service deployments, as energy consumption is a critical factor that directly influences the Total Cost<br>of Ownership (TCO) for service providers.<\/p>\n\n\n\n<figure class=\"wp-block-image size-large\"><img decoding=\"async\" src=\"https:\/\/rebellions.ai\/wp-content\/uploads\/2025\/08\/image-9-1024x190.png\" alt=\"\" class=\"wp-image-92405\"\/><figcaption class=\"wp-element-caption\">Image size 512&#215;512, Diffusion step: 1 <br>ATOM\u2122\u2019s result is based on projected data. A100\u2019s result is based on the Hugging Face diffusers library.<\/figcaption><\/figure>\n\n\n\n<h2 class=\"wp-block-heading has-large-font-size\"><strong>Conclusion<\/strong><\/h2>\n\n\n\n<p class=\"wp-block-paragraph\">As businesses are increasingly dependent on AI services, finding the right AI chip that can scale sustainably presents a formidable challenge. The ideal AI chip must strike a precise balance between flexibility, power efficiency, and high performance, without sacrificing latency. ATOM\u2122 has been designed from the ground up to meet these demands, utilizing a CGRA architecture to ensure adaptability and high compute utilization. Its innovative Neural Engines, advanced multi-layered memory architecture, and robust synchronization capabilities optimize both latency and power efficiency. Furthermore, benchmark tests with T5-3B and SDXL-Turbo models demonstrate that ATOM\u2122 delivers up to 44% greater power efficiency than NVIDIA\u2019s A100. These results highlight ATOM\u2122\u2019s capacity to drastically reduce Total Cost of Ownership (TCO) and enhance profitability for AI services, establishing it as the optimal AI chip for a sustainable AI service.<\/p>\n","protected":false},"excerpt":{"rendered":"<p>Introduction Generative AI (GenAI) is transforming industries, necessitating the development of specialized hardware to manage its computational demands. AI accelerators&#8230;<\/p>\n","protected":false},"author":8,"featured_media":92719,"comment_status":"closed","ping_status":"closed","sticky":false,"template":"","format":"standard","meta":{"_acf_changed":false,"footnotes":""},"categories":[145],"tags":[],"class_list":["post-92667","post","type-post","status-publish","format-standard","has-post-thumbnail","hentry","category-white-papers"],"acf":[],"yoast_head":"<!-- This site is optimized with the Yoast SEO plugin v25.9 - https:\/\/yoast.com\/wordpress\/plugins\/seo\/ -->\n<title>ATOM\u2122 Architecture: Finding the Sweet Spot for GenAI - Rebellions<\/title>\n<meta name=\"robots\" content=\"index, follow, max-snippet:-1, max-image-preview:large, max-video-preview:-1\" \/>\n<link rel=\"canonical\" href=\"https:\/\/dev-rbln-kr.locomotion.co.kr\/?p=92667\" \/>\n<meta property=\"og:locale\" content=\"ko_KR\" \/>\n<meta property=\"og:type\" content=\"article\" \/>\n<meta property=\"og:title\" content=\"ATOM\u2122 Architecture: Finding the Sweet Spot for GenAI - Rebellions\" \/>\n<meta property=\"og:description\" content=\"Introduction Generative AI (GenAI) is transforming industries, necessitating the development of specialized hardware to manage its computational demands. AI accelerators...\" \/>\n<meta property=\"og:url\" content=\"https:\/\/dev-rbln-kr.locomotion.co.kr\/?p=92667\" \/>\n<meta property=\"og:site_name\" content=\"Rebellions\" \/>\n<meta property=\"article:published_time\" content=\"2024-07-11T05:34:44+00:00\" \/>\n<meta property=\"article:modified_time\" content=\"2025-08-21T08:00:39+00:00\" \/>\n<meta property=\"og:image\" content=\"https:\/\/dev-rbln-kr.locomotion.co.kr\/wp-content\/uploads\/2025\/05\/atom.png\" \/>\n\t<meta property=\"og:image:width\" content=\"1054\" \/>\n\t<meta property=\"og:image:height\" content=\"1036\" \/>\n\t<meta property=\"og:image:type\" content=\"image\/png\" \/>\n<meta name=\"author\" content=\"jiwon.kwak\" \/>\n<meta name=\"twitter:card\" content=\"summary_large_image\" \/>\n<meta name=\"twitter:creator\" content=\"@RebellionsAI\" \/>\n<meta name=\"twitter:site\" content=\"@RebellionsAI\" \/>\n<script type=\"application\/ld+json\" class=\"yoast-schema-graph\">{\"@context\":\"https:\/\/schema.org\",\"@graph\":[{\"@type\":\"Article\",\"@id\":\"https:\/\/dev-rbln-kr.locomotion.co.kr\/?p=92667#article\",\"isPartOf\":{\"@id\":\"https:\/\/dev-rbln-kr.locomotion.co.kr\/?p=92667\"},\"author\":{\"name\":\"jiwon.kwak\",\"@id\":\"https:\/\/dev-rbln-kr.locomotion.co.kr\/#\/schema\/person\/1b946c50a99f04d7b7193c47b212b6c5\"},\"headline\":\"ATOM\u2122 Architecture: Finding the Sweet Spot for GenAI\",\"datePublished\":\"2024-07-11T05:34:44+00:00\",\"dateModified\":\"2025-08-21T08:00:39+00:00\",\"mainEntityOfPage\":{\"@id\":\"https:\/\/dev-rbln-kr.locomotion.co.kr\/?p=92667\"},\"wordCount\":1702,\"publisher\":{\"@id\":\"https:\/\/dev-rbln-kr.locomotion.co.kr\/#organization\"},\"image\":{\"@id\":\"https:\/\/dev-rbln-kr.locomotion.co.kr\/?p=92667#primaryimage\"},\"thumbnailUrl\":\"https:\/\/dev-rbln-kr.locomotion.co.kr\/wp-content\/uploads\/2025\/05\/atom.png\",\"articleSection\":[\"White Papers\"],\"inLanguage\":\"ko-KR\"},{\"@type\":\"WebPage\",\"@id\":\"https:\/\/dev-rbln-kr.locomotion.co.kr\/?p=92667\",\"url\":\"https:\/\/dev-rbln-kr.locomotion.co.kr\/?p=92667\",\"name\":\"ATOM\u2122 Architecture: Finding the Sweet Spot for GenAI - Rebellions\",\"isPartOf\":{\"@id\":\"https:\/\/dev-rbln-kr.locomotion.co.kr\/#website\"},\"primaryImageOfPage\":{\"@id\":\"https:\/\/dev-rbln-kr.locomotion.co.kr\/?p=92667#primaryimage\"},\"image\":{\"@id\":\"https:\/\/dev-rbln-kr.locomotion.co.kr\/?p=92667#primaryimage\"},\"thumbnailUrl\":\"https:\/\/dev-rbln-kr.locomotion.co.kr\/wp-content\/uploads\/2025\/05\/atom.png\",\"datePublished\":\"2024-07-11T05:34:44+00:00\",\"dateModified\":\"2025-08-21T08:00:39+00:00\",\"breadcrumb\":{\"@id\":\"https:\/\/dev-rbln-kr.locomotion.co.kr\/?p=92667#breadcrumb\"},\"inLanguage\":\"ko-KR\",\"potentialAction\":[{\"@type\":\"ReadAction\",\"target\":[\"https:\/\/dev-rbln-kr.locomotion.co.kr\/?p=92667\"]}]},{\"@type\":\"ImageObject\",\"inLanguage\":\"ko-KR\",\"@id\":\"https:\/\/dev-rbln-kr.locomotion.co.kr\/?p=92667#primaryimage\",\"url\":\"https:\/\/dev-rbln-kr.locomotion.co.kr\/wp-content\/uploads\/2025\/05\/atom.png\",\"contentUrl\":\"https:\/\/dev-rbln-kr.locomotion.co.kr\/wp-content\/uploads\/2025\/05\/atom.png\",\"width\":1054,\"height\":1036},{\"@type\":\"BreadcrumbList\",\"@id\":\"https:\/\/dev-rbln-kr.locomotion.co.kr\/?p=92667#breadcrumb\",\"itemListElement\":[{\"@type\":\"ListItem\",\"position\":1,\"name\":\"Home\",\"item\":\"https:\/\/dev-rbln-kr.locomotion.co.kr\"},{\"@type\":\"ListItem\",\"position\":2,\"name\":\"ATOM\u2122 Architecture: Finding the Sweet Spot for GenAI\"}]},{\"@type\":\"WebSite\",\"@id\":\"https:\/\/dev-rbln-kr.locomotion.co.kr\/#website\",\"url\":\"https:\/\/dev-rbln-kr.locomotion.co.kr\/\",\"name\":\"\\bRebellions\",\"description\":\"Drive AI Innovation. Simple. Fast. At Scale.\",\"publisher\":{\"@id\":\"https:\/\/dev-rbln-kr.locomotion.co.kr\/#organization\"},\"alternateName\":\"\ub9ac\ubca8\ub9ac\uc628\",\"potentialAction\":[{\"@type\":\"SearchAction\",\"target\":{\"@type\":\"EntryPoint\",\"urlTemplate\":\"https:\/\/dev-rbln-kr.locomotion.co.kr\/?s={search_term_string}\"},\"query-input\":{\"@type\":\"PropertyValueSpecification\",\"valueRequired\":true,\"valueName\":\"search_term_string\"}}],\"inLanguage\":\"ko-KR\"},{\"@type\":\"Organization\",\"@id\":\"https:\/\/dev-rbln-kr.locomotion.co.kr\/#organization\",\"name\":\"Rebellions\",\"alternateName\":\"\ub9ac\ubca8\ub9ac\uc628\",\"url\":\"https:\/\/dev-rbln-kr.locomotion.co.kr\/\",\"logo\":{\"@type\":\"ImageObject\",\"inLanguage\":\"ko-KR\",\"@id\":\"https:\/\/dev-rbln-kr.locomotion.co.kr\/#\/schema\/logo\/image\/\",\"url\":\"https:\/\/dev-rbln-kr.locomotion.co.kr\/wp-content\/uploads\/2023\/02\/Web_Logo_Typing_05.gif\",\"contentUrl\":\"https:\/\/dev-rbln-kr.locomotion.co.kr\/wp-content\/uploads\/2023\/02\/Web_Logo_Typing_05.gif\",\"width\":1920,\"height\":1080,\"caption\":\"Rebellions\"},\"image\":{\"@id\":\"https:\/\/dev-rbln-kr.locomotion.co.kr\/#\/schema\/logo\/image\/\"},\"sameAs\":[\"https:\/\/x.com\/RebellionsAI\",\"https:\/\/www.youtube.com\/@Rebellions_inc\",\"https:\/\/www.linkedin.com\/company\/rebellions-ai\/\"]},{\"@type\":\"Person\",\"@id\":\"https:\/\/dev-rbln-kr.locomotion.co.kr\/#\/schema\/person\/1b946c50a99f04d7b7193c47b212b6c5\",\"name\":\"jiwon.kwak\",\"image\":{\"@type\":\"ImageObject\",\"inLanguage\":\"ko-KR\",\"@id\":\"https:\/\/dev-rbln-kr.locomotion.co.kr\/#\/schema\/person\/image\/\",\"url\":\"https:\/\/secure.gravatar.com\/avatar\/bad9ff0489ff35bd54e60b39bd7b60aae7eb336d677d5991a83975ec819e19a4?s=96&d=blank&r=g\",\"contentUrl\":\"https:\/\/secure.gravatar.com\/avatar\/bad9ff0489ff35bd54e60b39bd7b60aae7eb336d677d5991a83975ec819e19a4?s=96&d=blank&r=g\",\"caption\":\"jiwon.kwak\"}}]}<\/script>\n<!-- \/ Yoast SEO plugin. -->","yoast_head_json":{"title":"ATOM\u2122 Architecture: Finding the Sweet Spot for GenAI - Rebellions","robots":{"index":"index","follow":"follow","max-snippet":"max-snippet:-1","max-image-preview":"max-image-preview:large","max-video-preview":"max-video-preview:-1"},"canonical":"https:\/\/dev-rbln-kr.locomotion.co.kr\/?p=92667","og_locale":"ko_KR","og_type":"article","og_title":"ATOM\u2122 Architecture: Finding the Sweet Spot for GenAI - Rebellions","og_description":"Introduction Generative AI (GenAI) is transforming industries, necessitating the development of specialized hardware to manage its computational demands. AI accelerators...","og_url":"https:\/\/dev-rbln-kr.locomotion.co.kr\/?p=92667","og_site_name":"Rebellions","article_published_time":"2024-07-11T05:34:44+00:00","article_modified_time":"2025-08-21T08:00:39+00:00","og_image":[{"width":1054,"height":1036,"url":"https:\/\/dev-rbln-kr.locomotion.co.kr\/wp-content\/uploads\/2025\/05\/atom.png","type":"image\/png"}],"author":"jiwon.kwak","twitter_card":"summary_large_image","twitter_creator":"@RebellionsAI","twitter_site":"@RebellionsAI","schema":{"@context":"https:\/\/schema.org","@graph":[{"@type":"Article","@id":"https:\/\/dev-rbln-kr.locomotion.co.kr\/?p=92667#article","isPartOf":{"@id":"https:\/\/dev-rbln-kr.locomotion.co.kr\/?p=92667"},"author":{"name":"jiwon.kwak","@id":"https:\/\/dev-rbln-kr.locomotion.co.kr\/#\/schema\/person\/1b946c50a99f04d7b7193c47b212b6c5"},"headline":"ATOM\u2122 Architecture: Finding the Sweet Spot for GenAI","datePublished":"2024-07-11T05:34:44+00:00","dateModified":"2025-08-21T08:00:39+00:00","mainEntityOfPage":{"@id":"https:\/\/dev-rbln-kr.locomotion.co.kr\/?p=92667"},"wordCount":1702,"publisher":{"@id":"https:\/\/dev-rbln-kr.locomotion.co.kr\/#organization"},"image":{"@id":"https:\/\/dev-rbln-kr.locomotion.co.kr\/?p=92667#primaryimage"},"thumbnailUrl":"https:\/\/dev-rbln-kr.locomotion.co.kr\/wp-content\/uploads\/2025\/05\/atom.png","articleSection":["White Papers"],"inLanguage":"ko-KR"},{"@type":"WebPage","@id":"https:\/\/dev-rbln-kr.locomotion.co.kr\/?p=92667","url":"https:\/\/dev-rbln-kr.locomotion.co.kr\/?p=92667","name":"ATOM\u2122 Architecture: Finding the Sweet Spot for GenAI - Rebellions","isPartOf":{"@id":"https:\/\/dev-rbln-kr.locomotion.co.kr\/#website"},"primaryImageOfPage":{"@id":"https:\/\/dev-rbln-kr.locomotion.co.kr\/?p=92667#primaryimage"},"image":{"@id":"https:\/\/dev-rbln-kr.locomotion.co.kr\/?p=92667#primaryimage"},"thumbnailUrl":"https:\/\/dev-rbln-kr.locomotion.co.kr\/wp-content\/uploads\/2025\/05\/atom.png","datePublished":"2024-07-11T05:34:44+00:00","dateModified":"2025-08-21T08:00:39+00:00","breadcrumb":{"@id":"https:\/\/dev-rbln-kr.locomotion.co.kr\/?p=92667#breadcrumb"},"inLanguage":"ko-KR","potentialAction":[{"@type":"ReadAction","target":["https:\/\/dev-rbln-kr.locomotion.co.kr\/?p=92667"]}]},{"@type":"ImageObject","inLanguage":"ko-KR","@id":"https:\/\/dev-rbln-kr.locomotion.co.kr\/?p=92667#primaryimage","url":"https:\/\/dev-rbln-kr.locomotion.co.kr\/wp-content\/uploads\/2025\/05\/atom.png","contentUrl":"https:\/\/dev-rbln-kr.locomotion.co.kr\/wp-content\/uploads\/2025\/05\/atom.png","width":1054,"height":1036},{"@type":"BreadcrumbList","@id":"https:\/\/dev-rbln-kr.locomotion.co.kr\/?p=92667#breadcrumb","itemListElement":[{"@type":"ListItem","position":1,"name":"Home","item":"https:\/\/dev-rbln-kr.locomotion.co.kr"},{"@type":"ListItem","position":2,"name":"ATOM\u2122 Architecture: Finding the Sweet Spot for GenAI"}]},{"@type":"WebSite","@id":"https:\/\/dev-rbln-kr.locomotion.co.kr\/#website","url":"https:\/\/dev-rbln-kr.locomotion.co.kr\/","name":"\bRebellions","description":"Drive AI Innovation. Simple. Fast. At Scale.","publisher":{"@id":"https:\/\/dev-rbln-kr.locomotion.co.kr\/#organization"},"alternateName":"\ub9ac\ubca8\ub9ac\uc628","potentialAction":[{"@type":"SearchAction","target":{"@type":"EntryPoint","urlTemplate":"https:\/\/dev-rbln-kr.locomotion.co.kr\/?s={search_term_string}"},"query-input":{"@type":"PropertyValueSpecification","valueRequired":true,"valueName":"search_term_string"}}],"inLanguage":"ko-KR"},{"@type":"Organization","@id":"https:\/\/dev-rbln-kr.locomotion.co.kr\/#organization","name":"Rebellions","alternateName":"\ub9ac\ubca8\ub9ac\uc628","url":"https:\/\/dev-rbln-kr.locomotion.co.kr\/","logo":{"@type":"ImageObject","inLanguage":"ko-KR","@id":"https:\/\/dev-rbln-kr.locomotion.co.kr\/#\/schema\/logo\/image\/","url":"https:\/\/dev-rbln-kr.locomotion.co.kr\/wp-content\/uploads\/2023\/02\/Web_Logo_Typing_05.gif","contentUrl":"https:\/\/dev-rbln-kr.locomotion.co.kr\/wp-content\/uploads\/2023\/02\/Web_Logo_Typing_05.gif","width":1920,"height":1080,"caption":"Rebellions"},"image":{"@id":"https:\/\/dev-rbln-kr.locomotion.co.kr\/#\/schema\/logo\/image\/"},"sameAs":["https:\/\/x.com\/RebellionsAI","https:\/\/www.youtube.com\/@Rebellions_inc","https:\/\/www.linkedin.com\/company\/rebellions-ai\/"]},{"@type":"Person","@id":"https:\/\/dev-rbln-kr.locomotion.co.kr\/#\/schema\/person\/1b946c50a99f04d7b7193c47b212b6c5","name":"jiwon.kwak","image":{"@type":"ImageObject","inLanguage":"ko-KR","@id":"https:\/\/dev-rbln-kr.locomotion.co.kr\/#\/schema\/person\/image\/","url":"https:\/\/secure.gravatar.com\/avatar\/bad9ff0489ff35bd54e60b39bd7b60aae7eb336d677d5991a83975ec819e19a4?s=96&d=blank&r=g","contentUrl":"https:\/\/secure.gravatar.com\/avatar\/bad9ff0489ff35bd54e60b39bd7b60aae7eb336d677d5991a83975ec819e19a4?s=96&d=blank&r=g","caption":"jiwon.kwak"}}]}},"_links":{"self":[{"href":"https:\/\/dev-rbln-kr.locomotion.co.kr\/index.php?rest_route=\/wp\/v2\/posts\/92667","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/dev-rbln-kr.locomotion.co.kr\/index.php?rest_route=\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/dev-rbln-kr.locomotion.co.kr\/index.php?rest_route=\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/dev-rbln-kr.locomotion.co.kr\/index.php?rest_route=\/wp\/v2\/users\/8"}],"replies":[{"embeddable":true,"href":"https:\/\/dev-rbln-kr.locomotion.co.kr\/index.php?rest_route=%2Fwp%2Fv2%2Fcomments&post=92667"}],"version-history":[{"count":0,"href":"https:\/\/dev-rbln-kr.locomotion.co.kr\/index.php?rest_route=\/wp\/v2\/posts\/92667\/revisions"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/dev-rbln-kr.locomotion.co.kr\/index.php?rest_route=\/wp\/v2\/media\/92719"}],"wp:attachment":[{"href":"https:\/\/dev-rbln-kr.locomotion.co.kr\/index.php?rest_route=%2Fwp%2Fv2%2Fmedia&parent=92667"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/dev-rbln-kr.locomotion.co.kr\/index.php?rest_route=%2Fwp%2Fv2%2Fcategories&post=92667"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/dev-rbln-kr.locomotion.co.kr\/index.php?rest_route=%2Fwp%2Fv2%2Ftags&post=92667"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}