Publications

publications by categories in reversed chronological order. generated by jekyll-scholar.

† marks equal authorship (authors contribute equally to the publication).

2026

  1. CXL AnySSD: A composable CXL SSD using a CXL Type-2 device and any SSDs
    Yang Zhou, Houxiang Ji, Bikrant Sharma, Jiyuan Zhang, Yu Li, Linjie Ma, Jaeyong Lee, Myoungjun Chun, Jihong Kim, Sudarsun Kannan, and Nam Sung Kim
    IEEE/ACM International Symposium on Microarchitecture (MICRO), Oct 2026
  2. AMG: AMX-GPU cooperative acceleration of billion-scale ANNS via GEMM reformulation
    Minho Kim, Houxiang Ji, Hwanjun Lee, Jaeyoung Kang, Sungju Kim, Chunkyun Shin, Daehoon Kim, and Nam Sung Kim
    IEEE/ACM International Symposium on Microarchitecture (MICRO), Oct 2026
  3. Elastic LLM serving on FPGAs via partial reconfiguration: a case study using the Altera FPGA AI Suite
    Nachuan Wang, Mike Montano, Hyungyo Kim, Ren Wang, Phillip Swart, and Nam Sung Kim
    IEEE/ACM International Symposium on Microarchitecture (MICRO), Oct 2026
  4. NiF: hybrid in-flash and in-storage acceleration for approximate nearest neighbor search
    Myoungjun Chun, Jaeyong Lee, Gwuiin Kim, Minhyeong Kim, Inhyuk Choi, Myungsuk Kim, Nam Sung Kim, and Jihong Kim
    IEEE/ACM International Symposium on Microarchitecture (MICRO), Oct 2026
  5. ARC: adaptive reconfigurable CXL hotness monitoring
    Chihun Song, Hwayong Nam, Haijian Wang, Eojin Na, Haneul Park, Taehyung Lee, Jung Ho Ahn, and Nam Sung Kim
    IEEE/ACM International Symposium on Microarchitecture (MICRO), Oct 2026
  6. Rethinking compression for CXL memory expanders at hyperscale
    Haneul Park, Grant Ayers*, Nam Sung Kim*, Philip Levis*, and Brian Morris*
    IEEE/ACM International Symposium on Microarchitecture (MICRO), Oct 2026
    *The alphabetical order
  7. SOSP
    Borges: A low-latency distributed shared log on a CXL memory/SSD hybrid
    Haowei Chen, Yiming Xiang, Zhipeng Jia, Yan Sun, Nam Sung Kim, and Emmett Witchel
    ACM Symposium on Operating Systems Principles (SOSP), Sep 2026
  8. CAL
    HyGIN: Hybrid CPU/GPU-initiated communication for mixture-of-experts training
    Jinghan Huang and Nam Sung Kim
    July–December 2026
  9. TOCS
    Tigon: A distributed database for a CXL pod
    Yibo Huang, Haowei Chen, Newton Ni, Yan Sun, Nam Sung Kim, Vijay Chidambaram, Dixin Tang, and Emmett Witchel
    ACM Transactions on Computer Systems (TOCS), to appear, 2026
  10. MAC: Metadata acceleration for sustainable performance in big-data systems with CXL DRAM
    Dusol Lee, Yan Sun, Houxiang Ji, Vinit Gupta, Austin Antony Cruz, Inhyuk Choi, Nam Sung Kim, and Jihong Kim
    USENIX Symposium on Operating Systems Design and Implementation (OSDI), Jul 2026
  11. HCS
    XCENA MX1 CXL computational memory device
    Harry Kim, Jinin So, Dohun Kim, Sungwoo Chang, Wonjae Lee, Seunghak Lee, Hojin Nam, Jinoh Ahn, Grant Mackey, Brian Hirano, and Nam Sung Kim
    IEEE Hot Chips Symposium (HCS) , Aug 2026
  12. CAL
    A Simulator for LLM inference systems exploiting CXL memory pools
    Jinghan Huang, Hongkun Zeng, Mike Montano, Jaehong Cho, Hyunmin Choi, Jinin So, Junhyeok Im, Handeok Lee, Jongse Park, and Nam Sung Kim
    July–December 2026
  13. CAL
    KiF: Accelerating low-batch LLM inference using in-flash KV cache
    Haeeun Jung, Jaeyong Lee, Sanggu Lee, Huiwon Yun, Dongsuk Jeon, Nam Sung Kim, and Jihong Kim
    July–December 2026
  14. PVAC: A RowHammer mitigation architecture exploiting per-victim-row counting
    Jumin Kim, Seungmin Baek, Hwayong Nam, Minbok Wi, Nam Sung Kim, and Jung Ho Ahn
    IEEE/ACM International Symposium on Computer Architecture (ISCA) , Jun 2026
  15. CAL
    MMC: Metadata migration for efficient memory management in CXL DRAM systems
    Dusol Lee, Taehyung Lee, Hanseon Lee, Nam Sung Kim, and Jihong Kim
    July–December 2026
  16. CAL
    Characterizing system-level trade-offs of Intel DSA for user-space IPC offloading
    Misun Park, Richi Dubey, Yifan Yuan, Nam Sung Kim, and Ada Gavrilovska
    July–December 2026
  17. CAL
    Capacity-latency tradeoffs in CXL memory expander at hyperscale
    Haneul Park, Grant Ayers, Nam Sung Kim, Philip Levis, and Brian Morris
    July–December 2026
  18. CAL
    CXL-Tracer: Accurate, full-system memory tracing on commercial hardware via CXL memory
    Srikar Reddy Vanavasam, Austin Antony Cruz, and Nam Sung Kim
    July–December 2026
  19. ISPASS
    Compiler and system optimizations for gem5 simulator
    Haneul Park, Siddharth Agarwal, Pradyun Narkadamilli, Kiung Jung, Yongjun Park, Ipoom Jeong, and Nam Sung Kim
    IEEE International Symposium on Performance Analysis of System and Software (ISPASS) , Apr 2026
  20. ReScue: Reliable and Secure CXL memory
    Chihun Song, Austin Antony Cruz, Michael Jaemin Kim, Minbok Wi, Gaohan Ye, Kyungsan Kim, Sangyeol Lee, Jung Ho Ahn, and Nam Sung Kim
    IEEE International Symposium on High-Performance Computer Architecture , Jan 2026
  21. LILo: Harnessing the on-chip accelerators in Intel CPUs for compressed LLM inference acceleration
    Hyungyo Kim, Qirong Xia, Jinghan Huang, Nachuan Wang, Jung Ho Ahn, Younjoo Lee, Wajdi K Feghali, Ren Wang, and Nam Sung Kim
    IEEE International Symposium on High-Performance Computer Architecture , Jan 2026
  22. MemSOS: OS-guided selective memory mirroring
    Junghoon Kim, Jongheon Jeong, Seokwon Moon, Seong Hoon Seo, Yeonhong Park, Jinkyu Jeong, Nam Sung Kim, and Jae W. Lee
    IEEE International Symposium on High-Performance Computer Architecture , Jan 2026
  23. TCAS
    Characterizing the intrinsic bank-level accuracy vs. energy trade-off of SRAM-based analog in-memory computing architectures in 28 nm CMOS
    Shuo Li, Chihun Song, Hyungyo Kim, Nam Sung Kim, and Naresh R. Shanbhag
    IEEE Transactions on Circuits and Systems I: Regular Papers , 2026
  24. S&P
    SoK: Systematizing a decade of architectural RowHammer defenses through the lens of streaming algorithms
    Michael Jaemin Kim, Seungmin Baek, Jumin Kim, Hwayong Nam, Nam Sung Kim, and Jung Ho Ahn
    IEEE Symposium on Security and Privacy (S&P) , Mar 2026
  25. TiNA: Tiered network buffer architecture for fast networking in chiplet-based CPU
    Siddharth Agarwal, Tianchen Wang, Jinghan Huang, Saksham Agarwal, and Nam Sung Kim
    ACM International Conference on Architectural Support for Programming Languages and Operating Systems (ASPLOS) , Apr 2026

2025

  1. EPEPS
    Temperature-Dependent SPICE Models for UCIe Interconnects
    Ram Krishna, Ashita Victor, Srujan Penta, Atom Watanabe, Xu Chen, Muhannad S. Bakir, Nam Sung Kim, and Elyse Rosenbaum
    IEEE Conference on Electrical Performance of Electronics Packaging and Systems , Oct 2025
  2. NetZIP: algorithm/hardware co-design of in-network lossless compression for distributed large model
    Jinghan Huang, Hyungyo Kim, Nachuan Wang, Jaeyoung Kang, Hrishi Shah, Eun Kyung Lee, Minjia Zhang, Fan Lai, and Nam Sung Kim
    2025 58th IEEE/ACM International Symposium on Microarchitecture (MICRO) , Oct 2025
  3. Re-architecting end-host networking with CXL: coherence, memory, and offloading
    Houxiang Ji, Yifan Yuan, Yang Zhou, Ipoom Jeong, Ren Wang, Saksham Agarwal, and Nam Sung Kim
    2025 58th IEEE/ACM International Symposium on Microarchitecture (MICRO) , Oct 2025
  4. DRAM fault classification through large-scale field monitoring for robust memory RAS management
    Hoeju Chung, Euisang Oh, Seungmin Baek, Hyeongshin Yoon, Jaesung Yoo, Sanghwan Lee, Yongjun Lee, Arhatha Bramhanand, Brett Dodds, Yang Zhou, and Nam Sung Kim
    2025 58th IEEE/ACM International Symposium on Microarchitecture (MICRO) , Oct 2025
  5. Stratum: system-hardware co-design with tiered monolithic 3D-DRAM for efficient MoE serving
    Yue Pan, Zihan Xia, Po-Kai Hsu, Lanxiang Hu, Hyungyo Kim, Janak Sharda, Minxuan Zhou, Nam Sung Kim, Shimeng Yu, Tajana Rosing, and Mingu Kang
    2025 58th IEEE/ACM International Symposium on Microarchitecture (MICRO) , Oct 2025
  6. CAL
    CABANA: Cluster-aware query batching for accelerating billion-scale ANNS with Intel AMX
    Minho Kim, Houxiang Ji, Jaeyoung Kang, Hwanjun Lee, Daehoon Kim, and Nam Sung Kim
    Dec 2025
  7. CAL
    HINT: A hardware platform for intra-host NIC traffic and SmartNIC emulation
    Jiaqi Lou, Yu Li, Srikar Vanavasam, and Nam Sung Kim
    Dec 2025
  8. CAL
    Time series machine learning models for precise SSD access latency prediction
    Bikrant Das Sharma, Houxiang Ji, Ipoom Jeong, and Nam Sung Kim
    Dec 2025
  9. ATC
    Para-ksm: Parallelized memory deduplication with data streaming accelerator
    Houxiang Ji, Minho Kim, Seonmu Oh, Daehoon Kim, and Nam Sung Kim
    USENIX Annual Technical Conference (ATC) , Jul 2025
  10. CAL
    srNAND: A novel NAND flash organization for Enhanced Small Read Throughput in SSDs
    Jeongho Lee, Sanjun Kim, Jaeyong Lee, Jaeyong Kang, Sungjin Lee, Nam Sung Kim, and Jihong Kim
    2025
  11. Dynamic Load Balancer in Intel® Xeon® Scalable Processor: Performance analyses, enhancements, and guidelines
    Jiaqi Lou, Srikar Vanavasam, Yifan Yuan, Ren Wang, and Nam Sung Kim
    IEEE/ACM International Symposium on Computer Architecture (ISCA) , 2025
  12. A4: Microarchitecture-aware LLC management for datacenter servers with emerging I/O devices
    Haneul Park, Jiaqi Lou, Sangjin Lee, Yifan Yuan, KyoungSoo Park, Yongseok Son, Ipoom Jeong, and Nam Sung Kim
    IEEE/ACM International Symposium on Computer Architecture (ISCA) , 2025
  13. LIA: A single-GPU LLM inference acceleration with cooperative AMX-enabled CPU-GPU computation and CXL offloading
    Hyungyo Kim, Nachuan Wang, Qirong Xia, Jinghan Huang, Amir Yazdanbakhsh, and Nam Sung Kim
    IEEE/ACM International Symposium on Computer Architecture (ISCA) , 2025
  14. Universal predicate pushdown to smart storage
    Ipoom Jeong, Jinghan Huang, Chuxuan Hu, Dohyun Park, Jaeyoung Kang, Nam Sung Kim, and Yongjoo Park
    IEEE/ACM International Symposium on Computer Architecture (ISCA) , 2025
  15. Hybrid SLC-MLC RRAM mixed-signal processing-in-memory architecture for Transformer acceleration via gradient redistribution
    Chang Eun Song, Priyansh Bhatnagar, Zihan Xia, Nam Sung Kim, Tajana S Rosing, and Mingu Kang
    IEEE/ACM International Symposium on Computer Architecture (ISCA) , 2025
  16. ISPASS
    Intel In-Memory Analytics Accelerator: performance characterization and guidelines
    Jaeyoung Kang, Qirong Xia, Ipoom Jeong, Yongjoo Park, and Nam Sung Kim
    IEEE International Symposium on Performance Analysis of Systems and Software (ISPASS) , 2025
  17. DAC
    A full-system, programmable, and extensible in-memory computing simulation framework for deep learning
    Kaining Zhou, Jian Huang, Nam Sung Kim, and Naresh Shanbhag
    IEEE/ACM Design Automation Conference (DAC) , 2025
  18. TVLSI
    Exploiting chiplet integration technology for fast high-capacity DRAM modules
    Zihan Xia, Chihun Song, Ram Krishna, Ashita Victor, Srujan Penta, Muhannad S. Bakir, Elyse Rosenbaum, Nam Sung Kim, and Mingu Kang
    IEEE Transactions on Very Large Scale Integration Systems (TVLSI) , 2025
  19. Marionette: A RowHammer Attack via Row Coupling
    Seungmin Baek, Minbok Wi, Seonyong Park, Hwayong Nam, Michael Jaemin Kim, Nam Sung Kim, and Jung Ho Ahn
    Proceedings of the 30th ACM International Conference on Architectural Support for Programming Languages and Operating Systems , 2025
  20. CAL
    Hardware-Accelerated Kernel-Space Memory Compression Using Intel QAT
    Qirong Xia, Houxiang Ji, Yang Zhou, and Nam Sung Kim
    IEEE Computer Architecture Letters , 2025
  21. M5: Mastering Page Migration and Memory Management for CXL-based Tiered Memory Systems
    Yan Sun, Jongyul Kim, Zeduo Yu, Jiyuan Zhang, Siyuan Chai, Michael Jaemin Kim, Hwayong Nam, Jaehyun Park, Eojin Na, Yifan Yuan, Ren Wang, Jung Ho Ahn, Tianyin Xu, and Nam Sung Kim
    2025
  22. CAL
    X-PPR: Post Package Repair for CXL Memory
    Chihun Song, Michael Jaemin Kim, Yan Sun, Houxiang Ji, Kyungsan Kim, TaeKyeong Ko, Jung Ho Ahn, and Nam Sung Kim
    IEEE Computer Architecture Letters , 2025
  23. Special Issue on Interconnects for Chiplet Integration Technologies
    Debendra Das Sharma and Nam Sung Kim
    IEEE Micro , 2025
  24. Warped-Compaction: Maximizing GPU Register File Bandwidth Utilization via Operand Compaction
    Eunbi Jeong, Ipoom Jeong, Myung Kuk Yoon, and Nam Sung Kim
    2025 IEEE International Symposium on High Performance Computer Architecture (HPCA) , 2025
  25. Application-transparent near-memory processing architecture with memory channel network
    Nam Sung Kim and Mohammad Alian
    2025
  26. CAL
    Cooperative Memory Deduplication with Intel Data Streaming Accelerator
    Houxiang Ji, Minho Kim, Seonmu Oh, Daehoon Kim, and Nam Sung Kim
    IEEE Computer Architecture Letters , 2025

2024

  1. AttAcc! Unleashing the power of PIM for batched transformer-based generative model inference
    Jaehyun Park, Jaewan Choi, Kwanhee Kyung, Michael Jaemin Kim, Yongsuk Kwon, Nam Sung Kim, and Jung Ho Ahn
    29th ACM International Conference on Architectural Support for Programming Languages and Operating Systems , 2024
  2. An lpddr-based cxl-pnm platform for tco-efficient inference of transformer-based large language models
    Sang-Soo Park, KyungSoo Kim, Jinin So, Jin Jung, Jonggeon Lee, Kyoungwan Woo, Nayeon Kim, Younghyun Lee, Hyungyo Kim, Yongsuk Kwon, Jinhyun Kim, Jieun Lee, YeonGon Cho, Yongmin Tai, Jeonghyeon Cho, Hoyoung Song, Jung Ho Ahn, and Nam Sung Kim
    2024 IEEE International Symposium on High-Performance Computer Architecture (HPCA) , 2024
  3. A quantitative analysis and guidelines of data streaming accelerator in modern intel xeon scalable processors
    Reese Kuper, Ipoom Jeong, Yifan Yuan, Ren Wang, Narayan Ranganathan, Nikhil Rao, Jiayu Hu, Sanjay Kumar, Philip Lantz, and Nam Sung Kim
    29th ACM International Conference on Architectural Support for Programming Languages and Operating Systems , 2024
  4. Tandem processor: Grappling with emerging operators in neural networks
    Soroush Ghodrati, Sean Kinzer, Hanyang Xu, Rohan Mahapatra, Yoonsung Kim, Byung Hoon Ahn, Dong Kai Wang, Lavanya Karthikeyan, Amir Yazdanbakhsh, Jongse Park, Nam Sung Kim, and Hadi Esmaeilzadeh
    29th ACM International Conference on Architectural Support for Programming Languages and Operating Systems , 2024
  5. CAL
    Exploiting Intel® Advanced Matrix Extensions (AMX) for Large Language Model Inference
    Hyungyo Kim, Gaohan Ye, Nachuan Wang, Amir Yazdanbakhsh, and Nam Sung Kim
    IEEE Computer Architecture Letters , 2024
  6. Scalecache: A scalable page cache for multiple solid-state drives
    Kiet Tuan Pham, Seokjoo Cho, Sangjin Lee, Lan Anh Nguyen, Hyeongi Yeo, Ipoom Jeong, Sungjin Lee, Nam Sung Kim, and Yongseok Son
    2024
  7. DRAMScope: Uncovering DRAM Microarchitecture and Characteristics by Issuing Memory Commands
    Hwayong Nam, Seungmin Baek, Minbok Wi, Michael Jaemin Kim, Jaehyun Park, Chihun Song, Nam Sung Kim, and Jung Ho Ahn
    2024 ACM/IEEE 51st Annual International Symposium on Computer Architecture (ISCA) , 2024
  8. Intel accelerators ecosystem: An soc-oriented perspective: Industry product
    Yifan Yuan, Ren Wang, Narayan Ranganathan, Nikhil Rao, Sanjay Kumar, Philip Lantz, Vivekananthan Sanjeepan, Jorge Cabrera, Atul Kwatra, Rajesh Sankaran, Ipoom Jeong, and Nam Sung Kim
    2024 ACM/IEEE 51st Annual International Symposium on Computer Architecture (ISCA) , 2024
  9. TAROT: A CXL SmartNIC-Based Defense Against Multi-bit Errors by Row-Hammer Attacks
    Chihun Song, Michael Jaemin Kim, Tianchen Wang, Houxiang Ji, Jinghan Huang, Ipoom Jeong, Jaehyun Park, Hwayong Nam, Minbok Wi, Jung Ho Ahn, and Nam Sung Kim
    Proceedings of the 29th ACM International Conference on Architectural Support for Programming Languages and Operating Systems (ASPLOS) , Apr 2024
  10. Spade: Sparse pillar-based 3d object detection accelerator for autonomous driving
    Minjae Lee, Seongmin Park, Hyungmin Kim, Minyong Yoon, Janghwan Lee, Jun Won Choi, Nam Sung Kim, Mingu Kang, and Jungwook Choi
    2024 IEEE International Symposium on High-Performance Computer Architecture (HPCA) , Feb 2024
  11. Transforming the Hybrid Cloud for Emerging AI Workloads
    Deming Chen, Alaa Youssef, Ruchi Pendse, André Schleife, Bryan K Clark, Hendrik Hamann, Jingrui He, Teodoro Laino, Lav Varshney, Yuxiong Wang, Avirup Sil, Reyhaneh Jabbarvand, Tianyin Xu, Volodymyr Kindratenko, Carlos Costa, Sarita Adve, Charith Mendis, Minjia Zhang, Santiago Núñez-Corrales, Raghu Ganti, Mudhakar Srivatsa, Nam Sung Kim, Josep Torrellas, Jian Huang, Seetharami Seelam, Klara Nahrstedt, Tarek Abdelzaher, Tamar Eilam, Huimin Zhao, Matteo Manica, Ravishankar Iyer, Martin Hirzel, Vikram Adve, Darko Marinov, Hubertus Franke, Hanghang Tong, Elizabeth Ainsworth, Han Zhao, Deepak Vasisht, Minh Do, Fabio Oliveira, Giovanni Pacifici, Ruchir Puri, and Priya Nagpurkar
    arXiv preprint arXiv:2411.13239 , 2024
  12. Lupin: Tolerating Partial Failures in a CXL Pod
    Zhiting Zhu, Newton Ni, Yibo Huang, Yan Sun, Zhipeng Jia, Nam Sung Kim, and Emmett Witchel
    2024
  13. Hal: Hardware-assisted load balancing for energy-efficient snic-host cooperative computing
    Jinghan Huang, Jiaqi Lou, Srikar Vanavasam, Xinhao Kong, Houxiang Ji, Ipoom Jeong, Danyang Zhuo, Eun Kyung Lee, and Nam Sung Kim
    2024 ACM/IEEE 51st Annual International Symposium on Computer Architecture (ISCA) , 2024
  14. Harmonic: Hardware-assisted RDMA Performance Isolation for Public Clouds
    Jiaqi Lou, Xinhao Kong, Jinghan Huang, Wei Bai, Nam Sung Kim, and Danyang Zhuo
    21st USENIX Symposium on Networked Systems Design and Implementation (NSDI 24) , 2024
  15. Demystifying a CXL Type-2 Device: A Heterogeneous Cooperative Computing Perspective
    Houxiang Ji, Srikar Vanavasam, Yang Zhou, Qirong Xia, Jinghan Huang, Yifan Yuan, Ren Wang, Pekon Gupta, Bhushan Chitlur, Ipoom Jeong, and Nam Sung Kim
    2024 57th IEEE/ACM International Symposium on Microarchitecture (MICRO) , 2024
  16. Computer Architecture Having Selectable Parallel and Serial Communication Channels Between Processors and Memory
    Hao Wang and Nam Sung Kim
    2024
  17. Computer architecture having selectable parallel and serial communication channels between processors and memory
    Hao Wang and Nam Sung Kim
    2024
  18. FriendlyFoe: Adversarial Machine Learning as a Practical Architectural Defense against Side Channel Attacks
    Hyoungwook Nam, Raghavendra Pradyumna Pothukuchi, Bo Li, Nam Sung Kim, and Josep Torrellas
    2024
  19. Yield-Aware Interposer Design for UCIe Interconnects
    Ram Krishna, Ashita Victor, Srujan Penta, Xu Chen, Muhannad S Bakir, Nam Sung Kim, and Elyse Rosenbaum
    2024 IEEE 33rd Conference on Electrical Performance of Electronic Packaging and Systems (EPEPS) , 2024
  20. Systems and methods for hardware-based asynchronous persistence
    Ahmed Abulila, Nam Sung Kim, and Izzat El Hajj
    2024
  21. Tandem Processor: Grappling with Emerging Operators in Neural Networks
    Soroush Ghodrati Sean Kinzer Hanyang Xu, Rohan Mahapatra Yoonsung Kim, Byung Hoon Ahn Dong Kai Wang, Lavanya Karthikeyan Amir Yazdanbakhsh, Jongse Park, Nam Sung Kim, and Hadi Esmaeilzadeh
    2024

2023

  1. Demystifying cxl memory with genuine cxl-ready systems and devices
    Yan Sun, Yifan Yuan, Zeduo Yu, Reese Kuper, Chihun Song, Jinghan Huang, Houxiang Ji, Siddharth Agarwal, Jiaqi Lou, Ipoom Jeong, Ren Wang, Jung Ho Ahn, Tianyin Xu, and Nam Sung Kim
    2023 56th IEEE/ACM International Symposium on Microarchitecture (MICRO) , 2023
  2. Shadow: Preventing row hammer in dram with intra-subarray row shuffling
    Minbok Wi, Jaehyun Park, Seoyoung Ko, Michael Jaemin Kim, Nam Sung Kim, Eojin Lee, and Jung Ho Ahn
    2023 IEEE International Symposium on High-Performance Computer Architecture (HPCA) , 2023
  3. Rambda: Rdma-driven acceleration framework for memory-intensive µs-scale datacenter applications
    Yifan Yuan, Jinghan Huang, Yan Sun, Tianchen Wang, Jacob Nelson, Dan RK Ports, Yipeng Wang, Ren Wang, Charlie Tai, and Nam Sung Kim
    2023 IEEE International Symposium on High-Performance Computer Architecture (HPCA) , 2023
  4. CAL
    Unleashing the potential of pim: Accelerating large batched inference of transformer-based generative models
    Jaewan Choi, Jaehyun Park, Kwanhee Kyung, Nam Sung Kim, and Jung Ho Ahn
    IEEE Computer Architecture Letters , 2023
  5. ATC
    STYX: Exploiting SmartNIC capability to reduce datacenter memory tax
    Houxiang Ji, Mark Mansi, Yan Sun, Yifan Yuan, Jinghan Huang, Reese Kuper, Michael M Swift, and Nam Sung Kim
    2023 USENIX Annual Technical Conference (USENIX ATC 23) , 2023
  6. How to kill the second bird with one ecc: The pursuit of row hammer resilient dram
    Michael Jaemin Kim, Minbok Wi, Jaehyun Park, Seoyoung Ko, Jaeyoung Choi, Hwayoung Nam, Nam Sung Kim, Jung Ho Ahn, and Eojin Lee
    2023
  7. CAL
    X-ray: Discovering dram internal structure and error characteristics by issuing memory commands
    Hwayong Nam, Seungmin Baek, Minbok Wi, Michael Jaemin Kim, Jaehyun Park, Chihun Song, Nam Sung Kim, and Jung Ho Ahn
    IEEE Computer Architecture Letters , 2023
  8. HotOS
    Towards a Manageable Intra-Host Network
    Xinhao Kong, Jiaqi Lou, Wei Bai, Nam Sung Kim, and Danyang Zhuo
    2023
  9. Making sense of using a smartnic to reduce datacenter tax from slo and tco perspectives
    Jinghan Huang, Jiaqi Lou, Yan Sun, Tianchen Wang, Eun Kyung Lee, and Nam Sung Kim
    2023 IEEE International Symposium on Workload Characterization (IISWC) , 2023
  10. ISPASS
    Analyzing Energy Efficiency of a Server with a SmartNIC under SLO Constraints
    Jinghan Huang, Jiaqi Lou, Yan Sun, Tianchen Wang, Eun Kyung Lee, and Nam Sung Kim
    2023 IEEE International Symposium on Performance Analysis of Systems and Software (ISPASS) , 2023
  11. CAL
    Nohammer: Preventing row hammer with last-level cache management
    Seunghak Lee, Ki-Dong Kang, Gyeongseo Park, Nam Sung Kim, and Daehoon Kim
    IEEE Computer Architecture Letters , 2023
  12. Defensive ml: Defending architectural side-channels with adversarial obfuscation
    Hyoungwook Nam, Raghavendra Pradyumna Pothukuchi, Bo Li, Nam Sung Kim, and Josep Torrellas
    arXiv preprint arXiv:2302.01474 , 2023
  13. Computer architecture having selectable parallel and serial communication channels between processors and memory
    Hao Wang and Nam Sung Kim
    2023
  14. CAL
    LADIO: Leakage-aware direct I/O for I/O-intensive workloads
    Ipoom Jeong, Jiaqi Lou, Yongseok Son, Yongjoo Park, Yifan Yuan, and Nam Sung Kim
    IEEE Computer Architecture Letters , 2023
  15. Mesa: Microarchitecture extensions for spatial architecture generation
    Dong Kai Wang, Jiaqi Lou, Naiyin Jin, Edwin Mascarenhas, Rohan Mahapatra, Sean Kinzer, Soroush Ghodrati, Amir Yazdanbakhsh, Hadi Esmaeilzadeh, and Nam Sung Kim
    2023
  16. Special issue on emerging system interconnects
    John Kim and Nam Sung Kim
    IEEE Micro , 2023
  17. Dataflow-based general-purpose processor architectures
    Nam Sung Kim and Dong Kai Wang
    2023
  18. Triple-A: Early Operand Collector Allocation for Maximizing GPU Register Bank Utilization
    Ipoom Jeong, Eunbi Jeong, Nam Sung Kim, and Myung Kuk Yoon
    IEEE Embedded Systems Letters , 2023

2022

  1. MLSys
    Bns-gcn: Efficient full-graph training of graph convolutional networks with partition-parallelism and random boundary node sampling
    Cheng Wan, Youjie Li, Ang Li, Nam Sung Kim, and Yingyan Lin
    Proceedings of Machine Learning and Systems , 2022
  2. Pipegcn: Efficient full-graph training of graph convolutional networks with pipelined feature communication
    Cheng Wan, Youjie Li, Cameron R Wolfe, Anastasios Kyrillidis, Nam Sung Kim, and Yingyan Lin
    arXiv preprint arXiv:2203.10428 , 2022
  3. Aquabolt-XL HBM2-PIM LPDDR5-PIM with in-memory processing and AXDIMM with acceleration buffer
    Jin Hyun Kim, Shin-Haeng Kang, Sukhan Lee, Hyeonsu Kim, Yuhwan Ro, Seungwon Lee, David Wang, Jihyun Choi, Jinin So, YeonGon Cho, JoonHo Song, Jeonghyeon Cho, Kyomin Sohn, and Nam Sung Kim
    IEEE Micro , 2022
  4. Unlocking the power of inline Floating-Point operations on programmable switches
    Yifan Yuan, Omar Alama, Jiawei Fei, Jacob Nelson, Dan RK Ports, Amedeo Sapio, Marco Canini, and Nam Sung Kim
    19th USENIX Symposium on Networked Systems Design and Implementation (NSDI 22) , 2022
  5. Harmony: Overcoming the hurdles of GPU memory capacity to train massive DNN models on commodity servers
    Youjie Li, Amar Phanishayee, Derek Murray, Jakub Tarnawski, and Nam Sung Kim
    48th International Conference on Very Large Databases , 2022
  6. IDIO: Network-driven inbound network data orchestration on server processors
    Mohammad Alian, Siddharth Agarwal, Jongmin Shin, Neel Patel, Yifan Yuan, Daehoon Kim, Ren Wang, and Nam Sung Kim
    2022 55th IEEE/ACM International Symposium on Microarchitecture (MICRO) , 2022
  7. An FPGA-based RNN-T Inference Accelerator with PIM-HBM
    Shinhaeng Kang, Sukhan Lee, Byeongho Kim, Hweesoo Kim, Kyomin Sohn, Nam Sung Kim, and Eojin Lee
    2022
  8. Asap: architecture support for asynchronous persistence
    Ahmed Abulila, Izzat El Hajj, Myoungsoo Jung, and Nam Sung Kim
    2022
  9. Rethinking DRAM’s page mode with STT-MRAM
    Byoungchan Oh, Nilmini Abeyratne, Nam Sung Kim, Jeongseob Ahn, Ronald G Dreslinski, and Trevor Mudge
    IEEE Transactions on Computers , 2022
  10. Application-transparent near-memory processing architecture with memory channel network
    Nam Sung Kim and Mohammad Alian
    2022
  11. Network-driven packet context-aware power management for client-server architecture
    Nam Sung Kim and Mohammad Alian
    2022
  12. Coordinated Science Laboratory 70th Anniversary Symposium: The Future of Computing
    Klara Nahrstedt, Naresh Shanbhag, Vikram Adve, Nancy Amato, Romit Roy Choudhury, Carl Gunter, Nam Sung Kim, Olgica Milenkovic, Sayan Mitra, Lav Varshney, Yurii Vlasov, Sarita Adve, Rashid Bashir, Andreas Cangellaris, James DiCarlo, Katie Driggs-Campbell, Nick Feamster, Mattia Gazzola, Karrie Karahalios, Sanmi Koyejo, Paul Kwiat, Bo Li, Negar Mehr, Ravish Mehra, Andrew Miller, Daniela Rus, Alex Schwing, and Anshumali Shrivastava
    arXiv preprint arXiv:2210.08974 , 2022
  13. HAMS: Hardware Automated Memory-over-Storage for Large-scale Memory Expansion
    Jie Zhang, Miryeong Kwon, Donghyun Gouk, Sungjoon Koh, Nam Sung Kim, Mahmut Taylan Kandemir, and Myoungsoo Jung
    13rd Annual Non-Volatile Memories Workshop (NVMW), 2022 , 2022

2021

  1. Hardware architecture and software stack for PIM based on commercial DRAM technology: Industrial product
    Sukhan Lee, Shin-haeng Kang, Jaehoon Lee, Hyeonsu Kim, Eojin Lee, Seungwoo Seo, Hosang Yoon, Seungwon Lee, Kyounghwan Lim, Hyunsung Shin, Jinhyun Kim, O Seongil, Anand Iyer, David Wang, Kyomin Sohn, and Nam Sung Kim
    2021 ACM/IEEE 48th Annual International Symposium on Computer Architecture (ISCA) , 2021
  2. 25.4 a 20nm 6gb function-in-memory dram based on hbm2 with a 1.2 tflops programmable computing unit using bank-level parallelism for machine learning applications
    Young-Cheon Kwon, Suk Han Lee, Jaehoon Lee, Sang-Hyuk Kwon, Je Min Ryu, Jong-Pil Son, O Seongil, Hak-Soo Yu, Haesuk Lee, Soo Young Kim, Youngmin Cho, Jin Guk Kim, Jongyoon Choi, Hyun-Sung Shin, Jin Kim, BengSeng Phuah, HyoungMin Kim, Myeong Jun Song, Ahn Choi, Daeho Kim, SooYoung Kim, Eun-Bong Kim, David Wang, Shinhaeng Kang, Yuhwan Ro, Seungwoo Seo, JoonHo Song, Jaeyoun Youn, Kyomin Sohn, and Nam Sung Kim
    2021 IEEE International Solid-State Circuits Conference (ISSCC) , 2021
  3. Near-memory processing in action: Accelerating personalized recommendation with axdimm
    Liu Ke, Xuan Zhang, Jinin So, Jong-Geon Lee, Shin-Haeng Kang, Sukhan Lee, Songyi Han, YeonGon Cho, Jin Hyun Kim, Yongsuk Kwon, KyungSoo Kim, Jin Jung, Ilkwon Yun, Sung Joo Park, Hyunsun Park, Joonho Song, Jeonghyeon Cho, Kyomin Sohn, Nam Sung Kim, and Hsien-Hsin S Lee
    IEEE Micro , 2021
  4. Don’t forget the I/O when allocating your LLC
    Yifan Yuan, Mohammad Alian, Yipeng Wang, Ren Wang, Ilia Kurakin, Charlie Tai, and Nam Sung Kim
    2021 ACM/IEEE 48th Annual International Symposium on Computer Architecture (ISCA) , 2021
  5. 25.2 a 16gb sub-1v 7.14 gb/s/pin lpddr5 sdram applying a mosaic architecture with a short-feedback 1-tap dfe an fss bus with low-level swing and an adaptively controlled body biasing in a 3 rd-generation 10nm dram
    Yong-Hun Kim, Hyung-Jin Kim, Jaemin Choi, Min-Su Ahn, Dongkeon Lee, Seung-Hyun Cho, Dong-Yeon Park, Young-Jae Park, Min-Soo Jang, Yong-Jun Kim, Jinyong Choi, Sung-Woo Yoon, Jae-Woo Jung, Jae-Koo Park, Jae-Woo Lee, Dae-Hyun Kwon, Hyung-Seok Cha, Si-Hyeong Cho, Seong-Hoon Kim, Jihwa You, Kyoung-Ho Kim, Dae-Hyun Kim, Byung-Cheol Kim, Young-Kwan Kim, Jun-Ho Kim, Seouk-Kyu Choi, Chan-Young Kim, Byong-Wook Na, Hye-In Choi, Reum Oh, Jeong-Don Ihm, Seung-Jun Bae, Nam Sung Kim, and Jung-Bae Lee
    2021 IEEE International Solid-State Circuits Conference (ISSCC) , 2021
  6. Nmap: Power management based on network packet processing mode transition for latency-critical workloads
    Ki-Dong Kang, Gyeongseo Park, Hyosang Kim, Mohammad Alian, Nam Sung Kim, and Daehoon Kim
    2021
  7. Dml: Dynamic partial reconfiguration with scalable task scheduling for multi-applications on fpgas
    Ashutosh Dhar, Edward Richter, Mang Yu, Wei Zuo, Xiaohao Wang, Nam Sung Kim, and Deming Chen
    IEEE Transactions on Computers , 2021
  8. Revamping storage class memory with hardware automated memory-over-storage solution
    Jie Zhang, Miryeong Kwon, Donghyun Gouk, Sungjoon Koh, Nam Sung Kim, Mahmut Taylan Kandemir, and Myoungsoo Jung
    2021 ACM/IEEE 48th Annual International Symposium on Computer Architecture (ISCA) , 2021
  9. Network-centric architecture and algorithms to accelerate distributed training of neural networks
    Nam Sung Kim, LI Youjie, and Alexander Gerhard Schwing
    2021
  10. Diag: a dataflow-inspired architecture for general-purpose processors
    Dong Kai Wang and Nam Sung Kim
    26th ACM International Conference on Architectural Support for Programming Languages and Operating Systems , 2021
  11. Osc: An online self-configuring big data framework for optimization of qos
    Zhendong Bei, Nam Sung Kim, Kai HWang, and Zhibin Yu
    IEEE Transactions on Computers , 2021
  12. QEI: Query acceleration can be generic and efficient in the cloud
    Yifan Yuan, Yipeng Wang, Ren Wang, Rangeen Basu Roy Chowhury, Charlie Tai, and Nam Sung Kim
    2021 IEEE International Symposium on High-Performance Computer Architecture (HPCA) , 2021
  13. A Journey to a Commercial-Grade Processing-In-Memory (PIM) Chip Development HPCA 2021
    Nam Sung Kim
    The 27th IEEE International Symposium on High-Performance Computer Architecture (PCA-27), Seoul, South Korea, URL: https://hpca-conf. org/2021/keynotes/, dated Mar , 2021
  14. Greendimm: Os-assisted dram power management for dram with a sub-array granularity power-down state
    Seunghak Lee, Ki-Dong Kang, Hwanjun Lee, Hyungwon Park, Younghoon Son, Nam Sung Kim, and Daehoon Kim
    2021
  15. In-Memory Near-Data Approximate Acceleration
    Nam Sung Kim, Hadi Esmaeilzadeh, and Amir Yazdanbakhsh
    2021
  16. Virtual-Cache: A cache-line borrowing technique for efficient GPU cache architectures
    Bingchao Li, Jizeng Wei, and Nam Sung Kim
    Microprocessors and Microsystems , 2021
  17. Memory module memory device and processing device having a processor mode and memory system
    O Seong-il, Nam Sung Kim, SON Young-Hoon, Chan-kyung Kim, Ho-young Song, Jung Ho Ahn, and Sang-joon Hwang
    2021
  18. FlatFlash system for byte granularity accessibility of memory in a unified memory-storage hierarchy
    Ahmed AbulilaAbulila, Vikram Sharma Mailthody, Zaid Qures, Hijian Huang, Nam Sung Kim, Jinjun Xiong, and Wen-mei Hwu
    2021

2020

  1. Planaria: Dynamic architecture fission for spatial multi-tenant acceleration of deep neural networks
    Soroush Ghodrati, Byung Hoon Ahn, Joon Kyung Kim, Sean Kinzer, Brahmendra Reddy Yatham, Navateja Alla, Hardik Sharma, Mohammad Alian, Eiman Ebrahimi, Nam Sung Kim, Cliff Young, and Hadi Esmaeilzadeh
    2020 53rd Annual IEEE/ACM International Symposium on Microarchitecture (MICRO) , 2020
  2. A 16-GB 640-GB/s HBM2E DRAM with a data-bus window extension technique and a synergetic on-die ECC scheme
    Ki Chul Chun, Yong Ki Kim, Yesin Ryu, Jaewon Park, Chi Sung Oh, Young Yong Byun, So Young Kim, Dong Hak Shin, Jun Gyu Lee, Byung-Kyu Ho, Min-Sang Park, Seong-Jin Cho, Seunghan Woo, Byoung Mo Moon, Beomyong Kil, Sungoh Ahn, Jae Hoon Lee, Soo Young Kim, Seouk-Kyu Choi, Jae-Seung Jeong, Sung-Gi Ahn, Jihye Kim, Jun Jin Kong, Kyomin Sohn, Nam Sung Kim, and Jung-Bae Lee
    IEEE Journal of Solid-State Circuits , 2020
  3. Mixed-signal charge-domain acceleration of deep neural networks through interleaved bit-partitioned arithmetic
    Soroush Ghodrati, Hardik Sharma, Sean Kinzer, Amir Yazdanbakhsh, Jongse Park, Nam Sung Kim, Doug Burger, and Hadi Esmaeilzadeh
    2020
  4. Data direct I/O characterization for future I/O system exploration
    Mohammad Alian, Yifan Yuan, Jie Zhang, Ren Wang, Myoungsoo Jung, and Nam Sung Kim
    2020 IEEE International Symposium on Performance Analysis of Systems and Software (ISPASS) , 2020
  5. Bit-Parallel Vector Composability for Neural Acceleration
    Hadi Esmaeilzadeh Soroush Ghodrati, Hardik Sharma, Cliff Young, and Nam Sung Kim
    IEEE/ACM Design Automation Conference , 2020
  6. Babelfish: Fusing address translations for containers
    Dimitrios Skarlatos, Umur Darbaz, Bhargava Gopireddy, Nam Sung Kim, and Josep Torrellas
    2020 ACM/IEEE 47th Annual International Symposium on Computer Architecture (ISCA) , 2020
  7. An 8.5-Gb/s/Pin 12-Gb LPDDR5 SDRAM with a hybrid-bank architecture low power and speed-boosting techniques
    Chang-Kyo Lee, Hyung-Joon Chi, Jin-Seok Heo, Jung-Hwan Park, Jin-Hun Jang, Dongkeon Lee, Jae-Hoon Jung, Dong-Hun Lee, Dae-Hyun Kim, Kihan Kim, Sang-Yun Kim, Dukha Park, Youngil Lim, Geuntae Park, Seung-Jun Lee, Seungki Hong, Dae-Hyun Kwon, Isak Hwang, Byongwook Na, Kyung-Ryun Kim, Seouk-Kyu Choi, Hyein Choi, Won-Il Bae, Jeong-Don Ihm, Seung-Jun Bae, Nam Sung Kim, and Jung-Bae Lee
    IEEE Journal of Solid-State Circuits , 2020
  8. Freac cache: Folded-logic reconfigurable computing in the last level cache
    Ashutosh Dhar, Xiaohao Wang, Hubertus Franke, Jinjun Xiong, Jian Huang, Wen-mei Hwu, Nam Sung Kim, and Deming Chen
    2020 53rd Annual IEEE/ACM International Symposium on Microarchitecture (MICRO) , 2020
  9. Leveraging dynamic partial reconfiguration with scalable ILP based task scheduling
    Ashutosh Dhar, Mang Yu, Wei Zuo, Xiaohao Wang, Nam Sung Kim, and Deming Chen
    2020 33rd International Conference on VLSI Design and 2020 19th International Conference on Embedded Systems (VLSID) , 2020
  10. CAL
    IDIO: Orchestrating Inbound Network Data on Server Processors
    Mohammad Alian, Jongmin Shin, Ki-Dong Kang, Ren Wang, Alexandros Daglis, Daehoon Kim, and Nam Sung Kim
    IEEE Computer Architecture Letters , 2020
  11. BDS-GCN: Efficient full-graph training of graph convolutional nets with partition-parallelism and boundary sampling
    Cheng Wan, Youjie Li, Nam Sung Kim, and Yingyan Lin
    2020
  12. CAL
    FastDrain: Removing page victimization overheads in NVMe storage stack
    Jie Zhang, Miryeong Kwon, Sanghyun Han, Nam Sung Kim, Mahmut Kandemir, and Myoungsoo Jung
    IEEE Computer Architecture Letters , 2020
  13. Graphic processor unit providing reduced storage costs for similar operands
    Nam Sung Kim and Zhenhong Liu
    2020
  14. Errata to “Exploring Fault-Tolerant Erasure Codes for Scalable All-Flash Array Clusters”
    Sungjoon Koh, Jie Zhang, Miryeong Kwon, Jungyeon Yoon, David Donofrio, Nam Sung Kim, and Myoungsoo Jung
    IEEE Transactions on Parallel & Distributed Systems , 2020

2019

  1. Flatflash: Exploiting the byte-accessibility of ssds within a unified memory-storage hierarchy
    Ahmed Abulila, Vikram Sharma Mailthody, Zaid Qureshi, Jian Huang, Nam Sung Kim, Jinjun Xiong, and Wen-mei Hwu
    24th ACM International Conference on Architectural Support for Programming Languages and Operating Systems , 2019
  2. Netdimm: Low-latency near-memory network interface architecture
    Mohammad Alian and Nam Sung Kim
    2019
  3. Memory module memory device and processing device having a processor mode and memory system
    O Seong-il, Nam Sung Kim, SON Young-Hoon, Chan-kyung Kim, Ho-young Song, Jung Ho Ahn, and Sang-joon Hwang
    2019
  4. LL-PCM: Low-latency phase change memory architecture
    Nam Sung Kim, Choungki Song, Woo Young Cho, Jian Huang, and Myoungsoo Jung
    2019
  5. An efficient GPU cache architecture for applications with irregular memory access patterns
    Bingchao Li, Jizeng Wei, Jizhou Sun, Murali Annavaram, and Nam Sung Kim
    ACM Transactions on Architecture and Code Optimization (TACO) , 2019
  6. AxMemo: Hardware-compiler co-design for approximate code memoization
    Zhenhong Liu, Amir Yazdanbakhsh, Dong Kai Wang, Hadi Esmaeilzadeh, and Nam Sung Kim
    2019
  7. Near-memory and in-storage fpga acceleration for emerging cognitive computing workloads
    Ashutosh Dhar, Sitao Huang, Jinjun Xiong, Damir Jamsek, Bruno Mesnet, Jian Huang, Nam Sung Kim, Wen-mei Hwu, and Deming Chen
    2019 IEEE Computer Society Annual Symposium on VLSI (ISVLSI) , 2019
  8. CAL
    Exploiting os-level memory offlining for dram power management
    Seunghak Lee, Nam Sung Kim, and Daehoon Kim
    IEEE Computer Architecture Letters , 2019
  9. Practical near-data processing to evolve memory and storage devices into mainstream heterogeneous computing systems
    Nam Sung Kim and Pankaj Mehra
    2019
  10. An energy-efficient programmable mixed-signal accelerator for machine learning algorithms
    Mingu Kang, Prakalp Srivastava, Vikram Adve, Nam Sung Kim, and Naresh R Shanbhag
    IEEE micro , 2019
  11. CAL
    Network packet processing mode-aware power management for data center servers
    Ki-Dong Kang, Gyeongseo Park, Nam Sung Kim, and Daehoon Kim
    IEEE Computer Architecture Letters , 2019
  12. SMART: STT-MRAM architecture for smart activation and sensing
    Byoungchan Oh, Nilmini Abeyratne, Nam Sung Kim, Ronald G Dreslinski, and Trevor Mudge
    2019
  13. A2M: Approximate Algebraic Memory Using Polynomials Rings
    Dong Kai Wang and Nam Sung Kim
    2019 IEEE/ACM International Symposium on Low Power Electronics and Design (ISLPED) , 2019
  14. Ghost routers: energy-efficient asymmetric multicore processors with symmetric NoCs
    Hyojun Son, Hanjoon Kim, Hao Wang, Nam Sung Kim, and John Kim
    2019

2018

  1. Pipe-SGD: A decentralized pipelined SGD framework for distributed deep net training
    Youjie Li, Mingchao Yu, Songze Li, Salman Avestimehr, Nam Sung Kim, and Alexander Schwing
    Advances in Neural Information Processing Systems , 2018
  2. FlashShare: Punching Through Server Storage Stack from Kernel to Firmware for Ultra-Low Latency SSDs
    Jie Zhang, Miryeong Kwon, Donghyun Gouk, Sungjoon Koh, Changlim Lee, Mohammad Alian, Myoungjun Chun, Mahmut Taylan Kandemir, Nam Sung Kim, Jihong Kim, and Myoungsoo Jung
    13th USENIX Symposium on Operating Systems Design and Implementation (OSDI 18) , 2018
  3. Ganax: A unified mimd-simd acceleration for generative adversarial networks
    Amir Yazdanbakhsh, Hajar Falahati, Philip J Wolfe, Hadi Esmaeilzadeh, and Kambiz Samadi
    2018 ACM/IEEE 45th annual international symposium on computer architecture (ISCA) , 2018
  4. A network-centric hardware/algorithm co-design to accelerate distributed training of deep neural networks
    Youjie Li, Jongse Park, Mohammad Alian, Yifan Yuan, Zheng Qu, Peitian Pan, Ren Wang, Alexander Schwing, Hadi Esmaeilzadeh, and Nam Sung Kim
    2018 51st Annual IEEE/ACM International Symposium on Microarchitecture (MICRO) , 2018
  5. Amber: Enabling precise full-system simulation with detailed modeling of all SSD resources
    Donghyun Gouk, Miryeong Kwon, Jie Zhang, Sungjoon Koh, Wonil Choi, Nam Sung Kim, Mahmut Kandemir, and Myoungsoo Jung
    2018 51st Annual IEEE/ACM International Symposium on Microarchitecture (MICRO) , 2018
  6. Flexigan: An end-to-end solution for fpga acceleration of generative adversarial networks
    Amir Yazdanbakhsh, Michael Brzozowski, Behnam Khaleghi, Soroush Ghodrati, Kambiz Samadi, Nam Sung Kim, and Hadi Esmaeilzadeh
    2018 IEEE 26th Annual International Symposium on Field-Programmable Custom Computing Machines (FCCM) , 2018
  7. Gradiveq: Vector quantization for bandwidth-efficient gradient aggregation in distributed cnn training
    Mingchao Yu, Zhifeng Lin, Krishna Narra, Songze Li, Youjie Li, Nam Sung Kim, Alexander Schwing, Murali Annavaram, and Salman Avestimehr
    Advances in Neural Information Processing Systems , 2018
  8. PROMISE: An end-to-end design of a programmable mixed-signal accelerator for machine-learning algorithms
    Prakalp Srivastava, Mingu Kang, Sujan K Gonugondla, Sungmin Lim, Jungwook Choi, Vikram Adve, Nam Sung Kim, and Naresh Shanbhag
    2018 ACM/IEEE 45th Annual International Symposium on Computer Architecture (ISCA) , 2018
  9. Application-transparent near-memory processing architecture with memory channel network
    Mohammad Alian, Seung Won Min, Hadi Asgharimoghaddam, Ashutosh Dhar, Dong Kai Wang, Thomas Roewer, Adam McPadden, Oliver O’Halloran, Deming Chen, Jinjun Xiong, Daehoon Kim, Wen-mei Hwu, and Nam Sung Kim
    2018 51st Annual IEEE/ACM International Symposium on Microarchitecture (MICRO) , 2018
  10. In-dram near-data approximate acceleration for gpus
    Amir Yazdanbakhsh, Choungki Song, Jacob Sacks, Pejman Lotfi-Kamran, Hadi Esmaeilzadeh, and Nam Sung Kim
    2018
  11. Leveraging power-performance relationship of energy-efficient modern DRAM devices
    Sukhan Lee, Hyunyoon Cho, Young Hoon Son, Yuhwan Ro, Nam Sung Kim, and Jung Ho Ahn
    IEEE Access , 2018
  12. SiMul: An algorithm-driven approximate multiplier design for machine learning
    Zhenhong Liu, Amir Yazdanbakhsh, Taejoon Park, Hadi Esmaeilzadeh, and Nam Sung Kim
    IEEE Micro , 2018
  13. Computer architecture having selectable parallel and serial communication channels between processors and memory
    Hao Wang and Nam Sung Kim
    2018
  14. Cta-aware prefetching and scheduling for gpu
    Gunjae Koo, Hyeran Jeon, Zhenhong Liu, Nam Sung Kim, and Murali Annavaram
    2018 IEEE International Parallel and Distributed Processing Symposium (IPDPS) , 2018
  15. Memory system and method for error correction of memory
    Jung Ho Ahn and KIM Namsung
    2018
  16. Load-triggered warp approximation on GPU
    Zhenhong Liu, Daniel Wong, and Nam Sung Kim
    2018
  17. CIAO: Cache interference-aware throughput-oriented architecture and scheduling for GPUs
    Jie Zhang, Shuwen Gao, Nam Sung Kim, and Myoungsoo Jung
    2018 IEEE International Parallel and Distributed Processing Symposium (IPDPS) , 2018
  18. Exploring fault-tolerant erasure codes for scalable all-flash array clusters
    Sungjoon Koh, Jie Zhang, Miryeong Kwon, Jungyeon Yoon, David Donofrio, Nam Sung Kim, and Myoungsoo Jung
    IEEE Transactions on Parallel and Distributed Systems , 2018
  19. Shared row buffer system for asymmetric memory
    Hao Wang and Nam Sung Kim
    2018
  20. 3D-Xpath: High-density managed dram architecture with cost-effective alternative paths for memory transactions
    Sukhan Lee, Kiwon Lee, Minchul Sung, Mohammad Alian, Chankyung Kim, Wooyeong Cho, Reum Oh, Seongil O, Jung Ho Ahn, and Nam Sung Kim
    2018
  21. Simulating PCI-Express interconnect for future system exploration
    Mohammad Alian, Krishna Parasuram Srinivasan, and Nam Sung Kim
    2018 IEEE International Symposium on Workload Characterization (IISWC) , 2018
  22. Ganax: A unified mimd-simd acceleration for generative adversarial networks. In 2018 ACM/IEEE 45th Annual International Symposium on Computer Architecture (ISCA)
    Amir Yazdanbakhsh, Kambiz Samadi, Nam Sung Kim, and Hadi Esmaeilzadeh
    2018
  23. VIP: Virtual performance-state for efficient power management of virtual machines
    Ki-Dong Kang, Mohammad Alian, Daehoon Kim, Jaehyuk Huh, and Nam Sung Kim
    2018
  24. Design and implementation of SSD-assisted backup and recovery for database systems
    Yongseok Son, Moonsub Kim, Sunggon Kim, Heon Young Yeom, Nam Sung Kim, and Hyuck Han
    IEEE Transactions on Knowledge and Data Engineering , 2018
  25. Practical challenges in supporting function in memory
    Nam Sung Kim
    2018 IEEE Asian Solid-State Circuits Conference (A-SSCC) , 2018
  26. A load balancing technique for memory channels
    Byoungchan Oh, Nam Sung Kim, Jeongseob Ahn, Bingchao Li, Ronald G Dreslinski, and Trevor Mudge
    2018
  27. Approximate ultra-low voltage many-core processor design
    Nam Sung Kim and Ulya R Karpuzcu
    2018
  28. CAL
    Semi-coherent dma: An alternative i/o coherency management for embedded systems
    Seungwon Min, Mohammad Alian, Wen-Mei Hwu, and Nam Sung Kim
    IEEE Computer Architecture Letters , 2018

2017

  1. dist-gem5: Distributed simulation of computer clusters
    Alian Mohammad, Umur Darbaz, Gabor Dozsa, Stephan Diestelhorst, Daehoon Kim, and Nam Sung Kim
    2017 IEEE International Symposium on Performance Analysis of Systems and Software (ISPASS) , 2017
  2. Defect analysis and cost-effective resilience architecture for future DRAM devices
    Sanguhn Cha, O Seongil, Hyunsung Shin, Sangjoon Hwang, Kwangil Park, Seong Jin Jang, Joo Sun Choi, Gyo Young Jin, Young Hoon Son, Hyunyoon Cho, Jung Ho Ahn, and Nam Sung Kim
    2017 IEEE International Symposium on High Performance Computer Architecture (HPCA) , 2017
  3. CAL
    SimpleSSD: Modeling solid state drives for holistic system simulation
    Myoungsoo Jung, Jie Zhang, Ahmed Abulila, Miryeong Kwon, Narges Shahidi, John Shalf, Nam Sung Kim, and Mahmut Kandemir
    IEEE Computer Architecture Letters , 2017
  4. Smart gait-aid glasses for Parkinson’s disease patients
    DaeHan Ahn, Hyerim Chung, Ho-Won Lee, Kyunghun Kang, Pan-Woo Ko, Nam Sung Kim, and Taejoon Park
    IEEE Transactions on Biomedical Engineering , 2017
  5. Elastic-cache: GPU cache architecture for efficient fine-and coarse-grained cache-line management
    Bingchao Li, Jizhou Sun, Murali Annavaram, and Nam Sung Kim
    2017 IEEE International Parallel and Distributed Processing Symposium (IPDPS) , 2017
  6. Pageforge: a near-memory content-aware page-merging architecture
    Dimitrios Skarlatos, Nam Sung Kim, and Josep Torrellas
    2017
  7. Heterogeneous computing meets near-memory acceleration and high-level synthesis in the post-moore era
    Nam Sung Kim, Deming Chen, Jinjun Xiong, and Wen-mei W Hwu
    IEEE Micro , 2017
  8. Collaborative (cpu+ gpu) algorithms for triangle counting and truss decomposition on the minsky architecture: Static graph challenge: Subgraph isomorphism
    Ketan Date, Keven Feng, Rakesh Nagi, Jinjun Xiong, Nam Sung Kim, and Wen-Mei Hwu
    2017 IEEE High Performance Extreme Computing Conference (HPEC) , 2017
  9. Ncap: Network-driven packet context-aware power management for client-server architecture
    Mohammad Alian, Ahmed HMO Abulila, Lokesh Jindal, Daehoon Kim, and Nam Sung Kim
    2017 IEEE International Symposium on High Performance Computer Architecture (HPCA) , 2017
  10. G-scalar: Cost-effective generalized scalar execution architecture for power-efficient gpus
    Zhenhong Liu, Syed Gilani, Murali Annavaram, and Nam Sung Kim
    2017 IEEE International Symposium on High Performance Computer Architecture (HPCA) , 2017
  11. Rebooting the data access hierarchy of computing systems
    W Hwu Wen-mei, Izzat El Hajj, Simon Garcia De Gonzalo, Carl Pearson, Nam Sung Kim, Deming Chen, Jinjun Xiong, and Zehra Sura
    2017 IEEE International Conference on Rebooting Computing (ICRC) , 2017
  12. Voltage regulator control for improved computing power efficiency
    Nam Sung Kim
    2017
  13. Resource and core scaling for improving performance of power-constrained multi-core processors
    Nam Sung Kim
    2017
  14. Understanding System Characteristics of Online Erasure Coding on Scalable Distributed and Large-Scale SSD Array Systems
    Sungjoon Koh, Jie Zhang, Miryeong Kwon, Jungyeon Yoonz, David Donofrioy, Nam Sung Kim, and Myoungsoo Jung
    2017 IEEE International Symposium on Workload Characterization (IISWC) , 2017
  15. Multiplication circuit providing dynamic truncation
    Srinivasan Narayanamoorthy and Nam Sung Kim
    2017
  16. Understanding power-performance relationship of energy-efficient modern DRAM devices
    Sukhan Lee, Yuhwan Ro, Young Hoon Son, Hyunyoon Cho, Nam Sung Kim, and Jung Ho Ahn
    2017 IEEE International Symposium on Workload Characterization (IISWC) , 2017
  17. Temporal codes in on-chip interconnects
    Michael Mishkin, Nam Sung Kim, and Mikko Lipasti
    2017 IEEE/ACM International Symposium on Low Power Electronics and Design (ISLPED) , 2017
  18. Memory controller for heterogeneous computer
    Hao Wang and Nam Sung Kim
    2017
  19. CNFET-based high throughput SIMD architecture
    Li Jiang, Tianjian Li, Naifeng Jing, Nam Sung Kim, Minyi Guo, and Xiaoyao Liang
    IEEE Transactions on Computer-Aided Design of Integrated Circuits and Systems , 2017
  20. Energy‐Efficient Approximate Speech Signal Processing for Wearable Devices
    Taejoon Park, Kyoosik Shin, and Nam Sung Kim
    ETRI Journal , 2017
  21. Pageforge
    Dimitrios Skarlatos, Nam Sung Kim, and Josep Torrellas
    Proceedings of the 50th Annual IEEE/ACM International Symposium on Microarchitecture , 2017
  22. Janus: Supporting heterogeneous power management in virtualized environments
    Daehoon Kim, Mohammad Alian, Jaehyuk Huh, and Nam Sung Kim
    2017
  23. Method and apparatus for power reduction during lane divergence
    Nam Sung Kim, James M O’connor, Michael J Schulte, and Vijay Janapa Reddi
    2017
  24. Memory fault patching using pre-existing memory structures
    David John Palframan, Nam Sung Kim, and Mikko Lipasti
    2017
  25. Energy-efficient approximate audio signal processing engine for wearable devices
    Electronics and Telecommunications Research Institute (ETRI) Journal , 2017
  26. Sequential circuit with error detection
    Keith A Bowman, James W Tschanz, Nam Sung Kim, Janice C Lee, Christopher B Wilkerson, Shih-Lien L Lu, Tanay Karnik, and Vivek K De
    2017
  27. Apparatus and method for adjusting bandwidth
    Ho-Young Kim, Nam-Sung Kim, and Daniel W Chang
    2017

2016

  1. Chameleon: Versatile and practical near-DRAM acceleration architecture for large memory systems
    Hadi Asghari-Moghaddam, Young Hoon Son, Jung Ho Ahn, and Nam Sung Kim
    2016 49th annual IEEE/ACM international symposium on Microarchitecture (MICRO) , 2016
  2. Approximating warps with intra-warp operand value similarity
    Daniel Wong, Nam Sung Kim, and Murali Annavaram
    2016 IEEE International Symposium on High Performance Computer Architecture (HPCA) , 2016
  3. Near-DRAM acceleration with single-ISA heterogeneous processing in standard memory modules
    Hadi Asghari-Moghaddam, Amin Farmahini-Farahani, Katherine Morrow, Jung Ho Ahn, and Nam Sung Kim
    IEEE Micro , 2016
  4. On effective and efficient quality management for approximate computing
    Ting Wang, Qian Zhang, Nam Sung Kim, and Qiang Xu
    2016
  5. ScalCore: Designing a core for voltage scalability
    Bhargava Gopireddy, Choungki Song, Josep Torrellas, Nam Sung Kim, Aditya Agrawal, and Asit Mishra
    2016 IEEE International Symposium on High Performance Computer Architecture (HPCA) , 2016
  6. VARIUS-TC: A modular architecture-level model of parametric variation for thin-channel switches
    S Karen Khatamifard, Michael Resch, Nam Sung Kim, and Ulya R Karpuzcu
    2016 IEEE 34th International Conference on Computer Design (ICCD) , 2016
  7. DUANG: Fast and lightweight page migration in asymmetric memory systems
    Hao Wang, Jie Zhang, Sharmila Shridhar, Gieseo Park, Myoungsoo Jung, and Nam Sung Kim
    2016 IEEE International Symposium on High Performance Computer Architecture (HPCA) , 2016
  8. Exploring new features of high-bandwidth memory for gpus
    Bingchao Li, Choungki Song, Jizeng Wei, Jung Ho Ahn, and Nam Sung Kim
    IEICE Electronics Express , 2016
  9. Multiplier circuit with dynamic energy consumption adjustment
    Nam Sung Kim
    2016
  10. Dynamic error handling for on-chip memory structures
    Nam Sung Kim
    2016
  11. Memory controller for heterogeneous computer
    Hao Wang and Nam Sung Kim
    2016
  12. Method and apparatus for soft error mitigation in computers
    David John Palframan, Nam Sung Kim, and Mikko Lipasti
    2016
  13. High efficiency computer floating point multiplier unit
    Nam Sung Kim, Syed Gilani, and Michael Schulte
    2016
  14. Snatch: Opportunistically reassigning power allocation between processor and memory in 3D stacks
    Dimitrios Skarlatos, Renji Thomas, Aditya Agrawal, Shibin Qin, Robert Pilawa-Podgurski, Ulya R Karpuzcu, Radu Teodorescu, Nam Sung Kim, and Josep Torrellas
    2016 49th Annual IEEE/ACM International Symposium on Microarchitecture (MICRO) , 2016
  15. Spinwise: A practical energy-efficient synchronization technique for cmps
    Hadi Asgharimoghaddam and Nam Sung Kim
    ACM SIGARCH Computer Architecture News , 2016
  16. VR-scale: Runtime dynamic phase scaling of processor voltage regulators for improving power efficiency
    Hadi Asghari-Moghaddam, Hamid Reza Ghasemi, Abhishek Arvind Sinkar, Indrani Paul, and Nam Sung Kim
    2016
  17. Fine-Grained Task Migration for Graph Algorithms using Processing in Memory
    Paula Aguilera, Dong Ping Zhang, Nam Sung Kim, and Nuwan Jayasena
    2016 IEEE International Parallel and Distributed Processing Symposium Workshops (IPDPSW) , 2016
  18. CNFET-based high throughput register file architecture
    Tianjian Li, Li Jiang, Naifeng Jing, Nam Sung Kim, and Xiaoyao Liang
    2016 IEEE 34th International Conference on Computer Design (ICCD) , 2016
  19. Write-after-read hazard prevention in GPGPUSIM
    Michael Mishkin, Nam Sung Kim, and Mikko Lipasti
    Workshop on Deplicating, Deconstructing, and Debunking (WDDD) , 2016
  20. Bit Serializing a Microprocessor for Ultra-low-power
    Matthew Tomei, Henry Duwe, Nam Sung Kim, and Rakesh Kumar
    2016
  21. Guest Editors’ Introduction: Approximate Computing
    Qiang Xu, Todd Mytkowicz, and Nam Sung Kim
    IEEE Design & Test , 2016
  22. pd-gem5: Simulation Infrastructure for Parallel/Distributed Computer Systems
    Mohammad Alian, Daehoon Kim, Nam Sung Kim, Y Kim, W Yang, O Mutlu, LE Olson, S Sethumadhavan, MD Hill, B Jacob, M Kleanthous, Y Sazeides, E Özer, C Nicopoulos, P Nikolaou, Z Hadjilambrou, BK Daya, LS Peh, and AP Chandrakasan
    no. IEEE , 2016
  23. Energy-efficient multicore processor architecture for parallel processing
    Nam-Sung Kim
    2016
  24. Signal processing circuit with multiple power modes
    Nam Sung Kim
    2016
  25. Heterogeneous Computing–A Path to Post-Moore Supercomputing: Architecture Circuits and Process
    Wayne Burleson, Shomit Das, Yasuko Eckert, and Nam Sung Kim
    International Workshop on Post Moore’s Era Supercomputing , 2016

2015

  1. Approximate computing: A survey
    Qiang Xu, Todd Mytkowicz, and Nam Sung Kim
    IEEE Design & Test , 2015
  2. NDA: Near-DRAM acceleration architecture leveraging commodity DRAM devices and standard memory modules
    Amin Farmahini-Farahani, Jung Ho Ahn, Katherine Morrow, and Nam Sung Kim
    2015 IEEE 21st International Symposium on High Performance Computer Architecture (HPCA) , 2015
  3. GPU register file virtualization
    Hyeran Jeon, Gokul Subramanian Ravi, Nam Sung Kim, and Murali Annavaram
    2015
  4. CiDRA: A cache-inspired DRAM resilience architecture
    Young Hoon Son, Sukhan Lee, O Seongil, Sanghyuk Kwon, Nam Sung Kim, and Jung Ho Ahn
    2015 IEEE 21st International Symposium on High Performance Computer Architecture (HPCA) , 2015
  5. COP: To compress and protect main memory
    David J Palframan, Nam Sung Kim, and Mikko H Lipasti
    ACM SIGARCH Computer Architecture News , 2015
  6. iPatch: Intelligent fault patching to improve energy efficiency
    David J Palframan, Nam Sung Kim, and Mikko H Lipasti
    2015 IEEE 21st International Symposium on High Performance Computer Architecture (HPCA) , 2015
  7. Ultra‐low‐power image signal processor for smart camera applications
    Zhenhong Liu, T Park, HS Park, and NS Kim
    Electronics Letters , 2015
  8. Decoupled control and data processing for approximate near-threshold voltage computing
    Ismail Akturk, Nam Sung Kim, and Ulya R Karpuzcu
    IEEE Micro , 2015
  9. ATC
    Bolt: Faster reconfiguration in operating systems
    Sankaralingam Panneerselvam, Michael Swift, and Nam Sung Kim
    2015 USENIX Annual Technical Conference (USENIX ATC 15) , 2015
  10. vCache: Architectural support for transparent and isolated virtual LLCs in virtualized environments
    Daehoon Kim, Hwanju Kim, Nam Sung Kim, and Jaehyuk Huh
    2015
  11. Workload-aware optimal power allocation on single-chip heterogeneous processors
    Jae Young Jang, Hao Wang, Euijin Kwon, Jae W Lee, and Nam Sung Kim
    IEEE Transactions on Parallel and Distributed Systems , 2015
  12. Alloy: Parallel-serial memory channel architecture for single-chip heterogeneous processor systems
    Hao Wang, Chang-Jae Park, Gyung-su Byun, Jung Ho Ahn, and Nam Sung Kim
    2015 IEEE 21st International Symposium on High Performance Computer Architecture (HPCA) , 2015
  13. Memory-link compression for graphic processor unit
    Nam Sung Kim
    2015
  14. Comparison of single-isa heterogeneous versus wide dynamic range processors for mobile applications
    Hamid Reza Ghasemi, Ulya R Karpuzcu, and Nam Sung Kim
    2015 33rd IEEE International Conference on Computer Design (ICCD) , 2015
  15. NDA: Near-DRAM acceleration architecture leveraging commodity DRAM devices and standard memory modules
    Amin Farmahini Farahani, Jung Ho Ahn, Katherine Morrow, and Nam Sung Kim
    HPCA , 2015
  16. Joint optimisation of computational accuracy and algorithm parameters for energy‐efficient recognition algorithms
    Heesung Lim, Taejoon Park, and Nam Sung Kim
    Electronics Letters , 2015
  17. Online and Operand-Aware Detection of Failures Utilizing False Alarm Vectors
    Amir Yazdanbakhsh, David Palframan, Azadeh Davoodi, Nam Sung Kim, and Mikko Lipasti
    2015

2014

  1. Energy-efficient approximate multiplication for digital signal processing and classification applications
    Srinivasan Narayanamoorthy, Hadi Asghari Moghaddam, Zhenhong Liu, Taejoon Park, and Nam Sung Kim
    IEEE transactions on very large scale integration (VLSI) systems , 2014
  2. Sleepscale: Runtime joint speed scaling and sleep states management for power efficient data centers
    Yanpei Liu, Stark C Draper, and Nam Sung Kim
    ACM SIGARCH Computer Architecture News , 2014
  3. CAL
    DRAMA: An architecture for accelerated processing near memory
    Amin Farmahini-Farahani, Jung Ho Ahn, Katherine Morrow, and Nam Sung Kim
    IEEE Computer Architecture Letters , 2014
  4. QoS-aware dynamic resource allocation for spatial-multitasking GPUs
    Paula Aguilera, Katherine Morrow, and Nam Sung Kim
    2014 19th Asia and South Pacific Design Automation Conference (ASP-DAC) , 2014
  5. Fair share: Allocation of GPU resources for both performance and fairness
    Paula Aguilera, Katherine Morrow, and Nam Sung Kim
    2014 IEEE 32nd International Conference on Computer Design (ICCD) , 2014
  6. Row-buffer decoupling: A case for low-latency DRAM microarchitecture
    O Seongil, Young Hoon Son, Nam Sung Kim, and Jung Ho Ahn
    2014 ACM/IEEE 41st International Symposium on Computer Architecture (ISCA) , 2014
  7. Process variation-aware workload partitioning algorithms for GPUs supporting spatial-multitasking
    Paula Aguilera, Jungseob Lee, Amin Farmahini-Farahani, Katherine Morrow, Michael Schulte, and Nam Sung Kim
    2014 Design, Automation & Test in Europe Conference & Exhibition (DATE) , 2014
  8. Precision-aware soft error protection for GPUs
    David J Palframan, Nam Sung Kim, and Mikko H Lipasti
    2014 IEEE 20th International Symposium on High Performance Computer Architecture (HPCA) , 2014
  9. Accordion: Toward soft near-threshold voltage computing
    Ulya R Karpuzcu, Ismail Akturk, and Nam Sung Kim
    2014 IEEE 20th International Symposium on High Performance Computer Architecture (HPCA) , 2014
  10. Memory scheduling towards high-throughput cooperative heterogeneous computing
    Hao Wang, Ripudaman Singh, Michael J Schulte, and Nam Sung Kim
    2014
  11. RCS: Runtime resource and core scaling for power-constrained multi-core processors
    Hamid Reza Ghasemi and Nam Sung Kim
    2014
  12. Leakage power management using programmable power gating transistors and on-chip aging and temperature tracking circuit
    Nam Sung Kim
    2014
  13. Optimization of a cell counting algorithm for mobile point-of-care testing platforms
    DaeHan Ahn, Nam Sung Kim, SangJun Moon, Taejoon Park, and Sang Hyuk Son
    Sensors , 2014
  14. Energy-efficient reconfigurable cache architectures for accelerator-enabled embedded systems
    Amin Farmahini-Farahani, Nam Sung Kim, and Katherine Morrow
    2014 IEEE International Symposium on Performance Analysis of Systems and Software (ISPASS) , 2014
  15. Maximizing throughput of power/thermal-constrained processors by balancing power consumption of cores
    Abhishek A Sinkar, Hao Wang, and Nam Sung Kim
    Fifteenth International Symposium on Quality Electronic Design , 2014
  16. Multiplier supporting accuracy and energy trade‐offs for recognition applications
    Nam Sung Kim, Taejoon Park, Srinivasan Narayanamoorthy, and Hadi Asgharimoghaddam
    Electronics letters , 2014
  17. Memory cell supply voltage control based on error detection
    Muhammad Khellah, Dinesh Somasekhar, Yibin Ye, Nam Sung Kim, and Vivek De
    2014
  18. Sleepscale: Runtime joint speed scaling and sleep states management for power efficient data centers. In 2014 ACM/IEEE 41st International Symposium on Computer Architecture (ISCA)
    Yanpei Liu, Stark C Draper, and Nam Sung Kim
    2014
  19. Energy-efficient pixel-arithmetic
    Syed Zohaib Gilani, Nam Sung Kim, and Michael Schulte
    IEEE Transactions on Computers , 2014
  20. Quantitative comparison of the power reduction techniques for samsung reconfigurable processor
    Hoyoung Kim, Soojung Ryu, Abhishek Sinkar, and Nam Sung Kim
    2014 IEEE International Symposium on Circuits and Systems (ISCAS) , 2014
  21. Low-cost scratchpad memory organizations using heterogeneous cell sizes for low-voltage operations
    Syed Gilani, Taejoon Park, and Nam Sung Kim
    Microprocessors and Microsystems , 2014
  22. Energy efficient processor having heterogeneous cache
    Nam Sung Kim and Stark C Draper
    2014

2013

  1. GPUWattch: Enabling energy optimizations in GPGPUs
    Jingwen Leng, Tayler Hetherington, Ahmed ElTantawy, Syed Gilani, Nam Sung Kim, Tor M Aamodt, and Vijay Janapa Reddi
    International Symposium on Computer Architecture , 2013
  2. Energysmart: Toward energy-efficient manycores for near-threshold computing
    Ulya R Karpuzcu, Abhishek Sinkar, Nam Sung Kim, and Josep Torrellas
    2013 IEEE 19th International Symposium on High Performance Computer Architecture (HPCA) , 2013
  3. Coping with parametric variation at near-threshold voltages
    Ulya R Karpuzcu, Nam Sung Kim, and Josep Torrellas
    IEEE Micro , 2013
  4. Reevaluating the latency claims of 3D stacked memories
    Daniel W Chang, Gyungsu Byun, Hoyoung Kim, Minwook Ahn, Soojung Ryu, Nam S Kim, and Michael Schulte
    2013 18th Asia and South Pacific Design Automation Conference (ASP-DAC) , 2013
  5. Exploiting GPU peak-power and performance tradeoffs through reduced effective pipeline latency
    Syed Zohaib Gilani, Nam Sung Kim, and Michael J Schulte
    2013
  6. Improving throughput of power-constrained many-core processors based on unreliable devices
    Hao Wang and Nam Sung Kim
    IEEE Micro , 2013
  7. CAD
    Improving platform energy-chip area trade-off in near-threshold computing environment
    Hao Wang, Abhishek A Sinkar, and Nam Sung Kim
    2013 IEEE/ACM International Conference on Computer-Aided Design (ICCAD) , 2013
  8. Resilient high-performance processors with spare ribs
    David J Palframan, Nam Sung Kim, and Mikko H Lipasti
    IEEE Micro , 2013
  9. Queuing theoretic analysis of power-performance tradeoff in power-efficient computing
    Yanpei Liu, Stark C Draper, and Nam Sung Kim
    2013 47th Annual Conference on Information Sciences and Systems (CISS) , 2013
  10. CAD
    Dynamic bandwidth scaling for embedded DSPs with 3D-stacked DRAM and wide I/Os
    Daniel W Chang, Young Hoon Son, Jung Ho Ahn, Hoyoung Kim, Minwook Ahn, Michael J Schulte, and Nam Sung Kim
    2013 IEEE/ACM International Conference on Computer-Aided Design (ICCAD) , 2013
  11. REEL: Reducing effective execution latency of floating point operations
    Vignyan Reddy, Syed Zohaib Gilani, Erika Gunadi, Nam Sung Kim, Michael J Schulte, and Mikko H Lipasti
    International Symposium on Low Power Electronics and Design (ISLPED) , 2013
  12. IMPROVING MEMORY RELIABILITY POWER AND PERFORMANCE USING MIXED-CELL DESIGNS.
    Alaa R Alameldeen, Nam Sung Kim, Samira M Khan, Hamid Reza Ghasemi, Chris Wilkerson, Jaydeep Kulkarni, and Daniel A Jiménez
    Intel Technology Journal , 2013
  13. Optimal Power Allocation for Multiprogrammed Workloads on Single-Chip Heterogeneous Processors
    Euijin Kwon, Jae Young Jang, Jae W Lee, and Nam Sung Kim
    2013 , 2013

2012

  1. The case for GPGPU spatial multitasking
    Jacob T Adriaens, Katherine Compton, Nam Sung Kim, and Michael J Schulte
    IEEE International Symposium on High-Performance Comp Architecture , 2012
  2. Lossless and lossy memory I/O link compression for improving performance of GPGPU workloads
    Vijay Sathish, Michael J Schulte, and Nam Sung Kim
    2012
  3. VARIUS-NTV: A microarchitectural model to capture the increased sensitivity of manycores to process variations at near-threshold voltages
    Ulya R Karpuzcu, Krishna B Kolluru, Nam Sung Kim, and Josep Torrellas
    IEEE/IFIP International Conference on Dependable Systems and Networks (DSN 2012) , 2012
  4. Power-efficient computing for compute-intensive GPGPU applications
    Syed Zohaib Gilani, Nam Sung Kim, and Michael J Schulte
    2012
  5. Workload and power budget partitioning for single-chip heterogeneous processors
    Hao Wang, Vijay Sathish, Ripudaman Singh, Michael J Schulte, and Nam Sung Kim
    2012
  6. Cost-effective power delivery to support per-core voltage domains for power-constrained processors
    Hamid Reza Ghasemi, Abhishek A Sinkar, Michael J Schulte, and Nam Sung Kim
    2012
  7. Workload-aware voltage regulator optimization for power efficient multi-core processors
    Abhishek A Sinkar, Hao Wang, and Nam Sung Kim
    2012 Design, Automation & Test in Europe Conference & Exhibition (DATE) , 2012
  8. A linear algebra core design for efficient level-3 blas
    Ardavan Pedram, Syed Zohaib Gilani, Nam Sung Kim, Robert Van De Geijn, Michael Schulte, and Andreas Gerstlauer
    2012 IEEE 23rd International Conference on Application-Specific Systems, Architectures and Processors , 2012
  9. Method and apparatus for optimizing clock speed and power dissipation in multicore architectures
    Nam Sung Kim
    2012
  10. Sequential circuit with error detection
    Keith Bowman, James Tachanz, Nam Sung Kim, Janice Lee, Chris Wilkerson, Shih-Lien L Lu, Tanay Karnlk, and Vivek De
    2012
  11. Virtual floating-point units for low-power embedded processors
    Syed Zohaib Gilani, Nam Sung Kim, and Michael Schulte
    2012 IEEE 23rd International Conference on Application-Specific Systems, Architectures and Processors , 2012
  12. Mitigating random variation with spare RIBs: Redundant intermediate bitslices
    David J Palframan, Nam Sung Kim, and Mikko H Lipasti
    IEEE/IFIP International Conference on Dependable Systems and Networks (DSN 2012) , 2012
  13. Parameter variation at near threshold voltage: The power efficiency versus resilience tradeoff
    Josep Torrellas, Nam Sung Kim, and Radu Teodorescu
    University of Illinois, Tech. Rep , 2012
  14. Clamping virtual supply voltage of power-gated circuits for active leakage reduction and gate-oxide reliability
    Abhishek Sinkar, Taejoon Park, and Nam Sung Kim
    IEEE transactions on very large scale integration (VLSI) systems , 2012

2011

  1. Improving throughput of power-constrained GPUs using dynamic voltage/frequency and core scaling
    Jungseob Lee, Vijay Sathisha, Michael Schulte, Katherine Compton, and Nam Sung Kim
    2011 International Conference on Parallel Architectures and Compilation Techniques , 2011
  2. Analyzing the impact of joint optimization of cell size redundancy and ECC on low-voltage SRAM array total area
    Nam Sung Kim, Stark C Draper, Shi-Ting Zhou, Sumeet Katariya, Hamid Reza Ghasemi, and Taejoon Park
    IEEE Transactions on Very Large Scale Integration (VLSI) Systems , 2011
  3. Low-voltage on-chip cache architecture using heterogeneous cell sizes for high-performance processors
    Hamid Reza Ghasemi, Stark C Draper, and Nam Sung Kim
    2011 IEEE 17th International Symposium on High Performance Computer Architecture , 2011
  4. Analyzing throughput of GPGPUs exploiting within-die core-to-core frequency variation
    Jungseob Lee, Paritosh Pratap Ajgaonkar, and Nam Sung Kim
    (IEEE ISPASS) IEEE International Symposium on Performance Analysis of Systems and Software , 2011
  5. Analyzing potential throughput improvement of power-and thermal-constrained multicore processors by exploiting DVFS and PCPG
    Jungseob Lee and Nam Sung Kim
    IEEE transactions on very large scale integration (VLSI) systems , 2011
  6. Scratchpad memory optimizations for digital signal processing applications
    Syed Z Gilani, Nam Sung Kim, and Michael Schulte
    2011 Design, Automation & Test in Europe , 2011
  7. Energy-efficient floating-point arithmetic for software-defined radio architectures
    Syed Zohaib Gilani, Nam Sung Kim, and Michael Schulte
    ASAP 2011-22nd IEEE International Conference on Application-specific Systems, Architectures and Processors , 2011
  8. Time redundant parity for low-cost transient error detection
    David J Palframan, Nam Sung Kim, and Mikko H Lipasti
    2011 Design, Automation & Test in Europe , 2011
  9. Memory cell supply voltage control based on error detection
    Khellah Muhammad, Dinesh Somasekhar, Yibin Ye, Nam Sung Kim, and Vivek De
    2011
  10. Energy-efficient floating-point arithmetic for digital signal processors
    Syed Zohaib Gilani, Nam Sung Kim, and Michael Schulte
    2011 Conference Record of the Forty Fifth Asilomar Conference on Signals, Systems and Computers (ASILOMAR) , 2011
  11. A low cost approach to calibrate on-chip thermal sensors
    Krishna Bharath, Chunhua Yao, Nam Sung Kim, Parameswaran Ramanathan, and Kewal K Saluja
    2011 12th International Symposium on Quality Electronic Design , 2011
  12. Analyzing the performance and energy impact of 3D memory integration on embedded DSPs
    Daniel W Chang, Nam S Kim, and Michael J Schulte
    2011 International Conference on Embedded Computer Systems: Architectures, Modeling and Simulation , 2011
  13. Maximizing frequency and yield of power-constrained designs using programmable power-gating
    Nam Sung Kim, Abhishek Sinkar, Jun Seomun, and Youngsoo Shin
    IEEE transactions on very large scale integration (VLSI) systems , 2011
  14. AVS-aware power-gate sizing for maximum performance and power efficiency of power-constrained processors
    Abhishek Sinkar and Nam Sung Kim
    16th Asia and South Pacific Design Automation Conference (ASP-DAC 2011) , 2011
  15. A Microarchitectural Model of Process Variation for Near-threshold Computing
    Ulya R Karpuzcu, Krishna Kolluru, Nam Sung Kim, and Josep Torrellas
    2011
  16. Spare RIBs: Redundant Intermediate Bitslices
    David J Palframan, Nam Sung Kim, and Mikko H Lipasti
    2011

2010

  1. Combating aging with the colt duty cycle equalizer
    Erika Gunadi, Abhisek A Sinkar, Nam Sung Kim, and Mikko H Lipasti
    2010 43rd Annual IEEE/ACM International Symposium on Microarchitecture , 2010
  2. Delay fault detection using latch with error sampling
    James W Tschanz, Keith A Bowman, Nam Sung Kim, Chris Wilkerson, Shih-Lien L Lu, and Tanay Karnik
    2010
  3. Runtime temperature-based power estimation for optimizing throughput of thermal-constrained multi-core processors
    Dongkeun Oh, Nam Sung Kim, Charlie Chung Ping Chen, Azadeh Davoodi, and Yu Hen Hu
    2010 15th Asia and South Pacific Design Automation Conference (ASP-DAC) , 2010
  4. CAD
    Optimal algorithm for profile-based power gating: A compiler technique for reducing leakage on execution units in microprocessors
    Danbee Park, Jungseob Lee, Nam Sung Kim, and Taewhan Kim
    2010 IEEE/ACM International Conference on Computer-Aided Design (ICCAD) , 2010
  5. Analyzing and minimizing effects of temperature variation and NBTI on active leakage power of power-gated circuits
    Abhishek Sinkar and Nam Sung Kim
    2010 11th International Symposium on Quality Electronic Design (ISQED) , 2010
  6. Memory cell bit valve loss detection and restoration
    Nam Sung Kim, Muhammad Kheliah, Yibin Ye, Dinesh Somasekhar, and Vivek De
    2010
  7. Workload-adaptive process tuning strategy for power-efficient multi-core processors
    Jungseob Lee, Chi-Chao Wang, Hamid Ghasemil, Lloyd Bircher, Yu Cao, and Nam Sung Kim
    2010
  8. Sleep transistor array apparatus and method with leakage control circuitry
    Nam Sung Kim and Vivek De
    2010
  9. The compatibility analysis of thread migration and DVFS in multi-core processor
    Dongkeun Oh, Charlie Chung Ping Chen, NamSung Kim, and Yu Hen Hu
    2010 11th International Symposium on Quality Electronic Design (ISQED) , 2010
  10. Analyzing impact of multiple ABB and AVS domains on throughput of power and thermal-constrained multi-core processors
    Jungseob Lee, Shi-Ting Zhou, and Nam Sung Kim
    2010 15th Asia and South Pacific Design Automation Conference (ASP-DAC) , 2010

2009

  1. Optimizing throughput of power-and thermal-constrained multicore processors using DVFS and per-core power-gating
    Jungseob Lee and Nam Sung Kim
    Design Automation Conference, 2009. DAC’09. 46th ACM/IEEE , 2009
  2. Process Temperature and Supply-Noise Tolerant 45< formula formulatype=
    Muhammad Khellah, Nam Sung Kim, Yibin Ye, Dinesh Somasekhar, Tanay Karnik, Nitin Borkar, Gunjan Pandya, Fatih Hamzaoglu, Tom Coan, Yih Wang, Kevin Zhang, Clair Webb, and Vivek De
    Solid-State Circuits, IEEE Journal of , 2009
  3. Optimizing total power of many-core processors considering voltage scaling limit and process variations
    Jungseob Lee and Nam Sung Kim
    2009
  4. Memory having bit line with resistor (s) between memory cells
    Muhammad M Khellah, Dinesh Somasekhar, Yibin Ye, Nam Sung Kim, and Vivek De
    2009
  5. Frequency and yield optimization using power gates in power-constrained designs
    Nam Sung Kim, Jun Seomun, Abhishek Sinkar, Jungseob Lee, Tae Hee Han, Ken Choi, and Youngsoo Shin
    2009
  6. SRAM dynamic stability estimation using MPFP and its applications
    DiaaEldin Khalil, Muhammad Khellah, Nam-Sung Kim, Yehea Ismail, Tanay Karnik, and Vivek De
    Microelectronics journal , 2009
  7. Analyzing potential power reduction with adaptive voltage positioning optimized for multicore processors
    Abhishek Sinkar and Nam Sung Kim
    2009
  8. Method and apparatus improving performance of a digital memory array device
    Min Huang, Chris Wilkerson, Nam Sung Kim, and Moinuddin K Qureshi
    2009
  9. Statistical static timing analysis considering leakage variability in power gated designs
    Michael J Anderson, Azadeh Davoodi, Jungseob Lee, Abhishek Sinkar, and Nam Sung Kim
    2009
  10. Sense amplifier method and arrangement
    Dinesh Somasekhar, Muhammad M Khellah, Yibin Ye, Nam Sung Kim, and K De Vivek
    2009

2008

  1. Energy-efficient and metastability-immune resilient circuits for dynamic variation tolerance
    Keith A Bowman, James W Tschanz, Nam Sung Kim, Janice C Lee, Chris B Wilkerson, Shih-Lien L Lu, Tanay Karnik, and Vivek K De
    IEEE Journal of Solid-State Circuits , 2008
  2. Accurate estimation of SRAM dynamic stability
    DiaaEldin Khalil, Muhammad Khellah, Nam-Sung Kim, Yehea Ismail, Tanay Karnik, and Vivek K De
    IEEE transactions on very large scale integration (VLSI) systems , 2008
  3. Energy-efficient and metastability-immune timing-error detection and instruction-replay-based recovery circuits for dynamic-variation tolerance
    Keith A Bowman, James W Tschanz, Nam Sung Kim, Janice C Lee, Chris B Wilkerson, Shih-Lien L Lu, Tanay Karnik, and Vivek K De
    2008 IEEE International Solid-State Circuits Conference-Digest of Technical Papers , 2008
  4. On-chip cache device scaling limits and effective fault repair techniques in future nanoscale technology
    David Roberts, Nam Sung Kim, and Trevor Mudge
    Microprocessors and Microsystems , 2008
  5. Memory with dynamically adjustable supply
    Fatih Hamzaoglu, Kevin Zhang, Nam Sung Kim, Muhammad M Khellah, Dinesh Somasekhar, Yibin Ye, Vivek K De, and Bo Zheng
    2008
  6. Energy-efficient and metastability-immune timing-error detection and recovery circuits for dynamic variation tolerance
    Keith A Bowman, James W Tschanz, Nam Sung Kim, Janice C Lee, Chris B Wilkerson, Shih-Lien L Lu, Tanay Karnik, and Vivek K De
    2008 IEEE international conference on integrated circuit design and technology and tutorial , 2008
  7. Address hashing to help distribute accesses across portions of destructive read cache memory
    Nam Sung Kim, Muhammad M Khellah, and Vivek De
    2008
  8. Memory driver circuits with embedded level shifters
    Muhammad M Khellah, Dinesh Somasekhar, Yibin Ye, Nam Sung Kim, and Vivek K De
    2008
  9. Analog phase control circuit and method
    Nam Sung Kim and Vivek De
    2008

2007

  1. Adaptive frequency and biasing techniques for tolerance to dynamic temperature-voltage variations and aging
    James Tschanz, Nam Sung Kim, Saurabh Dighe, Jason Howard, Gregory Ruhl, Sriram Vangal, Siva Narendra, Yatin Hoskote, Howard Wilson, Carol Lam, Matthew Shuman, Carlos Tokunaga, Dinesh Somasekhar, Stephen Tang, David Finan, Tanay Karnik, Nitin Borkar, Nasser Kurd, and Vivek De
    2007 IEEE International Solid-State Circuits Conference. Digest of Technical Papers , 2007
  2. CAD
    Yield-driven near-threshold SRAM design
    Gregory Chen, Dennis Sylvester, David Blaauw, Trevor Mudge, and Nam Sung Kim
    Proceedings of the 2007 IEEE/ACM international conference on Computer-aided design , 2007
  3. Memory cell having p-type pass device
    Muhammad M Khellah, Dinesh Somasekhar, Nam Sung Kim, Yibin Ye, Vivek K De, Kevin Zhang, and Bo Zheng
    2007
  4. Reducing aging effect on memory
    Nam Sung Kim, Shih-Lien L Lu, Chris Wilkerson, and Edward Grochowski
    2007
  5. SRAM dynamic stability estimation using MPFP
    DiaaEldin Khalil, Muhammad Khellah, Nam-Sung Kim, Yehea Ismail, Tanay Karnik, and Vivek De
    2007 Internatonal Conference on Microelectronics , 2007

2006

  1. Wordline & bitline pulsing schemes for improving SRAM cell stability in low-Vcc 65nm CMOS designs
    Muhammad Khellah, Yibin Ye, N Kim, Dinesh Somasekhar, Gunjan Pandya, Ali Farhang, Kevin Zhang, Clair Webb, and Vivek De
    2006 Symposium on VLSI Circuits, 2006. Digest of Technical Papers. , 2006
  2. A 256-Kb Dual- SRAM Building Block in 65-nm CMOS Process With Actively Clamped Sleep Transistor
    Muhammad Khellah, Dinesh Somasekhar, Yibin Ye, Nam Sung Kim, Jason Howard, Greg Ruhl, Murad Sunna, James Tschanz, Nitin Borkar, Fatih Hamzaoglu, Gunjan Pandya, Ali Farhang, Kevin Zhang, and Vivek De
    IEEE Journal of Solid-State Circuits , 2006
  3. Data processor memory circuit
    Krisztian Flautner, David T Blaauw, Trevor N Mudge, Nam S Kim, and Steven M Martin
    2006
  4. A 4.2 GHz 0.3 mm2 256kb Dual-V/sub cc/SRAM Building Block in 65nm CMOS
    M. Khellah, Nam Sung Kim, J. Howard, G. Ruhl, Yibin Ye, J. Tschanz, D. Somasekhar, N. Borkar, F. Hamzaoglu, G. Pandya, A. Farhang, K. Zhang, and V. De
    2006 IEEE International Solid State Circuits Conference-Digest of Technical Papers , 2006
  5. Session 8: leakage power analysis and optimization
    Nam Sung Kim, Naehyuck Chang, and Sanu Mathew
    Annual ACM IEEE Design Automation Conference: Proceedings of the 43 rd annual conference on Design automation , 2006

2005

  1. Quantitative analysis and optimization techniques for on-chip cache leakage power
    Nam Sung Kim, David Blaauw, and Trevor Mudge
    Very Large Scale Integration (VLSI) Systems, IEEE Transactions on , 2005
  2. CAD
    Total power-optimal pipelining and parallel processing under process variations in nanometer technology
    Nam Sung Kim, Taeho Kgil, Keith Bowman, Vivek De, and Trevor Mudge
    ICCAD-2005. IEEE/ACM International Conference on Computer-Aided Design, 2005. , 2005
  3. Total leakage optimization strategies for multi-level caches
    Robert Bai, Nam-Sung Kim, Dennis Sylvester, and Trevor Mudge
    2005
  4. Power-performance trade-offs in nanometer-scale multi-level caches considering total leakage
    Robert Bai, Nam-Sung Kim, Tae Ho Kgil, Dennis Sylvester, and Trevor Mudge
    Design, Automation and Test in Europe , 2005
  5. Erratum: Circuit and microarchitectural techniques for reducing cache leakage power (IEEE Transactions on Very Large Scale Integration (VLSI) Systems (Feb. 2004) 12: 2 (167-184))
    Nam Sung Kim, K Flautner, D Blaauw, and T Mudge
    IEEE Transactions on Very Large Scale Integration (VLSI) Systems , 2005

2004

  1. Razor: circuit-level correction of timing errors for low-power operation
    Dan Ernst, Shidhartha Das, Seokwoo Lee, David Blaauw, Todd Austin, Trevor Mudge, Nam Sung Kim, and Krisztián Flautner
    IEEE Micro , 2004
  2. Circuit and microarchitectural techniques for reducing cache leakage power
    Nam Sung Kim, Krisztian Flautner, David Blaauw, and Trevor Mudge
    IEEE Transactions on Very Large Scale Integration (VLSI) Systems , 2004
  3. Single-vDD and single-vT super-drowsy techniques for low-leakage high-performance instruction caches
    Nam Sung Kim, Krisztián Flautner, David Blaauw, and Trevor Mudge
    2004
  4. Microarchitectural power modeling techniques for deep sub-micron microprocessors
    Nam Sung Kim, Taeho Kgil, Valeria Bertacco, Todd Austin, and Trevor Mudge
    2004
  5. Shidhartha Das
    Dan Ernst and Nam Sung Kim
    Sanjay Pant, Rajeev Rao, Toan Pham , 2004
  6. Circuit and microarchitectural techniques for processor on-chip cache leakage power reduction
    Nam Sung Kim
    2004

2003

  1. Razor: A low-power pipeline based on circuit-level timing speculation
    Dan Ernst, Nam Sung Kim, Shidhartha Das, Sanjay Pant, Rajeev Rao, Toan Pham, Conrad Ziesler, David Blaauw, Todd Austin, Krisztian Flautner, and Trevor Mudge
    Proceedings. 36th Annual IEEE/ACM International Symposium on Microarchitecture, 2003. MICRO-36. , 2003
  2. Leakage current: Moore’s law meets static power
    Nam Sung Kim, Todd Austin, David Blaauw, Trevor Mudge, Jie S Hu, Mary Jane Irwin, Mahmut Kandemir, and Vijaykrishnan Narayanan
    computer , 2003
  3. Reducing register ports using delayed write-back queues and operand pre-fetch
    Nam Sung Kim and Trevor Mudge
    2003
  4. A 2.3 Gb/s fully integrated and synthesizable AES Rijndael core
    Nam Sung Kim, Trevor Mudge, and Richard Brown
    Proceedings of the IEEE 2003 Custom Integrated Circuits Conference, 2003. , 2003
  5. The microarchitecture of a low power register file
    Nam Sung Kim and Trevor Mudge
    2003
  6. Shidhartha Das Sanjay Pant Toan Pham Rajeev Rao Conrad Ziesler David Blaauw Todd Austin Trevor Mudge and Krisztián Flautner. Razor: A Low-Power Pipeline Based on Circuit-Level Timing Speculation
    Dan Ernst and Nam Sung Kim
    Proc. 36th Intl. Symp. Microarch , 2003
  7. Power analyzer for pocket computing (papc)
    T Mudge, NS Kim, J Ringenberg, and T Kgil
    University of Michigan, Tech. Rep. , 2003
  8. Reducing Register Ports Using Delayed Write-Back
    Nam Sung Kim and Trevor Mudge
    Conference Proceedings , 2003
  9. Leakage current: Moore’s law meets static power. computer 36 12 (2003) 68–75
    Nam Sung Kim, Todd Austin, David Baauw, Trevor Mudge, Krisztián Flautner, Jie S Hu, Mary Jane Irwin, Mahmut Kandemir, and Vijaykrishnan Narayanan
    Google Scholar Google Scholar Digital Library Digital Library , 2003

2002

  1. Drowsy caches: simple techniques for reducing leakage power
    Krisztián Flautner, Nam Sung Kim, Steve Martin, David Blaauw, and Trevor Mudge
    ACM SIGARCH Computer architecture news , 2002
  2. Drowsy instruction caches. leakage power reduction using dynamic voltage scaling and cache sub-bank prediction
    Nam Sung Kim, Krisztian Flautner, David Blaauw, and Trevor Mudge
    35th Annual IEEE/ACM International Symposium on Microarchitecture, 2002.(MICRO-35). Proceedings. , 2002
  3. Challenges for architectural level power modeling
    Nam Sung Kim, Todd Austin, Trevor Mudge, and Dirk Grunwald
    Power aware computing , 2002
  4. Low-energy data cache using sign compression and cache line bisection
    Nam Sung Kim, Todd Austin, and Trevor Mudge
    Proceedings of the 2nd Annual Workshop on Memory Performance Issues (WMPI’02) , 2002
  5. Leakage Power Reduction using Dynamic Voltage Scaling and Cahe Sub-bank Prediction
    Nam Sung Kim, Krisztián Flautner, David Blaauw, and Trevor Mudge
    35th Annual IEEE/ACM International Symposium on Microarchitecture, 2002.(MICRO-35). Proceedings. , 2002

2001

  1. Power Aware Computing chapter Challenges for Architectural Level Power Modeling
    Nam Sung Kim, Todd Austin, Trevor Mudge, and Dirk Grunwald
    2001
  2. VLSI Implementation of the Symmetric Key Block Cipher with the Advanced Encryption Standard-Rijndael
    Nam Sung Kim, Richard B Brown, and Trevor Mudge
    University of Michigan , 2001
  3. VLSI Implementation of Binaural Spatializer using FIR Head-Related Transfer Function (HRTF)
    Nam Sung Kim, Jiyoun Kim, Min-Gyu Cho, and Tae-Young Choi
    2001

1998

  1. Virtual chip: making functional models work on real target systems
    Namseung Kim, Hoon Choi, Seungjong Lee, Seungwang Lee, In-Cheolo Park, and Chong-Min Kyung
    Proceedings of the 35th annual Design Automation Conference , 1998