🔍 Reference Analysis

← Back to Paper Summaries
In Source: 163 Missing: 1201 Total Unique References: 1364 PDFs Analyzed: 27

In paper-source (163 papers)

All References by Citation Count (top 200 of 1364)

#TitleCitationsSourceYearVenue
1The gem5 Simulator6NO2011CAN
2Memory Access Scheduling5NO2000ISCA
3Fine-grained activation for power reduction in DRAM5NO2010IEEE Micro
4Architecting phase change memory as a scalable DRAM alternative5NO2009ISCA
5A Scalable Processing-in-memory Accelerator for Parallel Graph Processing5YES2015ISCA
6Rethinking DRAM design and organization for energy-constrained multi-cores4NO2010ISCA
7Adaptive granularity memory systems: A tradeoff between storage efficiency and throughput4YES2011ISCA
8Mini-rank: Adaptive DRAM architecture for improving memory power efficiency4NO2008MICRO
9A Case for Exploiting Subarray-Level Parallelism (SALP) in DRAM4YES2012ISCA
10DRAM Circuit Design: Fundamental and High-Speed Topics4NO2007Book
11Efficient Virtual Memory for Big Memory Servers4NO2013ISCA
12Rodinia: A benchmark suite for heterogeneous computing4NO2009IISWC
13EIE: Efficient Inference Engine on Compressed Deep Neural Network4YES2016ISCA
14Very Deep Convolutional Networks for Large- Scale Image Recognition,4YES2015ICLR
15Supporting x86-64 address translation for 100s of GPU lanes4YES2014HPCA
16DRAMSim2: A cycle accurate memory system simulator4NO2011CAL
17PIM-enabled Instructions: A Low-overhead, Locality-aware Processing-in-memory Architecture4NO2015ISCA
18Memory Systems: Cache, DRAM, Disk3NO2007Book
19Future scaling of processor-memory interfaces3NO2009SC
20Structural aspects of the system/360 model 85: II the cache3NO1968IBM Systems Journal
21Software caching and computation migration in Olden3YES1995Tech Report
22TOP-PIM: Throughput-oriented Programmable Processing in Memory3NO2014HPDC
23Shared Last-level TLBs for Chip Multiprocessors3NO2011HPCA
24Optimizing NUCA Organizations and Wiring Alternatives for Large Caches with CACTI 6.03NO2007MICRO
25USENIX Association3NO2004USENIX
26Minimalist Open-page: A DRAM Page-mode Scheduling Policy for the Many-core Era3NO2011MICRO
27CACTI-3DD: Architecture-level Modeling for 3D Die-stacked DRAM Main Memory3NO2012DATE
28Understanding the Energy Consumption of Dynamic Random Access Memories3NO2010MICRO
29Half-DRAM: A High-bandwidth and Low-power DRAM Architecture from the Rethinking of Fine-grained Activation3YES2014ISCA
30A permutation-based page interleaving scheme to reduce row-buffer conflicts and exploit data locality3NO2000MICRO
31SPEC CPU2006 Benchmark Descriptions3NO2006SIGARCH Computer Architecture News
32Eyeriss: A Spatial Architecture for Energy-efficient Dataflow for Convolutional Neural Networks3YES2016ISCA
33Architectural Support for Address Translation on GPUs: Designing Memory Management Units for CPU/GPUs with Unified Address Spaces3YES2014ASPLOS
34DRAM Errors in the Wild: A Large-Scale Field Study3NO2009SIGMETRICS
35Base-Delta-Immediate Compression: Practical Data Compression for On-Chip Caches3NO2012PACT
36ISAAC: A Convolutional Neural Network Accelerator with In-Situ Analog Arithmetic in Crossbars3YES2016ISCA
37Disaggregated Memory for Expansion and Sharing in Blade Servers2NO2009ISCA
38Towards Energy-Proportional Datacenter Memory with Mobile DRAM2NO2012ISCA
39BOOM: Enabling Mobile Memory based Low-Power Server DIMMs2NO2012ISCA
40Memory Power Management via Dynamic Voltage/Frequency Scaling2NO2011ICAC
41Skinflint DRAM System: Minimizing DRAM Chip Writes for Low Power2YES2013HPCA
42The Dirty-block Index2NO2014ISCA
43VLSI Memory Chip Design2NO2001Springer
44Spectre Attacks: Exploiting Speculative Execution2NO2018arXiv
45Meltdown2NO2018arXiv
46On the effectiveness of address-space randomization2NO2004CCS
47Efficient Address Translation for Architectures with Multiple Page Sizes2NO2017ASPLOS
48Characterization of silent stores2NO2000PACT
49Improving the reliability of on-chip L2 cache using redundancy2NO2007ICCD
50Dynamically exploiting narrow width operands to improve processor power and performance2NO1999HPCA
51Preventing PCM banks from seizing too much power2NO2011MICRO
52Experimental evaluation of on-chip microprocessor cache memories2NO1984ISCA
53DRAM Energy reduction by prefetching-based memory traffic clustering2NO2011GLSVLSI
54On the value locality of store instructions2NO2000ISCA
55Silent stores for free2NO2000MICRO
56Understanding and designing new server architectures for emerging warehouse-computing environments2NO2008ISCA
57Zesto: A cycle-level simulator for highly detailed microarchitecture exploration2NO2009ISPASS
58PowerNap: Eliminating server idle power2NO2009ASPLOS
59Memory-link compression schemes: A value locality perspective2NO2008IEEE TC
60Energy reduction for STT-RAM using early write termination2NO2009ICCAD
61CoLT: Coalesced Large-Reach TLBs2NO2012MICRO
62GPUs and the Future of Parallel Computing2NO2011IEEE Micro
63LazyPIM: An Efficient Cache Coherence Mechanism for Processing-in-Memory2NO2017IEEE CAL
64The Architecture of the DIVA Processing-in-memory Chip2NO2002ICS
653D-stacked Memory-side Acceleration: Accelerator and System Design2NO2013WoNDP
66Transparent Offloading and Mapping (TOM): Enabling Programmer-transparent Near-data Processing in GPU Systems2NO2016ISCA
67Accelerating Pointer Chasing in 3D-stacked Memory: Challenges, Mechanisms, Evaluation2NO2016ICCD
68Hybrid Memory Cube: New DRAM Architecture Increases Density and Performance2YES2012VLSIT
69FlexRAM: Toward an Advanced Intelligent Memory System2NO1999ICCD
70Ramulator: A Fast and Extensible DRAM Simulator2NO2016IEEE CAL
71EXECUBE: A New Architecture for Scaleable MPPs2NO1994ICPP
72Simultaneous Multi-Layer Access: Improving 3D-Stacked Memory Bandwidth at Low Cost2NO2016ACM TACO
73Active Pages: A Computation Model for Intelligent Memory2NO1998ISCA
74A Case for Intelligent RAM2NO1997IEEE Micro
75Scheduling Techniques for GPU Architectures with Processing-in-memory Capabilities2NO2016PACT
76Fast Bulk Bitwise AND and OR in DRAM2YES2015IEEE CAL
77A Logic-in-Memory Computer2NO1970IEEE Trans. Comput.
78Beyond the Wall: Near-Data Processing for Databases2YES2015DAMON
79Compute Caches2NO2017HPCA
80NDA: Near-DRAM Acceleration Architecture Leveraging Commodity DRAM Devices and Standard Memory Modules2NO2015HPCA
81Pinatubo: A Processing-in-Memory Architecture for Bulk Bitwise Operations in Emerging Non-Volatile Memories2YES2016DAC
82In-Datacenter Performance Analysis of a Tensor Processing Unit2NO2017ISCA
83Di- annao: A small-footprint high-throughput accelerator for ubiquitous machine-learning,2NO2014ASPLOS
84Dadiannao: A machine-learning supercom- puter,2NO2014MICRO
85Shidiannao: Shifting vision processing closer to the sensor,2NO2015ISCA
86Optimizing fpga-based accelerator design for deep convolutional neural networks,2YES2015fpga
87Flexflow: A flexible dataflow accelerator architecture for convolutional neural networks,2YES2017HPCA
88NVIDIA Tesla: A Unified Graphics and Computing Architecture2NO2008IEEE Micro
89Controller for a Synchronous DRAM that Maximizes Throughput by Allowing Memory Requests and Commands to be Issued Out of Order2YES1997US Patent 5630096
90Large-reach Memory Management Unit Caches2NO2013MICRO
91Inter-core Cooperative TLB for Chip Multiprocessors2NO2010ASPLOS
92Supporting Address Translation for Accelerator-Centric Architectures2NO2017HPCA
93Redundant Memory Mappings for Fast Access to Large Memories2NO2015ISCA
94Adaptive Cache Management for Energy-Efficient GPU Computing2YES2014MICRO
95iGPU: Exception Support and Speculative Execution on GPUs2NO2012ISCA
96Observations and opportunities in architecting shared virtual memory for heterogeneous systems2YES2016ISPASS
97ATLAS: A Scalable and High- Performance Scheduling Algorithm for Multiple Memory Controllers,2YES2010HPCA
98Thread Cluster Memory Scheduling: Exploiting Differences in Memory Access Behavior,2NO2010MICRO
99Reducing Memory Interference in Multicore Systems via Application-Aware Memory Channel Partitioning,2NO2011MICRO
100Exploiting inter-warp heterogeneity to improve gpgpu performance,2NO2015PACT
101Analyzing CUDA workloads using a detailed GPU simulator,2NO2009ISPASS
102Orchestrated scheduling and prefetching for GPGPUs,2NO2013ISCA
103Managing GPU concurrency in heterogeneous architectures,2NO2014MICRO
104Locality-driven dynamic GPU cache bypassing,2NO2015-
105Cache-conscious wavefront scheduling,2NO2012MICRO
106IEEE Computer Society2NO2014MICRO
107Microbank: Architecting Through-Silicon Interposer-Based Main Memory Systems2NO2014SC
108A White Paper on the Benefits of Chipkill-Correct ECC for PC Server Main Memory2YES1997IBM Microelectronics Division
109Architecting an Energy-Efficient DRAM System for GPUs,2NO2017HPCA
110USIMM: the Utah Simulated Memory Module2NO2012-
111A 1.2v 38nm 2.4gb/s/pin 2gb ddr4 sdram with bank group and 4 half-page architecture2NO2012ISSCC
112A non-volatile microcontroller with integrated floating-gate transistors2NO2011IEEE
113A Simpler, Safer Programming and Execution Model for Intermittent Systems2NO2015ACM
114Arrakis: The operating system is the control plane2NO2014OSDI
115A Full GPU Virtualization Solution with Mediated Pass-Through,2NO2014USENIX
116Imagenet classification with deep convolutional neural networks,2YES2012NeurIPS
117Going deeper with convolutions,2NO2015CVPR
118Cambricon-x: An accelerator for sparse neural networks,2NO2016MICRO
119Long short-term memory,2NO1997Neural Computation
120A Case for Toggle-Aware Compression for GPU Systems,2NO2016HPCA
121Nv-heaps: making persistent objects fast and safe with next-generation, non-volatile memories2NO2011ACM SIGPLAN Notices
122High-performance transactions for persistent memories2NO2016ASPLOS
123Dudetm: Building durable transactions with decoupling for persistent memory2YES2017ASPLOS
124An analysis of persistent memory use with whisper2NO2017ASPLOS
125Mnemosyne: Lightweight persistent memory2NO2011ACM SIGARCH Computer Architecture News
126Efficient Memory Integrity Verification and Encryption for Secure Processors2YES2003MICRO
127Incidental Computing on IoT Nonvolatile Processors2NO2017IEEE
128The Dynamic Granularity Memory System2NO2012ISCA
129Co-architecting Controllers and DRAM to Enhance DRAM Process Scaling2NO2014The Memory Forum
130DDR4 SDRAM STANDARD2NO2012JEDEC
131Redundancy techniques for high-density drams2NO1997IEEE International Conference on Innovative Systems in Silicon
132Archshield: Architectural framework for assisting dram scaling by tolerating high error rates2NO2013ACM SIGARCH Computer Architecture News
1338Gb DDR4 SDRAM2NO-SK hynix
134The Netflix Prize2NO2017KDD
135Graphicionado: A high-performance and energy-efficient accelerator for graph analytics2NO2016MICRO
136Practical Near-Data Processing for In- Memory Analytics Frameworks,2YES2015PACT
137GraphPIM: Enabling Instruction-Level PIM Offloading in Graph Computing Frameworks,2NO2017HPCA
138Nvsim: A circuit-level performance, energy, and area model for emerging nonvolatile memory,2YES2012IEEE
139Cnvlutin: Ineffectual-neuron-free deep neural network computing,2NO2016ISCA
140DaDianNao: A Machine-Learning Supercomputer2NO2014MICRO
141Enhancing lifetime and security of PCM-based main memory with start-gap wear leveling2YES2009MICRO
142Overcoming the challenges of crossbar resistive memory architectures2NO2015HPCA
143The Datacenter as a Computer1NO2009Book
144Tiered Memory: An Iso-Power Memory Architecture to Address the Memory Wall1NO2012IEEE TC
145Smart Refresh: An Enhanced Memory Controller Design for Reducing Energy in Conventional and 3D Die-Stacked DRAMs1YES2007MICRO
146A Comprehensive Approach to DRAM Power Management1NO2008HPCA
147Flikker: Saving DRAM Refresh-power through Critical Data Partitioning1NO2011ASPLOS
148Rethinking DRAM Power Modes for Energy Proportionality1NO2012MICRO
149NVMain: An Architectural-Level Main Memory Simulator for Emerging Non-volatile Memories1YES2012ISVLSI
150Power-Supply-Network Design in 3D Integrated Systems1NO2011ISQED
151Measurement, Analysis and Improvement of Supply Noise in 3D ICs1NO2011VLSI Symposium
152The Memory System: You Can't Avoid It, You Can't Ignore It, You Can't Fake It1NO2009Book
153An Optimized 3D-Stacked Memory Architecture by Exploiting Excessive, High-Density TSV Bandwidth1NO2010HPCA
154Pragmatic Integration of an SRAM Row Cache in Heterogeneous 3-D DRAM Architecture Using TSV1NO2011IEEE TVLSI
155Thermal Management of High Power Memory Module for Server Platforms1YES2008ITHERM
156The Datacenter As a Computer: An Introduction to the Design of Warehouse-Scale Machines1YES2009Morgan and Claypool Publishers
157Decoupled DIMM: Building High-bandwidth Memory System Using Low-speed DRAM Devices1NO2009ISCA
158More is Less: Improving the Energy Efficiency of Data Movement via Opportunistic Use of Sparse Codes1NO2015MICRO
159Improving Power and Data Efficiency with Threaded Memory Modules1NO2006ICCD
160Multicore DIMM: An Energy Efficient Memory Module with Independently Controlled DRAMs1NO2009CAL
161Multiple Sub-row Buffers in DRAM: Unlocking Performance and Energy Improvement Opportunities1NO2012ICS
162Energy Efficient Data Encoding in DRAM Channels Exploiting Data Value Similarity1YES2016ISCA
163Power Protocol: Reducing Power Dissipation on Off-Chip Data Buses1NO2002MICRO
164Bus-Invert Coding for Low-Power I/O1NO1995TVLSI
165DRAM-Aware Last-Level Cache Writeback: Reducing Write-Caused Interference in Memory Systems1NO2010Tech. Rep.
166The Virtual Write Queue: Coordinating DRAM and Last-Level Cache Policies1NO2010ISCA
167Conditional-Capture Flip-Flop for Statistical Power Reduction1NO2001JSSC
168Error Control Coding: Fundamentals and Applications1NO2004Pearson-Prentice Hall
169Decoupled Sectored Caches: Conciliating Low Tag Implementation Cost1NO1994ISCA
170A Data Cache with Multiple Caching Strategies Tuned to Different Types of Locality1NO1995ICS
171The Pool of Subsectors Cache Design1NO1999ICS
172Exploiting Spatial Locality in Data Caches Using Spatial Footprints1NO1998ISCA
173Accurate and Complexity-Effective Spatial Pattern Prediction1NO2004HPCA
174SimPoint 3.0: Faster and More Flexible Program Analysis1NO2005MoBS
175Using Storage Cells to Perform Computation1NO2014US Patent 8908465
176In-memory Computational Device1NO2015US Patent 9653166
177GateKeeper: A New Hardware Architecture for Accelerating Pre-Alignment in DNA Short Read Mapping1NO2017Bioinformatics
178A Bit-Parallel, General Integer-Scoring Sequence Alignment Algorithm1NO2013CPM
179Space/time Trade-offs in Hash Coding with Allowable Errors1NO1970ACM Communications
180LazyPIM: Efficient Support for Cache Coherence in Processing-in-Memory Architectures1YES2017arXiv
181Bitmap Index Design and Evaluation1NO1998SIGMOD
182Improving DRAM Performance by Parallelizing Refreshes with Accesses1NO2014HPCA
183Understanding Latency Variation in Modern DRAM Chips: Experimental Characterization, Analysis, and Optimization1YES2016SIGMETRICS
184Low-cost Inter-linked Subarrays (LISA): Enabling Fast Inter-subarray Data Movement in DRAM1NO2016HPCA
185Understanding Reduced-voltage Operation in Modern DRAM Devices: Experimental Characterization, Analysis, and Mechanisms1YES2017SIGMETRICS
186Linux Device Drivers1NO2005O'Reilly Media
187An Efficient and Scalable Semiconductor Architecture for Parallel Automata Processing1YES2014IEEE TPDS
188Computational RAM: Implementing Processors in Memory1NO1999IEEE DT
189Suppressing Power Supply Noise Using Data Scrambling in Double Data Rate Memory Systems1YES2009US Patent 8503678
190Programming the FlexRAM Parallel Intelligent Memory System1NO2003PPoPP
191Processing in Memory: The Terasys Massively Parallel PIM Array1YES1995Computer
192BitFunnel: Revisiting Signatures for Search1NO2017SIGIR
193A Dichromatic Framework for Balanced Trees1NO1978SFCS
194Error Detecting and Error Correcting Codes1NO1950BSTJ
195Optical Image Encryption Based on XOR Operations1NO1999SPIE OE
196ChargeCache: Reducing DRAM Latency by Exploiting Row Access Locality1YES2016HPCA
197SoftMC: A Flexible and Practical Open-source Infrastructure for Enabling Experimental DRAM Studies1YES2017HPCA
198One-Transistor Type DRAM1NO2009US Patent 7701751
199An Energy-efficient VLSI Architecture for Pattern Recognition via Deep Embedding of Computation in SRAM1YES2014ICASSP
200GRIM-filter: Fast Seed Filtering in Read Mapping Using Emerging Memory Technologies1NO2017arXiv