Boosting Performance of Directory-based Cache Coherence Protocols with Coherence Bypass at Subpage Granularity and A Novel On-chip Page Table
| dc.contributor.author | Soltaniyeh, Mohammadreza | |
| dc.contributor.author | Kadayif, Ismail | |
| dc.contributor.author | Ozturk, Ozcan | |
| dc.date.accessioned | 2025-01-27T20:59:54Z | |
| dc.date.available | 2025-01-27T20:59:54Z | |
| dc.date.issued | 2016 | |
| dc.department | Çanakkale Onsekiz Mart Üniversitesi | |
| dc.description | ACM International Conference on Computing Frontiers (CF) -- MAY 16-18, 2016 -- Como, ITALY | |
| dc.description.abstract | Chip multiprocessors (CMPs) require effective cache coherence protocols as well as fast virtual-to-physical address translation mechanisms for high performance. Directory-based cache coherence protocols are the state-of-the-art approaches in many-core CMPs to keep the data blocks coherent at the last level private caches. However, the area overhead and high associativity requirement of the directory structures may not scale well with increasingly higher number of cores. As shown in some prior studies, a significant percentage of data blocks are accessed by only one core, therefore, it is not necessary to keep track of these in the directory structure. In this study, we have two major contributions. First, we show that compared to the classification of cache blocks at page granularity as done in some previous studies, data block classification at subpage level helps to detect considerably more private data blocks. Consequently, it reduces the percentage of blocks required to be tracked in the directory significantly compared to similar page level classification approaches. This, in turn, enables smaller directory caches with lower associativity to be used in CMPs without hurting performance, thereby helping the directory structure to scale gracefully with the increasing number of cores. Memory block classification at subpage level, however, may increase the frequency of the Operating System's (OS) involvement in updating the maintenance bits belonging to subpages stored in page table entries, nullifying some portion of performance benefits of subpage level data classification. To overcome this, we propose a distributed on-chip page table as a our second contribution. | |
| dc.description.sponsorship | Assoc Comp Machinery | |
| dc.description.sponsorship | Scientific and Technological Research Council of Turkey (TUBITAK) [113E258] | |
| dc.description.sponsorship | This study was fully funded by the Scientific and Technological Research Council of Turkey (TUBITAK) with a grant 113E258. | |
| dc.identifier.doi | 10.1145/2903150.2903175 | |
| dc.identifier.endpage | 187 | |
| dc.identifier.isbn | 978-1-4503-4128-8 | |
| dc.identifier.scopus | 2-s2.0-84978519653 | |
| dc.identifier.scopusquality | N/A | |
| dc.identifier.startpage | 180 | |
| dc.identifier.uri | https://doi.org/10.1145/2903150.2903175 | |
| dc.identifier.uri | https://hdl.handle.net/20.500.12428/26875 | |
| dc.identifier.wos | WOS:000693994700022 | |
| dc.identifier.wosquality | N/A | |
| dc.indekslendigikaynak | Web of Science | |
| dc.indekslendigikaynak | Scopus | |
| dc.language.iso | en | |
| dc.publisher | Assoc Computing Machinery | |
| dc.relation.ispartof | Proceedings of The Acm International Conference on Computing Frontiers (Cf'16) | |
| dc.relation.publicationcategory | Konferans Öğesi - Uluslararası - Kurum Öğretim Elemanı | |
| dc.rights | info:eu-repo/semantics/closedAccess | |
| dc.snmz | KA_WoS_20250125 | |
| dc.subject | cache coherence | |
| dc.subject | directory cache | |
| dc.subject | many-core system | |
| dc.subject | virtual memory | |
| dc.subject | page table | |
| dc.title | Boosting Performance of Directory-based Cache Coherence Protocols with Coherence Bypass at Subpage Granularity and A Novel On-chip Page Table | |
| dc.type | Conference Object |











