EDBT 2026 Demo / reviewers in the wild / expert
Rabin A. Sugumar
dblp:32/2451
· DBLP profile ↗
6ranked-venue papers
3as first author
0since 2021 · last 2000
0000-0002-5561-0987ORCID · corroborated
Domains — the database's venue-derived domains; a paper can count in several
Systems, architecture and hardware · 6 · 3 first-authorSoftware engineering, systems software and programming languages · 2 · 1 first-author
Expertise — from the expertise taxonomy: the topics of the expert's papers under the CCF categories. A weight counts papers with recency: 1 for a paper about the topic, 0.3 when the topic is its context, halved every five years.
| Computer architecture, parallel and distributed computing, and storage systems
5 papers |
Processor architecture and microarchitecture · 40% Memory systems · 35% Performance modeling and evaluation · 22% | |
| Software engineering, system software, and programming languages
1 paper |
Compilers and program optimization · 100% |
Topics — the 16 heaviest of 18, each with the papers that count most for it
| Topic | Weight | Papers | Last | Evidence papers |
|---|---|---|---|---|
Memory systems
cache |
0.0 | 3 | 1995 | Set-Associative Cache Simulation Using Generalized Binomial Trees · ACM Trans. Comput. Syst. 1995 Cache performance in vector supercomputers · SC 1994 Efficient Simulation of Caches under Optimal Replacement with Applications to Miss Characterization · SIGMETRICS 1993 |
Processor architecture and microarchitecture
instruction set architecture |
0.0 | 1 | 2000 | Vector instruction set support for conditional operations · ISCA 2000 |
Processor architecture and microarchitecture › SIMD
vector instruction set |
0.0 | 1 | 2000 | Vector instruction set support for conditional operations · ISCA 2000 |
Performance modeling and evaluation › simulation
cache simulation |
0.0 | 2 | 1995 | Set-Associative Cache Simulation Using Generalized Binomial Trees · ACM Trans. Comput. Syst. 1995 Efficient Simulation of Caches under Optimal Replacement with Applications to Miss Characterization · SIGMETRICS 1993 |
Performance modeling and evaluation
simulation |
0.0 | 2 | 1995 | Set-Associative Cache Simulation Using Generalized Binomial Trees · ACM Trans. Comput. Syst. 1995 Efficient Simulation of Caches under Optimal Replacement with Applications to Miss Characterization · SIGMETRICS 1993 |
Memory systems › cache › cache organization
set-associative cache |
0.0 | 1 | 1995 | Set-Associative Cache Simulation Using Generalized Binomial Trees · ACM Trans. Comput. Syst. 1995 |
Memory systems › cache
cache behavior |
0.0 | 1 | 1994 | Cache performance in vector supercomputers · SC 1994 |
Memory systems › cache management
cache replacement |
0.0 | 1 | 1993 | Efficient Simulation of Caches under Optimal Replacement with Applications to Miss Characterization · SIGMETRICS 1993 |
Memory systems › cache management › cache replacement
optimal replacement |
0.0 | 1 | 1993 | Efficient Simulation of Caches under Optimal Replacement with Applications to Miss Characterization · SIGMETRICS 1993 |
Processor architecture and microarchitecture › instruction set architecture › instruction set extension
multimedia instruction set |
0.0 | 1 | 2000 | Vector instruction set support for conditional operations · ISCA 2000 |
Processor architecture and microarchitecture
SIMD |
0.0 | 1 | 2000 | Vector instruction set support for conditional operations · ISCA 2000 |
Memory systems
memory hierarchy |
0.0 | 1 | 1994 | Cache performance in vector supercomputers · SC 1994 |
High-performance computing › supercomputer architecture
vector supercomputer |
0.0 | 1 | 1994 | Cache performance in vector supercomputers · SC 1994 |
Performance modeling and evaluation
benchmarking |
0.0 | 1 | 1993 | Efficient Simulation of Caches under Optimal Replacement with Applications to Miss Characterization · SIGMETRICS 1993 |
Parallel and multicore computing
dataflow computing |
0.0 | 1 | 1993 | Predictability of load/store instruction latencies · MICRO 1993 |
Performance modeling and evaluation › benchmarking › benchmark suite
SPEC benchmarks |
0.0 | 1 | 1993 | Efficient Simulation of Caches under Optimal Replacement with Applications to Miss Characterization · SIGMETRICS 1993 |
Methods — techniques the papers use, named apart from their topics
top-down and bottom-up clustering · 0.0single-pass trace simulation · 0.0generalized binomial tree · 0.0trace-driven analysis · 0.0simulation · 0.0tree-based simulation · 0.0set-associative simulation · 0.0limited lookahead · 0.0
| Year | Publication | Venue | Position |
|---|---|---|---|
| 2000 | Vector instruction set support for conditional operationsabstractVector instruction sets are receiving renewed interest because of their applicability to multimedia. Current multimedia instruction sets use short vectors with SIMD implementations, but long vector, pipelined implementations have a number of advantages and are a logical next step in multimedia ISA development. Support for conditional operations (as occur in loops containing IF statements) is an important aspect of a vector ISA. Seven ISA alternatives for implementing conditional operations are systematically explored. Performance considerations are discussed through evaluation of a typical IF loop over a range of vector lengths and true conditional values. An approach using masked operations is shown to be one of the better methods, especially if its implementation is able to skip over blocks of false mask bits. Additional analyses of complex IF loops and parallel pipeline implementations support the masked operation approach. The paper concludes with a practical implementation of masked operations that skips over power-of-2-length blocks of false values. This implementation is simpler than skipping arbitrary-length blocks and provides similar performance. 1. James E. Smith 0001, Greg Faanes, Rabin A. Sugumar |
ISCA | 3 |
| 1995 | Set-Associative Cache Simulation Using Generalized Binomial TreesabstractSet-associative caches are widely used in CPU memory hierarchies, I/O subsystems, and file systems to reduce average access times. This article proposes an efficient simulation technique for simulating a group of set-associative caches in a single pass through the address trace, where all caches have the same line size but varying associativities and varying number of sets. The article also introduces a generalization of the ordinary binomial tree and presents a representation of caches in this class using the Generalized Binomial Tree (gbt). The tree representation permits efficient search and update of the caches. Theoretically, the new algorithm, GBF_LS, based on the gbt structure, always takes fewer comparisons than the two earlier algorithms for the same class of caches: all-associativity and generalized forest simulation. Experimentally, the new algorithm shows performance gains in the range of 1.2 to 3.8 over the earlier algorithms on address traces of the SPEC benchmarks. A related algorithm for simulating multiple alternative direct-mapped caches with fixed cache size, but varying line size, is also presented. Rabin A. Sugumar, Santosh G. Abraham |
ACM Trans. Comput. Syst. | 1 |
| 1994 | Cache performance in vector supercomputersabstractTraditional supercomputers use a flat multi-bank SRAM memory organization to supply high bandwidth at low latency. Most other computers use a hierarchical organization with a small SRAM cache and a slower, cheaper DRAM for the main memory. Such systems rely heavily on data locality for achieving optimum performance. This paper evaluates cache-based memory systems for vector supercomputers. We develop a simulation model for a cache-based version of the Cray Research C90 and use the NAS parallel benchmarks to provide a large-scale workload. We show that while caches reduce memory traffic and improve the performance of plain DRAM memory, they still lag behind cacheless SRAM. We identify the performance bottlenecks in DRAM-based memory systems and quantify their contribution to program performance degradation. We find the data fetch strategy to be a significant parameter affecting performance, we evaluate the performance of several fetch policies, and we show that small fetch sizes improve performance by maximizing the use of available memory bandwidth.> Leonidas I. Kontothanassis, Rabin A. Sugumar, Greg Faanes, James E. Smith 0001, Michael L. Scott |
SC | 2 |
| 1993 | Predictability of load/store instruction latenciesabstractPresents a model of coarse grain dataflow execution. The authors present one top down and two bottom up methods for generation of multithreaded code, and evaluate their effectiveness. The bottom up techniques start from a fine-grain dataflow graph and coalesce this into coarse-grain clusters. The top down technique generates clusters directly from the intermediate data dependence graph used for compiler optimizations. The authors discuss the relevant phases in the compilation process. They compare the effectiveness of the strategies by measuring the total number of clusters executed, the total number of instructions executed, cluster size, and number of matches per cluster. It turns out that the top down method generates more efficient code, and larger clusters. However the number of matches per cluster is larger for the top down method, which could incur higher cluster synchronization costs.> Santosh G. Abraham, Rabin A. Sugumar, Daniel Windheiser, Bob Rau |
MICRO | 2 |
| 1993 | Efficient Simulation of Caches under Optimal Replacement with Applications to Miss CharacterizationabstractCache miss characterization models such as the three Cs model are useful in developing schemes to reduce cache misses and their penalty. In this paper we propose the OPT model that uses cache simulation under optimal (OPT) replacement to obtain a finer and more accurate characterization of misses than the three Cs model. However, current methods for optimal cache simulation are slow and difficult to use. We present three new techniques for optimal cache simulation. First, we propose a limited lookahead strategy with error fixing, which allows one pass simulation of multiple optimal caches. Second, we propose a scheme to group entries in the OPT stack, which allows efficient tree based fully-associative cache simulation under OPT. Third, we propose a scheme for exploiting partial inclusion in set-associative cache simulation under OPT. Simulators based on these algorithms were used to obtain cache miss characterizations using the OPT model for nine SPEC benchmarks. The results indicate that miss ratios under OPT are substantially lower than those under LRU replacement, by up to 70% in fully-associative caches, and up to 32% in two-way set-associative caches. Rabin A. Sugumar, Santosh G. Abraham |
SIGMETRICS | 1 |
| 1991 | Parallel Simulation of Fully Associative Caches
Rabin A. Sugumar, Santosh G. Abraham |
ICPP (2) | 1 |