CS752 Reader1
This reader may still change slightly, and, if so, I will send
email notice of changes. Links need to be verified.
H&P is
John L. Hennessy and David A. Patterson,
Computer Architecture: A Quantitative Approach,
Morgan Kaufmann Publishers, Fourth Edition, 2006.
HJ&S is
Mark D. Hill, Norman P. Jouppi, and Gurindar S. Sohi,
Readings in Computer Architecture,
Morgan Kaufmann Publishers, 2000.
Technology, Cost, Performance, Power, etc.
-
H&P Chapter 1
-
Standard Performance Evaluation Corporation (SPEC).
URL:
http://www.specbench.org/. Read the "run and reporting rules"
for SPEC CPU2006 and SPEC jAppServer2004. You may skim the rest of the web site.
-
ITRS Roadmap -- Executive Summary, Go to
URL
http://www.itrs.net/Links/2005ITRS/Home2005.htm
and click on
Executive Summary,
89 pages. Read Introduction (pp. 1-10) and flip through
Grand Challenges (pp. 11-18), reading, at least, titles. Study the trends in the
Overall Roadmap Technology Characteristics (pp. 59-85). Don't try to memorize
the tables. Rather, identify key facts and trends. Such as "what is the overall
scaling trend?" and "what is the target yield range for volume production?".
-
Gordon E. Moore,
Cramming More Components onto Integrated Circuits,
Electronics, April 1965.
Reprinted in HJ&S pp. 56-59.
To be reviewed
-
Ethan Mollick, MIT Sloan School of Management,
Establishing Moore's Law,
IEEE Annals of Computing, July-September 2006 (Vol. 28, No. 3) pp. 62-75,
IEEE Xplore link.
Reference reading
-
David A. Patterson,
Latency lags bandwith,
Communications of the ACM,
October 2004.
Online PDF for University of Wisconsin only.
To be reviewed
Instruction Sets
-
H&P Appendix B.
-
William A. Wulf.
Compilers and Computer Architecture,
IEEE Computer,
July 1981.
Reprinted in HJ&S pp. 119-125.
Reference reading
-
Robert Colwell et al.
Instructions Sets and Beyond: Computers, Complexity, and Concurrency.
IEEE Computer,
September 1985.
Reprinted in HJ&S pp. 144-155.
To be reviewed
-
Burger et al.
Scaling to end of Silicon with EDGE architectures.
IEEE Computer, July 2004. PDF download.
To be reviewed
-
J. S. Emer and D. W. Clark.
A Characterization of Processor Performance in the VAX-11/780,
ISCA 1984,
Reprinted in HJ&S pp. 101-110.
Portal.acm online.
To be reviewed
-
B. Sprunt. Pentium 4 performance-monitoring features, IEEE Micro,
July 2002.
IEEE Xplore link.
Reference reading
-
IA-32 Intel(R) Architecture Software Developer's Manual,
Volume 1: Basic Architecture.
Online PDF for University of Wisconsin only.,
476 pages (reference).
Reference reading
Pipelining and pipeline analysis
-
H&P Appendix A has a review.
-
Dan Ernst, et al.,
A Low-Power Pipeline Based on Circuit-Level Timing Speculation,
MICRO 2003
PDF download
To be reviewed
-
Hartstein and Puzak,
Optimum Power/Performance Pipeline Depth, MICRO 2003.
PDF download
To be reviewed
-
Hrishikesh et al.,
The Optimum Logic Depth Per Pipeline Stage is 6 to 8 FO4 Inverter Delays,
ISCA 2002. PDF download
Reference reading
Dynamic ILP
-
H&P Chapter 2 and 3.1-3.3
-
Gurindar S. Sohi and S. Vajapeyam.
Instruction Issue Logic for High-Performance, Interruptible,
Multiple Functional Unit, Pipelined Computers,
ISCA 1987.
Reprinted in HJ&S pp. 244-251.
To be reviewed
-
T-Y. Yeh and Y. Patt.
Two-level Adaptive Training Branch Prediction,
ISCA 1991.
Reprinted in HJ&S pp. 228-237.
To be reviewed
-
Seznec, A., Felix, S., Krishnan, V., and Sazeides, Y.
Design tradeoffs for the alpha EV8 conditional branch predictor.
ISCA 2002.
IEEE Xplore link
To be reviewed
-
Tse-Yu Yeh and Patt, Y.N.
Alternative Implementations of Two-Level Adaptive Branch Prediction.
ISCA 1992
IEEE Xplore link
Reference reading
-
Sethumadhavan et al.,
Scalable Hardware Memory Disambiguation for High-ILP Processors, MICRO 2003.
PDF download
To be reviewed
-
J. E. Smith and A. R. Pleszkun.
Implementing Precise Interrupts in Pipelined Processors, IEEE Trans. on Computers,
May 1988.
Reprinted in HJ&S pp. 202-213.
Reference reading
-
Kenneth C. Yeager.
The MIPS R10000 Superscalar Microprocessor, IEEE Micro,
April 1996.
Reprinted in HJ&S pp. 275-287.
To be reviewed
-
D. Papworth.
Tuning the Pentium Pro Architecture, IEEE Micro,
April 1996.
Reprinted in HJ&S pp. 660-667.
Reference reading
- Robert E. Kessler. The Alpha 21264 Microprocessor, IEEE Micro,
March/April 1999, (Vol. 19, No. 2), pp. 24-36.
IEEE Xplore link
Reference reading
-
Simcha Gochman, Ronny Ronen, Ittai Anati, Ariel Berkovits, Tsvika Kurts,
Alon Naveh, Ali Saeed, Zeev Sperber, Robert C. Valentine,
The Intel (R) Pentium(R) M Processor: Microarchitecture and Performance Intel Technology Journal,
May 2003.
PDF download
Reference reading
-
Timothy J. Slegel, et al.,
IBM's S/390 G5 Microprocessor,
IEEE Micro, Mar/Apr 1999,
IEEE Xplore link
Reference reading
-
Borch et al.,
Loose Loops Sink Chips,
HPCA 2002,
IEEE Xplore link
Reference reading
Caches
-
H&P Chapter 5.
-
Norman P. Jouppi.
Improving Direct-Mapped Cache Performance by the Addition of a Small Fully-Associative Cache and Prefetch Buffers,
ISCA 1990 ,
Reprinted in HJ&S pp. 395-404.
To be reviewed
-
Norman P. Jouppi and Steven J.E. Wilton. Tradeoffs in two-level on-chip caching.
ISCA 1994. ACM link.
Reference reading
-
David H. Albonesi,
Selective Cache Ways: On-demand Cache Resource Allocation,
MICRO 1999.
IEEE Xplore link
To be reviewed
-
Changkyu Kim, Doug Burger, and Stephen W. Keckler.
An Adaptive, Non-Uniform Cache Structure for Wire-Dominated On-Chip Caches,
ASPLOS 2002
PDF download
To be reviewed
Miscellaneous