| IBM Journal of Research and Development, vol. 34, No. 1, Jan. 1990, pp. 1-70. |
| Aiken, A. and Nicolau, A., “Perfect Pipelining: A New Loop Parallelization Technique*,” pp. 221-235. |
| Goodman, J.R. And Hsu, W., “Code Scheduling and Register Allocation in Large Basis Blocks,” ACM, 1988, pp. 442-452. |
| Groves, R.D. and Oehler, R., “An IBM Second Generation RISC Processor Architecture,” IEEE, 1989, pp. 134-137. |
| Horst, R.W. et al., “Multiple Instruction Issue in the NonStop Cyclone Processor,” IEEE, 1990, pp. 216-226. |
| Jouppi, N.P., “Integration and Packaging Plateaus of Processor Performance,” IEEE, 1989, pp. 229-232. |
| Jouppi, N.P., “The Nonuniform Distribution of Instruction-Level and Machine Parallelism and Its Effect on Performance,” IEEE Transactions on Computers, vol. 38, No. 12, Dec. 1989, pp. 1645-1658. |
| Lam, M.S., “Instruction Scheduling For Superscalar Architectures,” Annu. Rev. Computer Science, vol. 4, 1990, pp. 173-201. |
| Murakami, K. et al., “SIMP (Single Instruction Stream/Multiple Instruction Pipelining): A Novel High-Speed Single-Processor Architecture,” ACM, 1989, pp. 78-85. |
| Smith, M.D. et al., “Boosting Beyond Static Scheduling in a Superscalar Processor,” IEEE, 1990, pp. 344-354. |
| Smith et al., “Implementation of Precise Interrupts in Pipelined Processors,” Proceedings of the 12th Annual International Symposium on Computer Architecture, Jun. 1985, pp. 36-44. |
| Wedig, R.G., Detection of Concurrency in Directly Executed Language Instruction Streams, (Dissertation), Jun. 1982, pp. 1-179. |
| Agerwala et al., “High Performance Reduced Instruction Set Processors,” IBM Research Division, Mar. 31, 1987, pp. 1-61. |
| Gross et al., “Optimizing Delayed Branches,” Proceedings of the 5th Annual Workshop on Microprogramming, Oct. 5-7, 1982, pp. 114-120. |
| Tjaden et al., “Representation of Concurrency with Ordering Matrices,” IEEE Trans. on Computers, vol. C-22, No. 8, Aug. 1973, pp. 752-761. |
| Weiss et al., “Instruction Issue Logic in Pipelined Supercomputers,” Reprinted from IEEE Trans. On Computers, vol. C-33, No. 11, Nov. 1984, pp. 1013-1022. |
| Tomasulo, R.M., “An Efficient Algorithm for Exploiting Multiple Arithmetic Units,” IBM Journal, vol. 11, Jan. 1967, pp. 25-33. |
| Tjaden et al., “Detection and Parallel Execution of Independent Instructions,” IEEE Trans. On Computers, vol. C-19, No. 10, Oct. 1970, pp. 889-895. |
| Pleszkun et al., “The Performance Potential of Multiple Functional Unit Processors,” Proceedings of the 15th Annual Symposium on Computer Architecture, Jun. 1988, pp. 37-44. |
| Pleszkun et al., “WISQ: A Restartable Architecture Using Queues,” Proceedings of the 14th International Symposium on Computer Architecture, Jun. 1987, pp. 290-299. |
| Patt et al., “Critical Issues Regarding HPS, A High Performance Microarchitecture,” Proceedings of the 18th Annual Workshop on Microprogramming, Dec. 1985, pp. 109-116. |
| Hwu et al., “Checkpoint Repair for High-Performance Out-of-Order Execution Machines,” IEEE Trans. on Computers, vol. C-36, No. 12, Dec. 1987, pp. 1496-1514. |
| Patt et al., “HPS, A New Microarchitecture: Rationale and Introduction,” Proceedings of the 18th Annual Workshop on Microprogramming, Dec. 1985, pp. 103-108. |
| Keller, R.M., “Look-Ahead Processors,” Computing Surveys, vol. 7, No. 4, Dec. 1975, pp. 177-195. |
| Jouppi et al., “Available Instruction-Level Parallelism for Superscalar and Superpipelined Machines,” Proceedings of the 3rd International Conference on Architectural Support for Programming Languages and Operating Systems, Apr. 1989, pp. 272-282. |
| Hwu et al., “HPSm, a High Performance Restricted Data Flow Architecture Having Minimal Functionality,” Proceedings from ISCA-13, Tokyo, Japan, Jun. 2-5, 1986, pp. 297-306. |
| Hwu et al., “Exploiting Parallel Microprocessor Microarchitectures with a Compiler Code Generator,” Proceedings of the 15th Annual Symposium on Computer Architecture, Jun. 1988, pp. 45-53. |
| Colwell et al., “A VLIW Architecture for a Trace Scheduling Compiler,” Proceedings of the 2nd International Conference on Architectural Support for Programming Languages and Operating Systems, Oct. 1987, pp. 180-192. |
| Uht, A.K., “An Efficient Hardware Algorithm to Extract Concurrency from General-Purpose Code,” Proceedings of the 19th Annual Hawaii International Conference on System Sciences, 1986, pp. 41-50. |
| Charlesworth, A.E., “An Approach to Scientific Array Processing: The Architectural Design of the AP-120B/FPS-164 Family,” Computer, vol. 14, Sep. 1981, pp. 18-27. |
| Lightner et al., “The Metaflow Architecture,” p. 11, 12, 63, 64, 67, and 68, Jun. 1991, IEEE Micro Magazine. |
| Michael D. Smith et al., “Limits on Multiple Instruction Issue,” Proceedings of the 3rd International Conference on Architectural Support for Programming Languages and Operating Systems, Computer Architecture News, No. 2, Apr. 17, 1989, pp. 290-302. |
| Keller, R., “Look-Ahead Processors,” Computing Surveys, vol. 7, No. 4, Dec. 1975. |
| Critical Issues Regarding HPS, A High Performance Microarchitecture, Yale N. Patt, Stephen W. Melvin, Wen-Mei Hwu and Michael C. Shebanow; The 18th Annual Workshop on Microprogramming, Pacific Grove,California, Dec. 3-6, 1985, IEEE Computer Order No. 653, pp. 109-116. |
| HPS, A New Microarchitecture: Rationale and Introduction, Yale N. Patt, Wen-Mei Hwu and Michael Shebanow; The 18th Annual Workshop on Microprocessing, Pacific Grove, California, Dec. 3-6, 1985, IEEE Computer Society Order No. 653, pp. 103-108. |
| Johnson, Mike, Superscalar Microprocessor Design, “Chapter 5—The Role of Exception Recovery,” pp. 87-102, “Chapter 6—Register Dataflow,” pp. 103-125, Prentice Hall, 1991. |
| Peleg et al., “Future Trends in Microprocessors: Out-of-Order Execution, Spec. Branching and Their CISC Performance Potential,” Mar. 1991. |
| Popescu Val et al., “The Metaflow Architecture,” IEEE Micro., vol. 11, No. 3, pp. 10-13, 63-73, Jun. 1991. |
| Dwyer, A Multiple, Out-of-Order Instruction Issuing System for Superscalar Processors, (All), Aug. 1991. |
| Hwu, Wen-Mei, Steve Melvin, Mike Shebanow, Chein Chen, Jia-Juin Wei, Yale Patt, “An HPS Implementation of VAX: Initial Design and Analysis,” Proceedings ofthe Nineteenth Annual Hawaii International Conference on System Sciences, pp. 282-291, 1986. |
| Hwu et al., “Experiments with HPS, A restricted Data Flow Microarchitecture for High Performance Computers,” COMPCON 86, 1986. |
| Hwu, Wen-Mei and Yale N. Patt, “HPSm, A High Performance Restricted Data Flow Architecture Having Minimal Functionality,” Proceedings of the 18th International Symposium on Computer Architecture, pp. 297-306, Jun. 1986. |
| Yale N. Patt, Stephen W. Melvin, Wen-Mei Hwu, Michael C. Shebanow, Chein Chen, Jiajuin Wei, “Run-Time Generation of HPS Microinstructions From a VAX Instruction Stream,” Proceedings of MICRO 19 Workshop, New York, New York, pp. 1-7, Oct. 1986. |
| Swenson, John A. and Yale N. Patt, “Hierarchical Registers for Scientific Computers,” St. Malo '88, University of California at Berkeley, pp. 346-353, 1988. |
| Butler Michael and Yale Patt, “An Improved Area-Efficient Register Alias Table for Implementing HPS,” University of Michigan, Ann Arbor, Michigan, pp. 1-15, Jan. 1990. |
| Uvieghara, G.A., W. Hwu, Y. Nakagome, D.K. Jeong, D. Lee, D.A. Hodges, Y. Patt, “An Experimental Single-Chip Data Flow CPU,” Symposium on ULSI Circuits Design Digest of Technical Papers, May 1990. |
| Melvin, Stephen and Yale Patt, “Exploiting Fine-Grained Parallesism Through a Combination of Hardware and Software Techniques,” Proceedings From ISCA-18, pp. 287-296, May 1990. |
| Butler, Michael, Tse-Yu Yeh, Yale Patt, Mitch Alsup, Hunter Scales and Michael Shebanow , “Single Instruction Stream Parallelism Is Greater Than Two,” Proceedings of ISCA-18, pp. 276-286, May 1990. |
| Uvieghara, Gregory A., Wen-Mei, W. Hwu, Yoshinobu Nakagome, Deog-Kyoon Jeong, David D. Lee, David A. Hodges and Yale Patt, “An Experimental Single-Chip Data Flow CPU,” IEEE Journal of Solid-State Circuits, vol. 27, No. 1, pp. 17-28, Jan. 1992. |
| Gee, Jeff, Stephen W. Melvin, Yale N. Patt, “The Implementation of Prolog via VAX 8600 Microcode,” Proceedings of Micro 19, New York City, pp. 1-7, Oct. 1986. |
| Hwu, Wen-Mei Hwu and Yale N. Patt, “Design Choices for the HPSm Microprocessor Chip,” Proceedings of the Twentieth Annual Hawaii International Conference on System Sciences, pp. 330-336, 1987. |
| Wilson, James E., Steve Melvin, Michael Shebanow, Wen-Mei Hwu and Yale N. Patt, “On Turning the Microarchitecture of an HPS Implementation of the VAX,” Proceedings of Micro 20, pp. 162-167, Dec. 1987. |
| Hwu, Wen-Mei and Yale N. Patt, “HPSm2: A Refined Single-Chip Microengine,” HICSS '88, pp. 30-40, 1988. |
| Butler, Michael and Yale Patt, “An Investigation of the Performance of Various Dynamic Scheduling Techniques,” Proceedings from MICRO-25, Dec. 1-4, 1992, pp. 1-9. |
| Kateveris, Hardware Support “Thesis,” 1984, p. 138-145. |
| Hennessy, John L. et al., “Computer Architecture A Quantitative Approach,” Ch. 6.4, 6.7 and p. 449, 1990. |
| Lightner, Bruce D. et al., “The Metaflow Lightning Chipset,” p. 13, 14 and 16, 1991, IEEE Publication. |