Looking for breakthrough ideas for innovation challenges? Try Patsnap Eureka!

Selectively deferring instructions issued in program order utilizing a checkpoint and instruction deferral scheme

a program order and instruction deferral technology, applied in the field of computer system performance improvement, can solve the problems of microprocessor systems spending a large fraction of time waiting, unable to meet the requirements of program order, so as to facilitate extending out-of-order execution, save chip resources, and facilitate the effect of renaming unit complexity

Inactive Publication Date: 2006-11-30
CHAUDHRY SHAILENDER +1
View PDF9 Cites 0 Cited by
  • Summary
  • Abstract
  • Description
  • Claims
  • Application Information

AI Technical Summary

Benefits of technology

This approach effectively hides memory latency without the complexity of traditional out-of-order designs, allowing for efficient execution of thousands of instructions within the same area as a traditional out-of-order issue queue, reducing chip resources and eliminating the need for register renaming.

Problems solved by technology

Hence, the disparity between microprocessor clock speeds and memory access speeds continues to grow, and is beginning to create significant performance problems.
This means that the microprocessor systems spend a large fraction of time waiting for memory references to complete instead of performing computational operations.
However, when a memory reference, such as a load operation generates a cache miss, the subsequent access to level-two (L2) cache or memory can require dozens or hundreds of clock cycles to complete, during which time the processor is typically idle, performing no useful work.
Unfortunately, existing out-of-order designs have a hardware complexity that grows quadratically with the size of the issue queue.
Practically speaking, this constraint limits the number of entries in the issue queue to one or two hundred, which is not sufficient to hide memory latencies as processors continue to get faster.
Moreover, constraints on the number of physical registers that are available for register renaming purposes during out-of-order execution also limits the effective size of the issue queue.
However, it suffers from the disadvantage of having to re-compute any computational operations that are performed while in scout-ahead mode.

Method used

the structure of the environmentally friendly knitted fabric provided by the present invention; figure 2 Flow chart of the yarn wrapping machine for environmentally friendly knitted fabrics and storage devices; image 3 Is the parameter map of the yarn covering machine
View more

Image

Smart Image Click on the blue labels to locate them in the text.
Viewing Examples
Smart Image
  • Selectively deferring instructions issued in program order utilizing a checkpoint and instruction deferral scheme
  • Selectively deferring instructions issued in program order utilizing a checkpoint and instruction deferral scheme
  • Selectively deferring instructions issued in program order utilizing a checkpoint and instruction deferral scheme

Examples

Experimental program
Comparison scheme
Effect test

Embodiment Construction

[0026] The following description is presented to enable any person skilled in the art to make and use the invention, and is provided in the context of a particular application and its requirements. Various modifications to the disclosed embodiments will be readily apparent to those skilled in the art, and the general principles defined herein may be applied to other embodiments and applications without departing from the spirit and scope of the present invention. Thus, the present invention is not intended to be limited to the embodiments shown, but is to be accorded the widest scope consistent with the principles and features disclosed herein.

Processor

[0027]FIG. 1A illustrates the design of a processor 100 in accordance with an embodiment of the present invention. Processor 100 can generally include any type of processor, including, but not limited to, a microprocessor, a mainframe computer, a digital signal processor, a personal organizer, a device controller and a computationa...

the structure of the environmentally friendly knitted fabric provided by the present invention; figure 2 Flow chart of the yarn wrapping machine for environmentally friendly knitted fabrics and storage devices; image 3 Is the parameter map of the yarn covering machine
Login to View More

PUM

No PUM Login to View More

Abstract

One embodiment of the present invention provides a system that facilitates deferring execution of instructions with unresolved data dependencies as they are issued for execution in program order. During a normal execution mode, the system issues instructions for execution in program order. Upon encountering an unresolved data dependency during execution of an instruction, the system generates a checkpoint that can subsequently be used to return execution of the program to the point of the instruction. Next, the system executes subsequent instructions in an execute-ahead mode, wherein instructions that cannot be executed because of an unresolved data dependency are deferred, and wherein other non-deferred instructions are executed in program order.

Description

BACKGROUND [0001] 1. Field of the Invention [0002] The present invention relates to techniques for improving the performance of computer systems. More specifically, the present invention relates to a method and an apparatus for speeding up program execution by selectively deferring execution of instructions with unresolved data-dependencies as they are issued for execution in program order. [0003] 2. Related Art [0004] Advances in semiconductor fabrication technology have given rise to dramatic increases in microprocessor clock speeds. This increase in microprocessor clock speeds has not been matched by a corresponding increase in memory access speeds. Hence, the disparity between microprocessor clock speeds and memory access speeds continues to grow, and is beginning to create significant performance problems. Execution profiles for fast microprocessor systems show that a large fraction of execution time is spent not within the microprocessor core, but within memory structures outs...

Claims

the structure of the environmentally friendly knitted fabric provided by the present invention; figure 2 Flow chart of the yarn wrapping machine for environmentally friendly knitted fabrics and storage devices; image 3 Is the parameter map of the yarn covering machine
Login to View More

Application Information

Patent Timeline
no application Login to View More
Patent Type & Authority Applications(United States)
IPC IPC(8): G06F9/30G06F9/00G06F9/38G06F9/45
CPCG06F9/3836G06F9/3838G06F9/3842G06F9/384G06F9/3865G06F9/3855G06F9/3857G06F9/3863G06F9/3858G06F9/3856
Inventor CHAUDHRY, SHAILENDERTREMBLAY, MARC
Owner CHAUDHRY SHAILENDER
Who we serve
  • R&D Engineer
  • R&D Manager
  • IP Professional
Why Patsnap Eureka
  • Industry Leading Data Capabilities
  • Powerful AI technology
  • Patent DNA Extraction
Social media
Patsnap Eureka Blog
Learn More
PatSnap group products