Split transactions pipeline operations on a serial bus, reducing dummy cycles and doubling read throughput in NOR Flash systems.
A register scoreboard paired with a time counter dispatches instructions to an execution pipeline based on preset execution times.
Segments FIFO instructions into independent groups and reorders them based on path length metrics to resolve scheduling constraints.
A centralized update control unit coordinates timing between master and slave IP blocks to prevent synchronization delays in SoC data rendering.
A dual branch processing method routes primary and secondary branch instructions through distinct hardware paths to optimize microprocessor instruction throughput.
Dynamic instruction reordering and fusion reduce temporal latencies in processor pipelines.
Unified shader cores and hierarchical segmentation maximize parallel processing efficiency while managing coordination complexity.
A deep neural network training method uses score history to select hard negative samples for each epoch.
Register renaming enables atomic commit in reduced-width VLIW processors, lowering power consumption and cost while maintaining compiler compatibility.
Segmenting the dependence matrix into issue and load miss portions allows prompt instruction deallocation while maintaining consumer tracking.
A heuristic method invalidates non-useful entries in an operand store compare history table using a useless prediction counter.
A branch instruction processing system selects the highest accuracy prediction method from multiple preset approaches to improve processor efficiency.
Control logic compresses multiple instruction operations into single retire queue entries to improve storage efficiency.
Execution circuitry detects address hazards during vector loop iterations to adjust processing levels.
A multiprocessor system enforces a total ordering of computational operations to establish a consistent architectural state for precise restartability.
Merge circuitry uses bitwise logical operations on logging bit vectors to select completed transactions for output pipelines.
Pre-fetching candidate contexts into a buffer breaks data dependency chains, reducing latency during video frame decoding.