PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
March 3, 2026Journal of King Saud University - Computer and Information Sciences0 citationsOpen Access

A fully hardware-managed scheduling architecture for AI accelerators

View Full Paper
LCLibo ChengLYLiang YangJSJian Shao

Key Points

  • Our system significantly reduces scheduling latency, and improves overall throughput by about 30%.
  • The architecture utilizes operator completion status registers to reduce software overhead effectively.
  • Experimental validation shows minimal area overhead of only 7.41% while enhancing performance.
  • A dedicated hardware scheduling unit and a lightweight runtime work together to streamline task dispatch.

Abstract

The proliferation of AI applications across diverse domains has driven the evolution of AI accelerators toward higher performance and energy efficiency. This paper addresses the critical challenge of task scheduling in AI accelerators by introducing a fully hardware-managed scheduling system. Our approach leverages Operator Completion Status Registers (OCSRs) and a novel computation-scheduling instruction set to minimize software overhead and maximize execution parallelism. The co-designed hardware-software solution comprises: (1) a dedicated hardware scheduling unit with a complete instruction pipeline, (2) a compiler that maps operators to scheduling instructions while managing OCSR allocation, and (3) a lightweight runtime for efficient task dispatch. Experimental results demonstrate that our system significantly reduces scheduling latency and improves overall throughput, achieving an average performance gain of approximately 30% across multiple CNN models while maintaining minimal area overhead of only 7.41%. The proposed architecture establishes a new paradigm for high-efficiency AI accelerator design.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Cheng et al. (2026) studied this question.

synapsesocial.com/papers/69a7681bbadf0bb9e87e39dahttps://doi.org/10.1007/s44443-026-00513-z
Ask AI
Helpful
Bookmark
Share
View Full Paper