Parallel graph-oriented applications expressed in the Bulk-Synchronous Parallel (BSP) and Token Dataflow compute models generate highly-structured communication workloads from messages propagating along graph edges. We can statially expose this structure to traffic compilers and optimization tools to reshape and reduce traffic for higher performance (or lower area, lower energy, lower cost). Such offline traffic optimization eliminates the need for complex, runtime NoC hardware and enables lightweight, scalable NoCs. We perform load balancing, placement, fanout routing, and fine-grained synchronization to optimize our workloads for large networks up to 2025 parallel elements for BSP model and 25 parallel elements for Token Dataflow. This al...
Abstract—As the number of cores and threads in manycore compute accelerators such as Graphics Proces...
Abstract—In this paper, a NoC traffic monitoring method is proposed for billion cycle application de...
International audienceAs the key interconnection technique of System on Chip (SoC), Network on Chip ...
Parallel graph-oriented applications expressed in the Bulk-Synchronous Parallel (BSP) and Token Data...
Parallel graph-oriented applications expressed in the Bulk-Synchronous Parallel (BSP) and Token Data...
Sparse graph problems are notoriously hard to accelerate on conventional platforms due to irregular ...
FPGA-based soft processors customized for operations on sparse graphs can deliver significant perfor...
Abstract—As the number of cores and threads in manycore compute accelerators such as Graphics Proces...
Graduation date: 2017General-purpose Graphics Processing Units (GPGPUs) have become a critical compo...
Chip multiprocessors (CMPs) combine increasingly many general-purpose processor cores on a single ch...
How do we develop programs that are easy to express, easy to reason about, and able to achieve high ...
International audienceThe ever increasing density of integration makes the NoC a relevant communicat...
2018-10-16Graph analytics has drawn much research interest because of its broad applicability from m...
FPGA-based token dataflow architectures with heterogeneous computation and communication subsystems ...
As benchmark programs for microprocessor architectures, network-on-chip (NoC) traffic patterns are e...
Abstract—As the number of cores and threads in manycore compute accelerators such as Graphics Proces...
Abstract—In this paper, a NoC traffic monitoring method is proposed for billion cycle application de...
International audienceAs the key interconnection technique of System on Chip (SoC), Network on Chip ...
Parallel graph-oriented applications expressed in the Bulk-Synchronous Parallel (BSP) and Token Data...
Parallel graph-oriented applications expressed in the Bulk-Synchronous Parallel (BSP) and Token Data...
Sparse graph problems are notoriously hard to accelerate on conventional platforms due to irregular ...
FPGA-based soft processors customized for operations on sparse graphs can deliver significant perfor...
Abstract—As the number of cores and threads in manycore compute accelerators such as Graphics Proces...
Graduation date: 2017General-purpose Graphics Processing Units (GPGPUs) have become a critical compo...
Chip multiprocessors (CMPs) combine increasingly many general-purpose processor cores on a single ch...
How do we develop programs that are easy to express, easy to reason about, and able to achieve high ...
International audienceThe ever increasing density of integration makes the NoC a relevant communicat...
2018-10-16Graph analytics has drawn much research interest because of its broad applicability from m...
FPGA-based token dataflow architectures with heterogeneous computation and communication subsystems ...
As benchmark programs for microprocessor architectures, network-on-chip (NoC) traffic patterns are e...
Abstract—As the number of cores and threads in manycore compute accelerators such as Graphics Proces...
Abstract—In this paper, a NoC traffic monitoring method is proposed for billion cycle application de...
International audienceAs the key interconnection technique of System on Chip (SoC), Network on Chip ...