Shyam Marjit, Harshit Singh, et al.
WACV 2025
In this paper, we present several algorithms for performing all-to-many personalized communication on distributed memory parallel machines. We assume that each processor sends a different message (of potentially different size) to a subset of all the processors involved in the collective communication. The algorithms are based on decomposing the communication matrix into a set of partial permutations. We study the effectiveness of our algorithms from both the view of static scheduling and runtime scheduling. © 1995 Academic Press, Inc.
Shyam Marjit, Harshit Singh, et al.
WACV 2025
Yannis Belkhiter, Dhaval Salwala, et al.
NFV-SDN 2025
Guo-Jun Qi, Charu Aggarwal, et al.
IEEE TPAMI
Saeel Sandeep Nachane, Ojas Gramopadhye, et al.
EMNLP 2024