| 1 |
LUNDSTROM M S , ALAM M A . Moore's law: the journey ahead. Science, 2022, 378 (6621): 722- 723.
doi: 10.1126/science.ade2191
|
| 2 |
LEE E A . The problem with threads. Computer, 2006, 39 (5): 33- 42.
|
| 3 |
GROPP W , LUSK E , DOSS N , et al. A high-performance, portable implementation of the MPI message passing interface standard. Parallel Computing, 1996, 22 (6): 789- 828.
doi: 10.1016/0167-8191(96)00024-5
|
| 4 |
DAGUM L , MENON R . OpenMP: an industry standard API for shared-memory programming. IEEE Computational Science and Engineering, 1998, 5 (1): 46- 55.
doi: 10.1109/99.660313
|
| 5 |
SANDERS J , KANDROT E , SANDERS J , et al. CUDA by example: an introduction to general purpose GPU programming. Boston, USA: Addison-Wesley Professional, 2010.
|
| 6 |
STONE J E , GOHARA D , SHI G C . OpenCL: a parallel programming standard for heterogeneous computing systems. Computing in Science & Engineering, 2010, 12 (3): 66- 73.
|
| 7 |
DENNIS J B . Data flow supercomputers. Computer, 1980, 13 (11): 48- 56.
|
| 8 |
FLYNN M J . Some computer organizations and their effectiveness. IEEE Transactions on Computers, 1972, 100 (9): 948- 960.
|
| 9 |
JOHNSTON W M , PAUL H J R , MILLAR R J . Advances in dataflow programming languages. ACM Computing Surveys, 2004, 36 (1): 1- 34.
doi: 10.1145/1013208.1013209
|
| 10 |
OVIEDO E I. Control flow, data flow and program complexity[D]. Buffalo, USA: State University of New York at Buffalo, 1984.
|
| 11 |
BAUER M, TREICHLER S, SLAUGHTER E, et al. Legion: expressing locality and independence with logical regions[C]//Proceedings of the International Conference for High Performance Computing, Networking, Storage and Analysis. New York, USA: ACM Press, 2012: 1-11.
|
| 12 |
PHEATT C . Intel threading building blocks. Journal of Computing Sciences in Colleges, 2008, 23 (4): 298.
|
| 13 |
|
| 14 |
BEN-NUN T, FINE L J, ZIOGAS A N, et al. Stateful dataflow multigraphs: a data-centric model for performance portability on heterogeneous architectures[C]//Proceedings of the International Conference for High Performance Computing, Networking, Storage and Analysis. New York, USA: ACM Press, 2019: 1-14.
|
| 15 |
LATTNER C, AMINI M, BONDHUGULA U, et al. MLIR: scaling compiler infrastructure for domain specific computation[C]// Proceedings of the IEEE/ACM International Symposium on Code Generation and Optimization (CGO). Seoul, Korea: IEEE Press, 2021: 2-14.
|
| 16 |
BEN-NUN T, ATES B, CALOTOIU A, et al. Bridging control-centric and data-centric optimization[C]//Proceedings of the 21st ACM/IEEE International Symposium on Code Generation and Optimization. New York, USA: ACM Press, 2023: 173-185.
|
| 17 |
SUETTLERLEIN J, ZUCKERMAN S, GAO G R. An implementation of the Codelet model[C]//Proceedings of Euro-Par 2013. Berlin, Germany: Springer, 2013: 633-644.
|
| 18 |
SUETTERLEIN J. DARTS: a runtime based on the Codelet execution model[D]. Newark, USA: University of Delaware, 2014.
|
| 19 |
|
| 20 |
MOSES W S, CHELINI L, ZHAO R Z, et al. Polygeist: raising C to polyhedral MLIR[C]//Proceedings of the 30th International Conference on Parallel Architectures and Compilation Techniques (PACT). Atlanta, USA: IEEE Press, 2021: 45-59.
|
| 21 |
|
| 22 |
CHEN L , TANG S L , FU Y , et al. AceMesh: a structured data driven programming language for high performance computing. CCF Transactions on High Performance Computing, 2020, 2 (4): 309- 322.
doi: 10.1007/s42514-020-00047-4
|
| 23 |
BOSILCA G , BOUTEILLER A , DANALIS A , et al. PaRSEC: exploiting heterogeneity to enhance scalability. Computing in Science & Engineering, 2013, 15 (6): 36- 45.
|
| 24 |
KABRICK R, PERDOMO D A R, RASKAR S, et al. CODIR: towards an MLIR codelet model dialect[C]// Proceedings of the 4th Annual Workshop on Emerging Parallel and Distributed Runtime Systems and Middleware (IPDRM). Washington D. C., USA: IEEE Press, 2020: 33-40.
|
| 25 |
李金熹, 尹首一, 魏少军, 等. 基于MLIR的数据流模型. 计算机工程与科学, 2024, 46 (7): 1151- 1157.
|
|
LI J X , YIN S Y , WEI S J , et al. A codelet model based on MLIR. Computer Engineering and Science, 2024, 46 (7): 1151- 1157.
|