期刊文献+
共找到152篇文章
< 1 2 8 >
每页显示 20 50 100
FTCSEM—A FORTRAN-based parallelized 1D CSEM forward and inversion program for arbitrary source-receiver geometry 认领 引用
1
作者 Wei-ying Chen Si-xu Han +2 位作者 Wan-ting Song Yu-lian Zhu Zheng Liu 《Applied Geophysics》 SCIE CSCD 2026年第2期693-709,870,共17页
This study introduces FTCSEM,a FORTRAN-based,parallelized one-dimensional controlledsource electromagnetic(CSEM)forward modeling and inversion software capable of accommodating arbitrary source-receiver confi guration... This study introduces FTCSEM,a FORTRAN-based,parallelized one-dimensional controlledsource electromagnetic(CSEM)forward modeling and inversion software capable of accommodating arbitrary source-receiver confi gurations.In comparison to existing one-dimensional CSEM tools,FTCSEM incorporates several signifi cant enhancements:it supports transmitters of diverse shapes,quantities,and spatial locations;permits receivers to be positioned flexibly on the surface,subsurface,or in the atmosphere;facilitates simulations and inversions in both frequency and time domains;integrates an adaptive regularized inversion algorithm with multiple model constraints;and leverages GPU-accelerated parallel computing to attain high computational efficiency.Validation through numerical experiments and field data inversion confirms the program’s accuracy and practical applicability.The findings indicate that FTCSEM performs robustly in complex geoelectric environments,multi-source and multi-receiver arrangements,as well as multi-component joint inversion scenarios,thereby offering a versatile and powerful tool for advancing CSEM research and applications. 展开更多
关键词 CSEM forward modeling regularized inversion parallel computing program
暂未订购 下载PDF
Optimization Techniques for GPU-Based Parallel Programming Models in High-Performance Computing 认领 引用
2
作者 Shuntao Tang Wei Chen 《信息工程期刊(中英文版)》 2024年第1期7-11,共5页
This study embarks on a comprehensive examination of optimization techniques within GPU-based parallel programming models,pivotal for advancing high-performance computing(HPC).Emphasizing the transition of GPUs from g... This study embarks on a comprehensive examination of optimization techniques within GPU-based parallel programming models,pivotal for advancing high-performance computing(HPC).Emphasizing the transition of GPUs from graphic-centric processors to versatile computing units,it delves into the nuanced optimization of memory access,thread management,algorithmic design,and data structures.These optimizations are critical for exploiting the parallel processing capabilities of GPUs,addressingboth the theoretical frameworks and practical implementations.By integrating advanced strategies such as memory coalescing,dynamic scheduling,and parallel algorithmic transformations,this research aims to significantly elevate computational efficiency and throughput.The findings underscore the potential of optimized GPU programming to revolutionize computational tasks across various domains,highlighting a pathway towards achieving unparalleled processing power and efficiency in HPC environments.The paper not only contributes to the academic discourse on GPU optimization but also provides actionable insights for developers,fostering advancements in computational sciences and technology. 展开更多
关键词 Optimization Techniques GPU-Based Parallel Programming Models High-Performance Computing
Scheduling Step-Deteriorating Jobs on Parallel Machines by Mixed Integer Programming 认领 引用 被引量:5
3
作者 郭鹏 程文明 +1 位作者 曾鸣 梁剑 《Journal of Donghua University(English Edition)》 EI CAS 2015年第5期709-714,719,共6页
Production scheduling has a major impact on the productivity of the manufacturing process. Recently, scheduling problems with deteriorating jobs have attracted increasing attentions from researchers. In many practical... Production scheduling has a major impact on the productivity of the manufacturing process. Recently, scheduling problems with deteriorating jobs have attracted increasing attentions from researchers. In many practical situations,it is found that some jobs fail to be processed prior to the pre-specified thresholds,and they often consume extra deteriorating time for successful accomplishment. Their processing times can be characterized by a step-wise function. Such kinds of jobs are called step-deteriorating jobs. In this paper,parallel machine scheduling problem with stepdeteriorating jobs( PMSD) is considered. Due to its intractability,four different mixed integer programming( MIP) models are formulated for solving the problem under consideration. The study aims to investigate the performance of these models and find promising optimization formulation to solve the largest possible problem instances. The proposed four models are solved by commercial software CPLEX. Moreover,the near-optimal solutions can be obtained by black-box local-search solver LocalS olver with the fourth one. The computational results show that the efficiencies of different MIP models depend on the distribution intervals of deteriorating thresholds, and the performance of LocalS olver is clearly better than that of CPLEX in terms of the quality of the solutions and the computational time. 展开更多
关键词 parallel machine step-deterioration mixed integer programming(MIP) scheduling models total completion time
暂未订购 下载PDF
The parallel 3D magnetotelluric forward modeling algorithm 认领 引用 被引量:34
4
作者 Tan Handong Tong Tuo Lin Changhong 《Applied Geophysics》 2006年第4期197-202,共6页
The workload of the 3D magnetotelluric forward modeling algorithm is so large that the traditional serial algorithm costs an extremely large compute time. However, the 3D forward modeling algorithm can process the dat... The workload of the 3D magnetotelluric forward modeling algorithm is so large that the traditional serial algorithm costs an extremely large compute time. However, the 3D forward modeling algorithm can process the data in the frequency domain, which is very suitable for parallel computation. With the advantage of MPI and based on an analysis of the flow of the 3D magnetotelluric serial forward algorithm, we suggest the idea of parallel computation and apply it. Three theoretical models are tested and the execution efficiency is compared in different situations. The results indicate that the parallel 3D forward modeling computation is correct and the efficiency is greatly improved. This method is suitable for large size geophysical computations. 展开更多
关键词 Magnetotelluric 3D forward modeling MPI parallel programming design 3D staggered-grid finite difference method parallel algorithm.
暂未订购 下载PDF
Parallel Machine Scheduling Models with Fuzzy Parameters and Precedence Constraints: A Credibility Approach 认领 引用
5
作者 侯福均 吴祈宗 《Journal of Beijing Institute of Technology》 EI CAS 2007年第2期231-236,共6页
A method for modeling the parallel machine scheduling problems with fuzzy parameters and precedence constraints based on credibility measure is provided. For the given n jobs to be processed on m machines, it is assum... A method for modeling the parallel machine scheduling problems with fuzzy parameters and precedence constraints based on credibility measure is provided. For the given n jobs to be processed on m machines, it is assumed that the processing times and the due dates are nonnegative fuzzy numbers and all the weights are positive, crisp numbers. Based on credibility measure, three parallel machine scheduling problems and a goal-programming model are formulated. Feasible schedules are evaluated not only by their objective values but also by the credibility degree of satisfaction with their precedence constraints. The genetic algorithm is utilized to find the best solutions in a short period of time. An illustrative numerical example is also given. Simulation results show that the proposed models are effective, which can deal with the parallel machine scheduling problems with fuzzy parameters and precedence constraints based on credibility measure. 展开更多
关键词 parallel machine scheduling programming model possibility measure credibility measure fuzzy number genetic algorithm
暂未订购 下载PDF
基于Map-Reduce的自适应双语短语挖掘系统 认领 引用
6
作者 李彬 杨世泉 陈文杰 《昆明学院学报》 2013年第3期83-87,共5页
对于跨语言信息检索,统计翻译等应用,双语短语都是极其重要的资源.提出了基于自适应模式的双语短语挖掘算法,该算法可以自动的学习当前Web页面的翻译模式,然后利用学习到的模式抽取当前页面中的双语短语.同时,将自适应双语短语挖掘算法... 对于跨语言信息检索,统计翻译等应用,双语短语都是极其重要的资源.提出了基于自适应模式的双语短语挖掘算法,该算法可以自动的学习当前Web页面的翻译模式,然后利用学习到的模式抽取当前页面中的双语短语.同时,将自适应双语短语挖掘算法与Map-Reduce并行编程模型融合起来,大大提高了系统的运行效率,并且通过实验验证了该方法的有效性. 展开更多
关键词 自适应模式 双语短语 Map-Reduce并行计算框架 分布式计算
暂未订购 下载PDF
Parallel Pipelines for DNA Sequence Alignment on a Cluster of Multicores:A Comparison of Communication Models 认领 引用
7
作者 Enzo Rucci Franco Chichizola +2 位作者 Marcelo Naiouf Laura De Giusti Armando De Giusti 《通讯和计算机(中英文版)》 2012年第12期1364-1371,共8页
HPC(high perfomance computing)based on clusters of multicores is one of the main research lines in parallel programming.It is important to study the impact of programming paradigms of shared memory,message passing or ... HPC(high perfomance computing)based on clusters of multicores is one of the main research lines in parallel programming.It is important to study the impact of programming paradigms of shared memory,message passing or a combination of both on these architectures in order to efficiently exploit the power of these architectures.The Smith-Waterman algorithm is used as study case for the local alignment of DNA sequences,which allows establishing the similarity degree between two sequences.In this paper,the Smith-Waterman algorithm is parallelized by means of a pipeline scheme due to the data dependencies that are inherent to the problem,using the various communication/synchronization models mentioned above and then carrying out a comparative analysis.Finally,experimental results are presented,as well as future research lines. 展开更多
关键词 Cluster of multicores communication models parallel programming pipeline Smith-Waterman
暂未订购 下载PDF
Parallel programming models for heterogeneous many‑cores:a comprehensive survey 认领 引用 被引量:10
8
作者 Jianbin Fang Chun Huang +1 位作者 Tao Tang Zheng Wang 《CCF Transactions on High Performance Computing》 EI 2020年第4期382-400,共19页
Heterogeneous many-cores are now an integral part of modern computing systems ranging from embedding systems to supercomputers.While heterogeneous many-core design offers the potential for energy-efficient high-perfor... Heterogeneous many-cores are now an integral part of modern computing systems ranging from embedding systems to supercomputers.While heterogeneous many-core design offers the potential for energy-efficient high-performance,such potential can only be unlocked if the application programs are suitably parallel and can be made to match the underlying heterogeneous platform.In this article,we provide a comprehensive survey for parallel programming models for heterogeneous many-core architectures and review the compiling techniques of improving programmability and portability.We examine various software optimization techniques for minimizing the communicating overhead between heterogeneous computing devices.We provide a road map for a wide variety of different research areas.We conclude with a discussion on open issues in the area and potential research directions.This article provides both an accessible introduction to the fast-moving area of heterogeneous programming and a detailed bibliography of its main achievements. 展开更多
关键词 Heterogeneous computing Many-core architectures Parallel programming models
基于国产编程语言的并行水动力模型开发及初步调优 认领 引用
9
作者 王明阳 王静 +2 位作者 李娜 俞茜 宫啸天 《人民黄河》 CAS 北大核心 2026年第2期41-46,共6页
基于国产编程语言Taichi开发了具有跨平台并行计算能力的高性能二维水动力模型FRAS。FRAS拥有良好的并行计算灵活性,能与同构CPU-CPU和异构CPU-GPU计算架构良好兼容,还支持多核CPU、CUDA、OpenGL、Metal、Vulkan等多种并行加速技术,具... 基于国产编程语言Taichi开发了具有跨平台并行计算能力的高性能二维水动力模型FRAS。FRAS拥有良好的并行计算灵活性,能与同构CPU-CPU和异构CPU-GPU计算架构良好兼容,还支持多核CPU、CUDA、OpenGL、Metal、Vulkan等多种并行加速技术,具有良好的跨平台性能。采用非结构化网格离散二维空间,运用有限体积法对连续性方程和动量方程进行数值离散处理,将FRAS模型应用于辽宁省绕阳河的洪水计算,相较于原始串行代码,并行化处理后加速比为14.7。通过优化变量存储结构,计算性能因访存优化而提升约2倍,初步优化后程序加速比达30.1。 展开更多
关键词 二维水动力模型 并行计算 跨平台 Taichi编程语言
暂未订购 下载PDF
Parallel Implementations of Modeling Dynamical Systems by Using System of Ordinary Differential Equations 认领 引用
10
作者 Cao Hong-qing, Kang Li-shan, Yu Jing-xianState Key Laboratory of Software Engineering, Wuhan University, Wuhan 430072,Hubei,ChinaCollege of Chemistry and Molecular Sciences, Wuhan University, Wuhan 430072, Hubei, China 《Wuhan University Journal of Natural Sciences》 EI CAS 2003年第S1期229-233,共5页
First, an asynchronous distributed parallel evolutionary modeling algorithm (PEMA) for building the model of system of ordinary differential equations for dynamical systems is proposed in this paper. Then a series of ... First, an asynchronous distributed parallel evolutionary modeling algorithm (PEMA) for building the model of system of ordinary differential equations for dynamical systems is proposed in this paper. Then a series of parallel experiments have been conducted to systematically test the influence of some important parallel control parameters on the performance of the algorithm. A lot of experimental results are obtained and we make some analysis and explanations to them. 展开更多
关键词 parallel genetic programming evolutionary modeling system of ordinary differential equations
暂未订购 下载PDF
基于强化学习与遗传算法的机器人并行拆解序列规划方法 认领 引用 被引量:4
11
作者 汪开普 马晓艺 +2 位作者 卢超 殷旅江 李新宇 《国防科技大学学报》 EI CAS CSCD 北大核心 2025年第2期24-34,共11页
在拆解序列规划问题中,为了提高拆解效率、降低拆解能耗,引入了机器人并行拆解模式,构建了机器人并行拆解序列规划模型,并设计了基于强化学习的遗传算法。为了验证模型的正确性,构造了混合整数线性规划模型。算法构造了基于目标导向的... 在拆解序列规划问题中,为了提高拆解效率、降低拆解能耗,引入了机器人并行拆解模式,构建了机器人并行拆解序列规划模型,并设计了基于强化学习的遗传算法。为了验证模型的正确性,构造了混合整数线性规划模型。算法构造了基于目标导向的编解码策略,以提高初始解的质量;采用Q学习来选择算法迭代过程中的最佳交叉策略和变异策略,以增强算法的自适应能力。在一个34项任务的发动机拆解案例中,通过与四种经典多目标算法对比,验证了所提算法的优越性;分析所得拆解方案,结果表明机器人并行拆解模式可以有效缩短完工时间,并降低拆解能耗。 展开更多
关键词 拆解序列规划 机器人并行拆解 混合整数线性规划模型 遗传算法 强化学习
暂未订购 下载PDF
基于开源程序的大跨桥梁模型更新 认领 引用 被引量:3
12
作者 郑俊浩 王达荣 +1 位作者 管仲国 林楷奇 《工程力学》 EI CSCD 北大核心 2025年第3期181-190,共10页
大跨桥梁作为重要的基础工程设施,准确评估其即时服役状态、建立高保真数值模型是提升城市交通防灾水平的重要途径,实现上述目标依赖于高效可靠的复杂结构模型更新与分析技术。但已有研究多依赖商业软件平台,存在价格昂贵、算法更新速... 大跨桥梁作为重要的基础工程设施,准确评估其即时服役状态、建立高保真数值模型是提升城市交通防灾水平的重要途径,实现上述目标依赖于高效可靠的复杂结构模型更新与分析技术。但已有研究多依赖商业软件平台,存在价格昂贵、算法更新速度较慢和软件接口复杂等局限性,限制了相关研究的深入发展。因此,该文基于开源有限元分析平台OpenSees和编程语言Python,开发了适用于复杂工程结构模型更新的开源程序框架。基于Python开发了工程结构模型更新所需的高效并行优化算法,进一步编写不同功能模块的软件接口,连接模型分析平台与并行优化算法,实现复杂工程结构模型的并行分析和模型更新。在此基础上,以一个简支梁损伤识别为例,验证了上述计算程序的有效性与计算精度。以苏通大桥振动台试验缩尺模型为研究对象,采用模态置信准则匹配有限元模型和振动台试验获得的模态振型和频率数据,构建模型更新所需的优化函数,采用高性能计算平台,开展试验桥梁的模型更新并验证上述程序框架的计算效率。结果表明,该框架可以实现大跨斜拉桥精细有限元模型的高效更新,匹配实测获得的结构模态数据,各阶计算频率与试验模型的误差在1%以下。进一步地,计算更新后的数值模型在PGA=0.1 g地震动作用下的动力时程响应,与试验数据对比验证了模型更新结果的准确性。研究成果可以为基于开源平台的复杂大跨桥梁的精细化与数据驱动建模提供参考。 展开更多
关键词 开源程序框架 并行优化算法 大跨度桥梁 精细化有限元模型 模型更新
暂未订购 下载PDF
A Multi-Criteria Decision Making for the Unrelated Parallel Machines Scheduling Problem 认领 引用
13
作者 Wei-Shung CHANG Chiuh-Cheng CHYU 《Journal of Software Engineering and Applications》 2009年第5期323-329,共7页
In this paper, we propose a multi-criteria machine-schedules decision making method that can be applied to a produc-tion environment involving several unrelated parallel machines and we will focus on three objectives:... In this paper, we propose a multi-criteria machine-schedules decision making method that can be applied to a produc-tion environment involving several unrelated parallel machines and we will focus on three objectives: minimizing makespan, total flow time, and total number of tardy jobs. The decision making method consists of three phases. In the first phase, a mathematical model of a single machine scheduling problem, of which the objective is a weighted sum of the three objectives, is constructed. Such a model will be repeatedly solved by the CPLEX in the proposed Multi-Objective Simulated Annealing (MOSA) algorithm. In the second phase, the MOSA that integrates job clustering method, job group scheduling method, and job group – machine assignment method, is employed to obtain a set of non-dominated group schedules. During this phase, CPLEX software and the bipartite weighted matching algorithm are used repeatedly as parts of the MOSA algorithm. In the last phase, the technique of data envelopment analysis is applied to determine the most preferable schedule. A practical example is then presented in order to demonstrate the applicability of the proposed decision making method. 展开更多
关键词 Multi-Objective Optimization Unrelated Parallel Machines Scheduling Simulated Annealing Algorithm Integer Programming Models Multi-Criteria Decision Making
暂未订购 下载PDF
Programming bare-metal accelerators with heterogeneous threading models:a case study of Matrix-3000 认领 引用 被引量:6
14
作者 Jianbin FANG Peng ZHANG +4 位作者 Chun HUANG Tao TANG Kai LU Ruibo WANG Zheng WANG 《Frontiers of Information Technology & Electronic Engineering》 SCIE EI CSCD 2023年第4期509-520,共12页
As the hardware industry moves toward using specialized heterogeneous many-core processors to avoid the effects of the power wall,software developers are finding it hard to deal with the complexity of these systems.In... As the hardware industry moves toward using specialized heterogeneous many-core processors to avoid the effects of the power wall,software developers are finding it hard to deal with the complexity of these systems.In this paper,we share our experience of developing a programming model and its supporting compiler and libraries for Matrix-3000,which is designed for next-generation exascale supercomputers but has a complex memory hierarchy and processor organization.To assist its software development,we have developed a software stack from scratch that includes a low-level programming interface and a high-level OpenCL compiler.Our low-level programming model offers native programming support for using the bare-metal accelerators of Matrix-3000,while the high-level model allows programmers to use the OpenCL programming standard.We detail our design choices and highlight the lessons learned from developing system software to enable the programming of bare-metal accelerators.Our programming models have been deployed in the production environment of an exascale prototype system. 展开更多
关键词 Heterogeneous computing Parallel programming models Programmability Compilers Runtime systems
暂未订购 下载PDF
一种异构多核系统动态调度协处理器设计 认领 引用 被引量:1
15
作者 曾树铭 倪伟 《合肥工业大学学报(自然科学版)》 CAS 北大核心 2025年第2期185-195,共11页
为研究异构多核片上系统(multi-processor system on chip,MPSoC)在密集并行计算任务中的潜力,文章设计并实现了一种适用于粗粒度数据特征、面向任务级并行应用的异构多核系统动态调度协处理器,采用了片上缓存、任务输出的多级写回管理... 为研究异构多核片上系统(multi-processor system on chip,MPSoC)在密集并行计算任务中的潜力,文章设计并实现了一种适用于粗粒度数据特征、面向任务级并行应用的异构多核系统动态调度协处理器,采用了片上缓存、任务输出的多级写回管理、任务自动映射、通讯任务乱序执行等机制。实验结果表明,该动态调度协处理器不仅能够实现任务级乱序执行等基本设计目标,还具有极低的调度开销,相较于基于动态记分牌算法的调度器,运行多个子孔径距离压缩算法的时间降低达17.13%。研究结果证明文章设计的动态调度协处理器能够有效优化目标场景下的任务调度效果。 展开更多
关键词 动态调度 硬件调度器 异构多核系统 任务级并行 编程模型 片上缓存 片上网络
暂未订购 下载PDF
AceMesh:a structured data driven programming language for high performance computing 认领 引用 被引量:5
16
作者 Li Chen Shenglin Tang +3 位作者 You Fu Xiran Gao Jie Guo Shangzhi Jiang 《CCF Transactions on High Performance Computing》 EI 2020年第4期309-322,共14页
Asynchronous task-based programming models are gaining popularity to address the programmability and performance challenges of contemporary large scale high performance computing systems.In this paper we present AceMe... Asynchronous task-based programming models are gaining popularity to address the programmability and performance challenges of contemporary large scale high performance computing systems.In this paper we present AceMesh,a taskbased,data-driven language extension targeting legacy MPI applications.Its language features include data-centric parallelizing template,aggregated task dependence for parallel loops.These features not only relieve the programmer from tedious refactoring details but also provide possibility for structured execution of complex task graphs,data locality exploitation upon data tile templates,and reducing system complexity incurred by complex array sections.We present the prototype implementation,including task shifting,data management and communication-related analysis and transformations.The language extension is evaluated on two supercomputing platforms.We compare the performance of AceMesh with existing programming models,and the results show that NPB/MG achieves at most 1.2X and 1.85X speedups on TaihuLight and TH-2,respectively,and the Tend_lin benchmark attains more than 2X speedup on average and attain at most 3.0X and 2.2X speedups on the two platforms,respectively. 展开更多
关键词 High performance computing Programming model MPI Task parallel Data driven Task dependence
任务并行编程模型研究与进展 认领 引用 被引量:31
17
作者 王蕾 崔慧敏 +1 位作者 陈莉 冯晓兵 《软件学报》 EI CSCD 北大核心 2013年第1期77-90,共14页
任务并行编程模型是近年来多核平台上广泛研究和使用的并行编程模型,旨在简化并行编程和提高多核利用率.首先,介绍了任务并行编程模型的基本编程接口和支持机制;然后,从3个角度,即并行性表达、数据管理和任务调度介绍任务并行编程模型... 任务并行编程模型是近年来多核平台上广泛研究和使用的并行编程模型,旨在简化并行编程和提高多核利用率.首先,介绍了任务并行编程模型的基本编程接口和支持机制;然后,从3个角度,即并行性表达、数据管理和任务调度介绍任务并行编程模型的研究问题、困难和最新研究成果;最后展望了任务并行未来的研究方向. 展开更多
关键词 任务并行 并行编程模型 任务窃取调度 并行性表达
暂未订购 下载PDF
MapReduce并行编程模型研究综述 认领 引用 被引量:191
18
作者 李建江 崔健 +2 位作者 王聃 严林 黄义双 《电子学报》 EI CAS CSCD 北大核心 2011年第11期2635-2642,共8页
MapReduce并行编程模型通过定义良好的接口和运行时支持库,能够自动并行执行大规模计算任务,隐藏底层实现细节,降低并行编程的难度.本文对MapReduce的国内外相关研究现状进行了综述,阐述和分析了当前国内外与MapReduce相关的典型研究成... MapReduce并行编程模型通过定义良好的接口和运行时支持库,能够自动并行执行大规模计算任务,隐藏底层实现细节,降低并行编程的难度.本文对MapReduce的国内外相关研究现状进行了综述,阐述和分析了当前国内外与MapReduce相关的典型研究成果的特点和不足,重点对MapReduce涉及的关键技术(包括:模型改进、模型针对不同平台的实现、任务调度、负载均衡和容错)的研究现状进行了深入的分析.本文最后还对MapReduce未来的发展趋势进行了展望. 展开更多
关键词 MapReduce 并行编程模型 运行时支持库 海量数据处理
暂未订购 下载PDF
基于MapReduce模型的并行科学计算 认领 引用 被引量:39
19
作者 郑启龙 房明 +3 位作者 汪胜 王向前 吴晓伟 王昊 《微电子学与计算机》 北大核心 2009年第8期13-17,共5页
随着多核处理器日渐普及,开发高效易用的并行编程模型成为新的挑战.MapReduce是Google开发的一种并行分布式计算模型,在其搜索业务中获得了巨大的成功.将MapReduce模型引入科学计算领域,并结合实例阐述了如何使用面向高性能计算的HPMR/H... 随着多核处理器日渐普及,开发高效易用的并行编程模型成为新的挑战.MapReduce是Google开发的一种并行分布式计算模型,在其搜索业务中获得了巨大的成功.将MapReduce模型引入科学计算领域,并结合实例阐述了如何使用面向高性能计算的HPMR/HPMR-s系统在分布式或共享存储系统中采用统一的方式描述并实现并行科学计算. 展开更多
关键词 并行编程模型 科学计算 MapReduce
暂未订购 下载PDF
异构并行编程模型研究与进展 认领 引用 被引量:13
20
作者 刘颖 吕方 +3 位作者 王蕾 陈莉 崔慧敏 冯晓兵 《软件学报》 EI CSCD 北大核心 2014年第7期1459-1475,共17页
近年来,异构系统硬件飞速发展.为了解决相应的编程和执行效率问题,异构并行编程模型已被广泛使用和研究.从异构并行编程接口与编译/运行时支持系统两个角度总结了异构并行编程模型最新的研究成果,它们为异构架构和上层应用带来的技术挑... 近年来,异构系统硬件飞速发展.为了解决相应的编程和执行效率问题,异构并行编程模型已被广泛使用和研究.从异构并行编程接口与编译/运行时支持系统两个角度总结了异构并行编程模型最新的研究成果,它们为异构架构和上层应用带来的技术挑战提供了相应的解决方案.最后,结合目前的研究现状以及异构系统的发展,提出了异构并行编程模型的未来方向. 展开更多
关键词 异构并行编程模型 异构系统 GPU 编程接口 编译 运行时系统
暂未订购 下载PDF
上一页 1 2 8 下一页 到第
在线咨询 使用帮助 返回顶部 意见反馈