检索结果-内蒙古大学图书馆

您好，读者！请登录

内蒙古大学图书馆

首页
概况
党建
资源
服务
科研支持
- 论文收录引用证明
- 科技查新
知识产权
档案馆
帮助

咨询与建议

建议与咨询留下您的常用邮箱和电话号码，以便我们向您反馈解决方案和替代方法

您的常用邮箱：*

您的手机号码：*

问题描述：

当前已输入0个字，您还可以输入200个字

全部搜索
期刊论文
图书
学位论文
标准
纸本馆藏
外文资源发现
数据库导航
超星发现

高级检索

分类表

所选分类

>> <<

限定检索结果

标题

标题
作者
主题词
出版物名称
出版社
机构
学科分类号
摘要
ISBN
ISSN
基金资助
索书号

作者

作者
标题
主题词
出版物名称
出版社
机构
学科分类号
摘要
ISBN
ISSN
基金资助
索书号

文献类型

828 篇 会议
294 篇 期刊文献
2 册 图书

馆藏范围

1,124 篇 电子文献
0 种 纸本馆藏

日期分布

学科分类号

712 篇 工学
- 556 篇 计算机科学与技术...
- 411 篇 软件工程
- 136 篇 信息与通信工程
- 82 篇 控制科学与工程
- 74 篇 电子科学与技术（可...
- 64 篇 生物工程
- 50 篇 机械工程
- 44 篇 电气工程
- 26 篇 动力工程及工程热...
- 24 篇 仪器科学与技术
- 24 篇 化学工程与技术
- 15 篇 网络空间安全
- 14 篇 材料科学与工程（可...
- 14 篇 土木工程
- 12 篇 力学（可授工学、理...
- 12 篇 交通运输工程
- 12 篇 农业工程
- 12 篇 环境科学与工程（可...
288 篇 理学
- 180 篇 数学
- 67 篇 生物学
- 50 篇 物理学
- 40 篇 统计学（可授理学、...
- 36 篇 系统科学
- 25 篇 化学
183 篇 管理学
- 126 篇 管理科学与工程(可...
- 60 篇 图书情报与档案管...
- 35 篇 工商管理
19 篇 法学
- 16 篇 社会学
13 篇 经济学
- 13 篇 应用经济学
13 篇 农学
11 篇 教育学
- 11 篇 教育学
8 篇 医学
4 篇 文学
3 篇 军事学
2 篇 艺术学

主题

44 篇 computational mo...
31 篇 concurrent compu...
31 篇 training
30 篇 laboratories
30 篇 algorithm design...
28 篇 computer archite...
28 篇 benchmark testin...
28 篇 feature extracti...
28 篇 kernel
27 篇 semantics
27 篇 distributed proc...
25 篇 graphics process...
25 篇 servers
25 篇 hardware
23 篇 parallel process...
23 篇 fault tolerance
23 篇 cloud computing
21 篇 task analysis
21 篇 throughput
21 篇 distributed comp...

机构

169 篇 national laborat...
134 篇 science and tech...
104 篇 college of compu...
81 篇 national laborat...
38 篇 national laborat...
36 篇 science and tech...
35 篇 school of comput...
34 篇 national laborat...
29 篇 national key lab...
22 篇 science and tech...
22 篇 national key lab...
18 篇 national laborat...
16 篇 science and tech...
16 篇 national laborat...
15 篇 school of comput...
14 篇 national laborat...
14 篇 laboratory of di...
13 篇 national key lab...
12 篇 college of compu...
12 篇 national key lab...

作者

44 篇 yong dou
41 篇 dou yong
41 篇 wang huaimin
36 篇 dongsheng li
36 篇 liu jie
35 篇 huaimin wang
31 篇 jie liu
30 篇 peng yuxing
29 篇 yuxing peng
29 篇 li dongsheng
29 篇 yijie wang
27 篇 xiaodong wang
26 篇 wang yijie
24 篇 yin gang
23 篇 wang ji
22 篇 zhigang luo
21 篇 xingming zhou
20 篇 gang yin
20 篇 qiao peng
20 篇 li kuan-ching

语言

1,047 篇 英文
59 篇 中文
18 篇 其他

检索条件"机构=The Science and Technology on Parallel and Distributed Processing Laboratory"

共 1124 条记录，以下是921-930 订阅

全选清除本页清除全部题录导出标记到"检索档案"

详细简洁

排序：

相关度排序

相关度排序
时效性降序
时效性升序

Sim-spm: A SimpleScalar-Based Simulator for Multi-level SPM Memory Hierarchy Architecture

Sim-spm: A SimpleScalar-Based Simulator for Multi-level SPM ...

引用

IEEE International Conference on High Performance Computing and Communications (HPCC)

作者： Xiaoguang Ren Yuhua Tang Tao Tang Sen Ye Huiquan Wang Jing Zhou National Laboratory for Parallel and Distributed Processing School of Computer National University of Defense Technology Changsha Hunan China

As a fast on-chip SRAM managed by software (the application and/or compiler), Scratchpad Memory (SPM) is widely used in many fields. This paper presents a Simple Scalar-based multi-level SPM memory hierarchy architecture simulator Sim-spm. We simulate the hardware of the multi-level SPM memory hierarchy successfully by extending Sim-outorder, which is an out-of-order simulator from Simple Scalar. Through the simulating memory method, the simulation framework of the multi-level SPM memory hierarchy has been built under the existing ISA (Instruction Set Architecture), which largely reduces the requirement to modify the existing compiler. The experimental results show that Sim-spm can accurately simulate the running state of the processor with a multi-level SPM memory hierarchy architecture, and it has a good prospect for the research of multi-level SPM memory hierarchy architecture.

关键词： Kernel Random access memory Libraries Memory management Memory architecture

来源：评论

学校读者我要写书评

暂无评论

Communication time models on fat-tree networks

Communication time models on fat-tree networks

引用

International Conference on Advanced Computer Theory and Engineering, ICACTE

作者： Yufei Lin Xinhai Xu Yisong Lin National Laboratory of Parallel and Distributed Processing Computer School National University of Defense Technology Changsha Hunan China

ISBN: (纸本)9781424465392;9781424465422

With the growth of supercomputer's scale, the communication time during executing is increasing. This phenomenon arouses the architecture researchers' interests. In this paper, based on the fat-tree topology, which is widely used in Infiniband, we present an one-to-all broadcast communication time model. After classifying applications into two kinds, we establish the ideal model and the bandwidth-limited model on the exponential-capacity binary fat-trees for the two kinds of applications. Through analyzing the models, we get the curves which describe the relationship between the communication time and the processor number. The conclusions we get in this paper can help system designers make better system design.

关键词： World Wide Web

来源：评论

学校读者我要写书评

暂无评论

Kernel Fusion: An Effective Method for Better Power Efficiency on Multithreaded GPU

Kernel Fusion: An Effective Method for Better Power Efficien...

引用

IEEE/ACM Int'l Conference on & Int'l Conference on Cyber, Physical and Social Computing (CPSCom) Green Computing and Communications (GreenCom)

作者： Guibin Wang YiSong Lin Wei Yi National Laboratory of Parallel and Distributed Processing School of Computer National University of Defense Technology Changsha Hunan China

ISBN: (纸本)9781424497799

As one of the most popular accelerators, Graphics processing Unit (GPU) has demonstrated high computing power in several application fields. On the other hand, GPU also produces high power consumption and has been one of the most largest power consumers in desktop and supercomputer systems. However, software power optimization method targeted for GPU has not been well studied. In this work, we propose kernel fusion method to reduce energy consumption and improve power efficiency on GPU architecture. Through fusing two or more independent kernels, kernel fusion method achieves higher utilization and much more balanced demand for hardware resources, which provides much more potential for power optimization, such as dynamic voltage and frequency scaling (DVFS). Basing on the CUDA programming model, this paper also gives several different fusion methods targeted for different situations. In order to make judicious fusion strategy, we deduce the process of fusing multiple independent kernels as a dynamic programming problem, which could be well solved with many existing tools and be simply embedded into compiler or runtime system. To reduce the overhead introduced by kernel fusion, we also propose effective method to reduce the usage of shared memory and coordinate the thread space of the kernels to be fused. Detailed experimental evaluation validates that the proposed kernel fusion method could reduce energy consumption without performance loss for several typical kernels.

关键词： Kernel Instruction sets Graphics processing unit Energy consumption Hardware Mathematical model Dynamic programming

来源：评论

学校读者我要写书评

暂无评论

Power-Efficient Work Distribution Method for CPU-GPU Heterogeneous System

Power-Efficient Work Distribution Method for CPU-GPU Heterog...

引用

International Symposium on parallel and distributed processing with Applications, ISPA

作者： Guibin Wang Xiaoguang Ren National Laboratory of Parallel and Distributed Processing School of Computer National University of Defense Technology Changsha Hunan China

As the system scales up continuously, the problem of power consumption for high performance computing (HPC) system becomes more severe. Heterogeneous system integrating two or more kinds of processors, could be better adapted to heterogeneity in applications and provide much higher energy efficiency in theory. Many studies have shown heterogeneous system is preferable on energy consumption to homogeneous system in a multi-programmed computing environment. However, how to exploit energy efficiency (Flops/Watt) of heterogeneous system for a single application or even for a single phase in an application has not been well studied. This paper proposes a power-efficient work distribution method for single application on a CPU-GPU heterogeneous system. The proposed method could coordinate inter-processor work distribution and per-processor's frequency scaling to minimize energy consumption under a given scheduling length constraint. We conduct our experiment on a real system, which equips with a multi-core CPU and a multi-threaded GPU. Experimental results show that, with reasonably distributing work over CPU and GPU, the method achieves 14% reduction in energy consumption than static mappings for several typical benchmarks. We also demonstrate that our method could adapt to changes in scheduling length constraint and hardware configurations.

关键词： Graphics processing unit Energy consumption Power demand Processor scheduling Benchmark testing Power measurement

来源：评论

学校读者我要写书评

暂无评论

Conflict graph based hardware transactional memory

Conflict graph based hardware transactional memory

引用

International Conference on Computer science and Information technology (CSIT)

作者： Kun Zeng National Laboratory for Parallel and Distributed Processing School of Computer National University of Defense Technology Changsha Hunan China

This paper proposes a novel transactional memory design: conflict graph based hardware transactional memory. It allows two conflicting transactions both to commit if they do not violate the condition of serializability. Simulation results show that conflict graph based hardware transactional memory outperforms the state-of-art transactional memory system.

关键词： Protocols

来源：评论

学校读者我要写书评

暂无评论

SemanticCast: Content-Based Data Distribution over Self-Organizing Semantic Overlay Networks

SemanticCast: Content-Based Data Distribution over Self-Orga...

引用

IEEE International Conference on parallel and distributed Computing, Applications and Technologies (PDCAT)

作者： Zhong Zheng Yijie Wang National Key Laboratory for Parallel and Distributed Processing School of Computer National University of Defense Technology Changsha China

Many applications demand distributing data with different contents efficiently in the network environment with unreliable links and a high node churn. Existing approaches mostly focus on optimizing either efficiency or robustness of data distribution, and fail to ensure both of them simultaneously. In this paper, we propose Semantic Cast - a content-based data distribution approach over self-organizing semantic overlay networks. Semantic Cast maintains a self-organizing semantic overlay based on view exchange (called Crowd). In Crowd, each node seeks neighbors with more similar interests by periodically exchanging its neighbor list (called view) with a chosen neighbor. Through these nodes' self-organizing behavior, various interest communities emerge in the overlay. For data distribution over Crowd, Semantic Cast adopts random walk to route data between interest communities, and adopts flooding to disseminate data inside the interested communities. The experimental results show that compared to existing approaches, Semantic Cast can support efficient content-based data distribution in the unreliable and highly dynamic network environment.

关键词： Semantics Peer to peer computing Communities Routing Robustness Convergence Subscriptions

来源：评论

学校读者我要写书评

暂无评论

Metal interference analysis and design rules of on-chip antennas for wireless interconnect

Metal interference analysis and design rules of on-chip ante...

引用

URSI International Symposium on Signals, Systems, and Electronics (ISSSE)

作者： Xiaowei He Minxuan Zhang Jinwen Li Shaoqing Li National Laboratory of Parallel and Distributed Processing School of Computer National University of Defense Technology Changsha Hunan China

The influence of on-chip metal interconnections, power grids, heat sink together with packaging, and metal dummy fills on the transmission characteristics of a 2mm-long integrated dipole antenna pair has been investigated in this paper. These metal structures and placements have been classified and particular simulations are performed to explore the interference effects of neighboring various metal structures on transmission gain, phase, impedance and radiation pattern for on-chip dipole antenna pair. By virtue of the experimental results and analyses, several experiential linear expressions for antenna pair gain and phase in interference circumstances are obtained using numerical fit. A set of design rules is concluded accordingly for guiding on-chip antenna layout and design targeting wireless interconnect.

关键词： Metals System-on-a-chip Transmitting antennas Dipole antennas Silicon

来源：评论

学校读者我要写书评

暂无评论

Window memory accesses method in alternate row/column matrix access systems

Window memory accesses method in alternate row/column matrix...

引用

International Conference on Computer Engineering and technology, ICCET

作者： Jie Zhou Yong Dou Yuanwu Lei Yazhuo Dong National Laboratory for Parallel & Distributed Processing National University of Defense Technology Changsha China Unit 91655 Beijing China

ISBN: (纸本)9781424463473;9781424463497;9781424463503

Many systems, such as Synthetic Aperture Radar (SAR) processing, two-dimensional image processing, 2d-FFT calculation, need access the row and column data of their matrix alternately. The DRAM memory should be used due to huge data in these systems. To improve the usage of memory bandwidth in such systems, this paper theoretically analyses the optimal window size to minimize the total number of opening/closing pages when performing in such instances by balancing the number of handling physical pages between row and column accesses. This paper presents a window-based optimal memory access method, and we implemented an FPGA-based SDRAM controller with eight simple ports, which is based on window accessing mechanism and supports commercialized SDARM. The experimental results show that the effective I/O bandwidth of external SDRAM using our window layout approach increases from 114.2MB/s of naive implementation to 730.2MB/s with over 6X speedup. In addition, we implemented two SAR processing systems with four FFT processing elements using our window-based SDRAM controller and Corner Turn method separately in FPGA chip. Results show window-based method can achieve a speedup of 2.6 compared to Corner Turn method.

关键词： SDRAM Synthetic aperture radar Bandwidth Image processing Random access memory Performance analysis Optimal control Commercialization Control systems Field programmable gate arrays

来源：评论

学校读者我要写书评

暂无评论

A Simulated Annealing Technique for Optimizing Time Warp Simulation

A Simulated Annealing Technique for Optimizing Time Warp Sim...

引用

International Conference on Computer Modeling and Simulation, ICCMS

作者： Wei Zhang Sina Meraji Jun Wang Carl Tropper National Laboratory of Parallel and Distributed Processing School of Computer Science National University of Defense Technology ChangSha China School of Computer Science McGill University Montreal Canada

According to Moore's law the complexity of VLSI circuits has doubled approximately every two years, resulting in simulation becoming the major bottleneck in the circuit design process. parallel and distributed simulations can be applied as fast, cost effective approaches to the simulation of large, complex circuits. In this paper, a simple yet effective simulated annealing-based approach is proposed to optimize the choice of a time window for optimistic parallel simulation. We chose gate level circuits simulations as our experimental vehicle. Our results show up to a 52% improvement in the simulation time using our simulated annealing algorithm. To the best of our knowledge, this is the first time that SA has been applied to optimize the performance of time warp simulations.

关键词： Time warp simulation Simulated annealing Circuit simulation Computational modeling Hardware design languages Discrete event simulation Computer simulation Voltage control Concurrent computing Computer science

来源：评论

学校读者我要写书评

暂无评论

Automatic synthesis of processor arrays with local memories on FPGAs

Automatic synthesis of processor arrays with local memories ...

引用

IEEE International Conference on Field-Programmable technology (FPT)

作者： Guiming Wu Yong Dou Miao Wang National Laboratory of Parallel and Distributed Processing National University of Defense Technology Changsha China Jiangnan Institute of Computing Technology Wuxi China

In this paper, we present an automatic synthesis framework to map loop nests to processor arrays with local memories on FPGAs. An affine transformation approach is firstly proposed to address space-time mapping problem. Then a data-driven architecture model is introduced to enable automatic generation of processor arrays by extracting this data-driven architecture model from transformed loop nests. Some techniques including memory allocation, communication generation and control generation are presented. Synthesizable RTL codes can be easily generated from the architecture model built by these techniques. A preliminary synthesis tool is implemented based on PLUTO, an automatic polyhedral source-to-source transformation and parallelization framework.

关键词： parallel processing Registers Field programmable gate arrays Radiation detectors Arrays Computational modeling

来源：评论

学校读者我要写书评

暂无评论

没有更多数据了...

全选清除本页清除全部题录导出标记到“检索档案”

共113页 << < 89 90 91 92 93 94 95 96 97 98 > >>

检索报告对象比较合并检索0

隐藏清空

合并搜索

回到顶部

执行限定条件

内容：

评分：

请选择保存的检索档案：

请选择收藏分类：

订阅名称：

通借通还

温馨提示：

图书名称：

借书校区：

取书校区：

手机号码：

邮箱地址：

一卡通帐号：

电话和邮箱必须正确填写，我们会与您联系确认。

联系人：

所在院系：

联系邮箱：

联系电话：

内蒙古自治区呼和浩特市赛罕区大学西街235号邮编: 010021

建议与咨询 留下您的常用邮箱和电话号码，以便我们向您反馈解决方案和替代方法

分类表

所选分类

限定检索结果

文献类型

馆藏范围

日期分布

学科分类号

主题

机构

作者

语言

请选择保存的检索档案： 新增检索档案 确定 取消

请选择收藏分类： 新增自定义分类 确定 取消

通借通还

建议与咨询留下您的常用邮箱和电话号码，以便我们向您反馈解决方案和替代方法

请选择保存的检索档案：

请选择收藏分类：