课题基金 / 基金详情

Collaborative Research: PPoSS: LARGE: A Full-Stack Architecture for Sparse Computation

Collaborative Research: PPoSS: LARGE: A Full-Stack Architecture for Sparse Computation
协作研究:PPoSS:LARGE:稀疏计算的全栈架构
批准号:
2216978
负责人:
Milind Kulkarni
金额:
$55.0万
依托单位:
依托单位国家:
美国
项目类别:
Continuing Grant
财政年份:
2022
资助国家:
美国
项目状态:
未结题
起止时间:
2022-10-01 至 2027-09-30

项目摘要

项目成果

Milind Kulkarni的其他基金

相似基金

相关文献

中文摘要
翻译
点击翻译按钮获取中文摘要
英文摘要
Computer systems have been designed and optimized primarily for dense computations, i.e., those that process regularly structured data. But current systems are ill-suited to sparse computations, i.e., those that process unstructured data. Sparse computations are very common because many relations and interactions are sparse. For example, most people are not friends and most neurons are not directly connected. Sparse computations take advantage of this sparsity by encoding and processing only meaningful relations, such as storing only the non-zero elements of a matrix. These applications are crucial in many domains, like deep learning, data analytics, and scientific computing, but their irregular structure makes them inefficient and hard to scale in currentsystems, wasting billions of dollars yearly. This project aims to redesign the computing stack to provide first-class support for sparse computations. The project's novelties include a full system stack that spans programming languages, compilers, and specialized hardware architectures and large-scale computer systems. The project's impacts include making future parallel systems much more versatile, scalable, energy efficient and easier to program.This project takes a coordinated approach across the system stack to unlock the performance and scalability of sparse computations, because they pose challenges that cannot be addressed at a single layer. For example, sparse computations have a rich space of choices in algorithm, data representation, and schedule, which current languages and compilers cannot capture or optimize properly. The right choice of algorithm and data representation are often unknown in advance and may change at run-time, thwarting the rigid division between current compilers and schedulers. Irregular, data-dependent control and memory accesses stymie compiler analysis, hinder parallelization, make poor use of hardware, and introduce numerous side channels that thwart security. Finally, their data-intensive nature is a poor match to the processors and accelerators pervasive in current clusters and datacenters, which optimize for compute operations rather than to minimize data movement. To tackle these challenges, this project will develop a full system stack spanning domain-specific languages, a tightly integrated compiler and scheduler, and specialized hardware architectures and high-performance, multi-node computer systems and networks. This stack is built around a unifying abstraction, anovel sparse intermediate representation that (1) encodes semantic information on key sparse data structures and their iterations, (2) enables optimizing compiler transformations and dynamic scheduling decisions, and (3) can be easily compiled to parallel architectures, including graphics processing units (GPUs), general-purpose processors, our proposed specialized architecture, and their combination. The full stack will be designed with security at the forefront, leveraging novel cross-layer techniques to achieve secure high performance. This system will be rigorously evaluated using a broad set of sparse applications and at a wide range of system scales, including large-scale clusters with hundreds of GPUs or tens of specialized processors. By innovating across the full software and hardware stack, these techniques will achieve performance, scalability, and efficiency gains that single-layer approaches cannot provide.This award reflects NSF's statutory mission and has been deemed worthy of support through evaluation using the Foundation's intellectual merit and broader impacts review criteria.
期刊论文(2)
专著(0)
科研奖励(0)
会议论文
RT-kNNS Unbound: Using RT Cores to Accelerate Unrestricted Neighbor Search
RT-kNNS Unbound:使用 RT 内核加速无限制邻居搜索
DOI: 10.1145/3577193.3593738
发表时间: 2023
期刊: ACM
影响因子: --
作者: [Nagarajan, Vani, Mandarapu, Durga, Kulkarni, Milind]
通讯作者: Kulkarni, Milind
RT-DBSCAN: Accelerating DBSCAN using Ray Tracing Hardware
RT-DBSCAN:使用光线追踪硬件加速 DBSCAN
DOI: 10.1109/ipdps54959.2023.00100
发表时间: 2023
期刊: IEEE
影响因子: --
作者: [Nagarajan, Vani, Kulkarni, Milind]
通讯作者: Kulkarni, Milind
Travel: Student Travel Grant for the Programming Languages Mentoring Workshop at PLDI 2022
  • 批准号:
    2227746
  • 项目类别:
    Standard Grant
  • 资助金额:
    $1.5万
  • 财政年份:
    2022
  • 负责人:
    Milind Kulkarni
  • 依托单位:
SHF: Small: A Composable, Sound Optimization Framework for Loops and Recursion
  • 批准号:
    1908504
  • 项目类别:
    Standard Grant
  • 资助金额:
    $45.0万
  • 财政年份:
    2019
  • 负责人:
    Milind Kulkarni
  • 依托单位:
SPX: Write Once, Run on Anything: Verified, Tuned Accelerator Kernels from High Level Specifications
  • 批准号:
    1919197
  • 项目类别:
    Standard Grant
  • 资助金额:
    $125.0万
  • 财政年份:
    2019
  • 负责人:
    Milind Kulkarni
  • 依托单位:
NSF Student Travel Grant for 2019 Midwest Programming Languages Summit (MWPLS)
  • 批准号:
    1942074
  • 项目类别:
    Standard Grant
  • 资助金额:
    $0.5万
  • 财政年份:
    2019
  • 负责人:
    Milind Kulkarni
  • 依托单位:
国内基金
海外基金
Research on Quantum Field Theory without a Lagrangian Description
  • 批准号:
    24ZR1403900
  • 项目类别:
    省市级项目
  • 资助金额:
    --
  • 批准年份:
    2024
  • 负责人:
    SATOSHI NAWATA
  • 依托单位:
Cell Research
Cell Research
Cell Research (细胞研究)