课题基金 / 基金详情

Optimizing hadoop to scale to big systems and big-data

Optimizing hadoop to scale to big systems and big-data
优化 hadoop 以扩展到大系统和大数据
批准号:
485325-2015
负责人:
Shriraman, Arrvindh
金额:
$5.77万
依托单位:
依托单位国家:
加拿大
项目类别:
Collaborative Research and Development Grants
财政年份:
2019
资助国家:
加拿大
项目状态:
已结题
起止时间:
2019-01-01 至 2020-12-31

项目摘要

项目成果

Shriraman, Arrvindh的其他基金

相似基金

相关文献

中文摘要
翻译
点击翻译按钮获取中文摘要
英文摘要
The overall `big-data' market is expected to be worth more than \$100 billion and growing roughly twice as fast as the software business as a whole [Deloitte]. The emergence of "big data", such as social media, has played critical roles in business analytics. As an IDC report indicates data analysis will play a key role in managing traffic in canadian cities (e.g., Toronto) and healthcare. Such big data is often stored in the new generation of NoSQL databases, and processed by massively parallel MapReduce/Hadoop based infrastructure. The industryis facing an essential challenge: can millions of existing successful applications based on SQL take the advantage of big data infrastructure? Many critical challenges remain including the design of optimal data storage formats and efficient communication of data between data storage nodes. In this project, we will be working with our partner, Simba, an industry leader of big data technology and has guided us towards the challenges that forsee from their client's perspective. We will be making three specific contributions i) Low-latency: We will be adapting Apache Hive (open source Hadoop-based NoSQL framework) to take advantage of RDMA (Remote Direct Memory Accesses) networks and scale out to take advantage of rack scale memory resources. ii) Reducing Bandwidth: We will be developing data-specific compression mechanisms to minimize the data movement across the compute rack and enable both an increase in database volume and data access speed, and iii) ``Flexibility": We will be developing a tool for enabling Apache Hive to dynamically enforce schemas and enable end-client SQL-based business analytic tools to interact with NoSQL database backends employed by Hadoop.We also plan to integrate our changes into the open source Apache Hive platform to benefit the wider Big data community.
期刊论文(0)
专著(0)
科研奖励(0)
会议论文
Self-Sketching Domain Specific Accelerators: Build Hardware from Software
  • 批准号:
    RGPIN-2018-06795
  • 项目类别:
    Discovery Grants Program - Individual
  • 资助金额:
    $5.97万
  • 财政年份:
    2022
  • 负责人:
    Shriraman, Arrvindh
  • 依托单位:
Self-Sketching Domain Specific Accelerators: Build Hardware from Software
  • 批准号:
    RGPIN-2018-06795
  • 项目类别:
    Discovery Grants Program - Individual
  • 资助金额:
    $2.99万
  • 财政年份:
    2021
  • 负责人:
    Shriraman, Arrvindh
  • 依托单位:
Self-Sketching Domain Specific Accelerators: Build Hardware from Software
  • 批准号:
    RGPIN-2018-06795
  • 项目类别:
    Discovery Grants Program - Individual
  • 资助金额:
    $2.99万
  • 财政年份:
    2020
  • 负责人:
    Shriraman, Arrvindh
  • 依托单位:
Self-Sketching Domain Specific Accelerators: Build Hardware from Software
  • 批准号:
    RGPIN-2018-06795
  • 项目类别:
    Discovery Grants Program - Individual
  • 资助金额:
    $2.99万
  • 财政年份:
    2019
  • 负责人:
    Shriraman, Arrvindh
  • 依托单位:
国内基金
海外基金
基于在线优化的Hadoop YARN平台下资源分配机制研究
  • 批准号:
    61802060
  • 项目类别:
    青年科学基金项目
  • 资助金额:
    27.0万元
  • 批准年份:
    2018
  • 负责人:
    徐欢乐
  • 依托单位:
基于hadoop技术的住院患者用药错误风险预警与干预
  • 批准号:
    2018JJ2597
  • 项目类别:
    省市级项目
  • 资助金额:
    --
  • 批准年份:
    2018
  • 负责人:
    丁四清
  • 依托单位:
多维气候大数据存储与处理关键技术研究
  • 批准号:
    61672312
  • 项目类别:
    面上项目
  • 资助金额:
    64.0万元
  • 批准年份:
    2016
  • 负责人:
    杨广文
  • 依托单位:
HDFS读、写性能概率建模与模型迁移方法研究
  • 批准号:
    61502379
  • 项目类别:
    青年科学基金项目
  • 资助金额:
    20.0万元
  • 批准年份:
    2015
  • 负责人:
    董博
  • 依托单位: