课题基金 / 基金详情

BIGDATA: F: DKA: Usable Multiple Scale Big Data Analytics through Interactive Visualization

BIGDATA: F: DKA: Usable Multiple Scale Big Data Analytics through Interactive Visualization
BIGDATA:F:DKA:通过交互式可视化进行可用的多尺度大数据分析
批准号:
1447416
负责人:
Christopher North
金额:
$99.89万
依托单位国家:
美国
项目类别:
Standard Grant
财政年份:
2014
资助国家:
美国
项目状态:
已结题
起止时间:
2014-09-01 至 2018-08-31

项目摘要

项目成果

Christopher North的其他基金

相似基金

相关文献

中文摘要
翻译
从大数据中获得更大的洞察力需要进行大量的分析,这会带来很大的可用性问题。大数据分析通常依赖于在多个级别的数据规模上运行的多个计算和统计模型来发现和表征值得注意的模式。这些模型联合或按顺序工作,对大数据进行筛选、分组、汇总和可视化,以便分析师可以评估数据。作为大文本分析中的一个简单示例,首先对海量文本进行采样以寻找相关或代表性的单词,然后通过复杂形式的建模(例如,主题建模)进一步对其进行约简,然后通过应用降维算法来可视化。随着数据大小的增加,模型的数量也在增加,分析过程中对人工交互的需求也在增加。通过互动,人类将专家的判断纳入分析过程,并从不同的角度高效地探索和理解大数据。然而,由于各种原因,与任何单个模特互动都是困难的,更不用说越来越多的模特了。因此,当前的人-机-交互研究与复杂的统计方法和快速计算相结合,开发了一个可用的、多模型的大数据分析框架。该框架包裹在软件中,专业用户和学生用户都可以访问;即,可以在当前的政府和工业大数据集中做出新的发现,以及通过新的教学模块培养未来的本科生和研究生级别的分析师。新的分析框架将视觉参数交互(V2PI)扩展到多尺度V2PI(MV2PI)。V2PI目前支持可用的小数据分析,并允许用户通过与可视化中的数据直接交互来调整模型参数。也就是说,V2PI定量地解释视觉交互,以更新底层模型参数并产生新的可视化效果。MV2PI现在将在统一交互空间中以多个级别的数据规模运行的多个模型链接在一起。在MV2PI中,可视化中的小规模数据交互传播到更大规模的模型(通过反转它们并更新其参数),并生成新的可视化。在文本分析示例中,如果用户将几个数据点拖在一起以假设一个集群,则倒向降维模型计算更新的维权重,在大范围上查询相关的新命中,识别改变的主题,并更新布局以显示对新集群的大数据支持。有了MV2PI,用户可以实时交互地探索模型之间的大规模数据和复杂的相互关系,并以一种可用的方式直接支持他们自然的认知感知过程。MV2PI的开发涉及:(1)制定明确的框架;(2)创建新的交互模型(例如,交互K-Means和交互潜在Dirichlet分配),该模型涵盖不同级别并支持MV2PI模型反演;(3)实施计算方法以支持高性能、实时的模型更新;以及(4)评估MV2PI软件框架的可用性和有效性。项目网站(http://www.apps.stat.vt.edu/bava/mv2pi.html))将包括关于MV2PI开发、软件访问、数据集、教育材料和出版物的信息。
英文摘要
Gaining big insight from big data requires big analytics, which poses big usability problems. Analyses of big data often rely on several computational and statistical models that operate on multiple levels of data scale to discover and characterize noteworthy patterns. The models work jointly or in sequence to filter, group, summarize, and visualize big data so that analysts may assess the data. As a simple example in big text analytics, massive text is first sampled for relevant or representative words, then further reduced by a complex form of modeling (e.g., topic modeling), then visualized by applying a dimension reduction algorithm. As the size of data increases, so does the number of models and, likewise, the need for human interaction in the analytical process. By interacting, humans include expert judgment into the analytical process, and efficiently explore and make sense of big data from varying perspectives. However, for a variety of reasons, interacting with any individual model is difficult, let alone a growing number of models. Thus, current human-computer-interaction research is merged with complex statistical methods and fast computation to develop a usable, multi-model analytic framework for big data. Wrapped in software, the framework will be accessible to both professional and student users alike; i.e., available to make new discoveries in current government and industrial big datasets, as well as, educate future analysts at the undergraduate and graduate levels given new teaching modules. The new analytic framework extends Visual-to-Parametric Interaction (V2PI) to Multi-scale V2PI (MV2PI). V2PI currently supports usable small-data analytics, and enables users to adjust model parameters by interacting directly with data in visualizations. That is, V2PI interprets visual interactions quantitatively to update underlying model parameters and produce new visualizations. MV2PI now links together several models that operate at multiple levels of data-scale in a unified interactive space. In MV2PI, small-scale data interactions in visualizations propagate to larger scale models (by inverting them and updating their parameters) and new visualizations are generated. In the text analytics example, if users drag several data points together to hypothesize a cluster, the inverted dimension reduction model computes updated dimension weights, queries relevant new hits at the large scale, identifies changed topics, and updates the layout to show big-data support for the new cluster. With MV2PI, users may interactively explore large-scale data and complex inter-relationships between models in real time, and in a usable fashion that directly supports their natural cognitive sensemaking process. Development of MV2PI involves: (1) formulation of an explicitly stated framework ; (2) creation of new interactive models (e.g., Interactive K-means and Interactive Latent Dirichlet Allocation) that cover different levels of scale and support MV2PI model inversion; (3) implementation of computational methods to support high-performance, real-time model updates; and (4) evaluation of MV2PI software framework for usability and effectiveness. The project web site (http://www.apps.stat.vt.edu/bava/mv2pi.html) will include information on MV2PI development, access to software, datasets, educational materials, and publications.
期刊论文(0)
专著(0)
科研奖励(0)
会议论文
Collaborative Research: CSSI Frameworks: SAGE3: Smart Amplified Group Environment for Harnessing the Data Revolution
The Dawn of Gravitational Wave Astronomy
  • 批准号:
    ST/P000924/1
  • 项目类别:
    Fellowship
  • 资助金额:
    $11.01万
  • 财政年份:
    2016
  • 负责人:
    Christopher North
  • 依托单位:
LEGO-LIGO
  • 批准号:
    ST/P001688/1
  • 项目类别:
    Research Grant
  • 资助金额:
    $0.49万
  • 财政年份:
    2016
  • 负责人:
    Christopher North
  • 依托单位:
HCC: Small: Semantic Interaction for Visual Text Analytics
国内基金
海外基金
HIV-1逆转录酶/整合酶双重抑制剂DKA-DAPYs的分子设计、合成及抗HIV活性研究
  • 批准号:
    21402148
  • 项目类别:
    青年科学基金项目
  • 资助金额:
    25.0万元
  • 批准年份:
    2014
  • 负责人:
    古双喜
  • 依托单位: