使用 4D 骨骼增强对齐 (Context-Aware Sequence Alignment using 4D Skeletal Augmentation) - 专知论文

会员服务 ·

0

CASA · state-of-the-art · 学成 · 估计/估计量 · Neural Networks ·

2022 年 4 月 26 日

Context-Aware Sequence Alignment using 4D Skeletal Augmentation

翻译：使用 4D 骨骼增强对齐

Taein Kwon,Bugra Tekin,Siyu Tang,Marc Pollefeys

from arxiv, Project page: http://www.taeinkwon.com/projects/casa. Accepted to CVPR 2022 Oral

Temporal alignment of fine-grained human actions in videos is important for numerous applications in computer vision, robotics, and mixed reality. State-of-the-art methods directly learn image-based embedding space by leveraging powerful deep convolutional neural networks. While being straightforward, their results are far from satisfactory, the aligned videos exhibit severe temporal discontinuity without additional post-processing steps. The recent advancements in human body and hand pose estimation in the wild promise new ways of addressing the task of human action alignment in videos. In this work, based on off-the-shelf human pose estimators, we propose a novel context-aware self-supervised learning architecture to align sequences of actions. We name it CASA. Specifically, CASA employs self-attention and cross-attention mechanisms to incorporate the spatial and temporal context of human actions, which can solve the temporal discontinuity problem. Moreover, we introduce a self-supervised learning scheme that is empowered by novel 4D augmentation techniques for 3D skeleton representations. We systematically evaluate the key components of our method. Our experiments on three public datasets demonstrate CASA significantly improves phase progress and Kendall's Tau scores over the previous state-of-the-art methods.

翻译：视频中细微的人类行为在时间上的配合对于计算机视觉、机器人和混杂现实中的许多应用非常重要。最先进的方法通过利用强大的深层神经神经网络直接学习基于图像的嵌入空间。其结果虽然不简单,但结果远不令人满意, 相配的视频在时间上表现出严重的不连续性, 而没有额外的处理步骤。人体和手部最近的进步在野生前景中提出了解决视频中人类行动协调任务的新方法的估计。在这项工作中,基于现成的人类形象估计器,我们提出了一个新的环境觉悟自我监督学习架构,以协调行动序列。我们命名CASA。具体地说, CASA使用自我注意和交叉注意机制来纳入人类行动的空间和时间背景,这可以解决时间不连续问题。此外,我们引入了一种自我超强的学习计划,通过新型的4D增强技术来增强3D骨架演示。我们系统地评估了我们的方法的关键组成部分。我们在三个公共数据集上的实验展示了CASA- TaI的阶段和Kenall的成绩。

0

相关内容

CASA

国际计算机动画和社会代理国际会议（CASA ）是世界上最古老的计算机动画和社交代理国际会议。会议主题包括但不限于计算机动画，虚拟代理，社交代理，虚拟现实和增强现实以及可视化。官网地址：http://dblp.uni-trier.de/db/conf/ca/

ICLR 2022杰出论文公布：7篇论文获得，清华朱军课题组摘得

ICLR 2022杰出论文公布：7篇论文获得，清华朱军课题组摘得

专知会员服务

60+阅读 · 2022年4月22日

【CVPR 2022】基于粗粒度和细粒度特征匹配的视频描述评估，EMScore: Evaluating Video Captioning via Coarse-Grained and Fine-Grained Embedding Matching

【CVPR 2022】基于粗粒度和细粒度特征匹配的视频描述评估，EMScore: Evaluating Video Captioning via Coarse-Grained and Fine-Grained Embedding Matching

专知会员服务

10+阅读 · 2022年3月19日

最新《Transformers模型》教程，64页ppt

最新《Transformers模型》教程，64页ppt

专知会员服务

321+阅读 · 2020年11月26日

100+篇《自监督学习(Self-Supervised Learning)》论文最新合集

100+篇《自监督学习(Self-Supervised Learning)》论文最新合集

专知会员服务

165+阅读 · 2020年3月18日

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

专知会员服务

59+阅读 · 2019年10月17日

强化学习最新教程，17页pdf

强化学习最新教程，17页pdf

专知会员服务

182+阅读 · 2019年10月11日

[综述]深度学习下的场景文本检测与识别

[综述]深度学习下的场景文本检测与识别

专知会员服务

78+阅读 · 2019年10月10日

【人工智能在2019：一年回顾】反人工智能，AI in 2019: A Year in Review

【人工智能在2019：一年回顾】反人工智能，AI in 2019: A Year in Review

专知会员服务

79+阅读 · 2019年10月10日

【哈佛大学商学院课程Fall 2019】机器学习可解释性

【哈佛大学商学院课程Fall 2019】机器学习可解释性

专知会员服务

105+阅读 · 2019年10月9日

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

专知会员服务

41+阅读 · 2019年10月9日

ACM MM 2022 Call for Papers

ACM MM 2022 Call for Papers

CCF多媒体专委会

5+阅读 · 2022年3月29日

【ICIG2021】Latest News & Announcements of the Plenary Talk1

【ICIG2021】Latest News & Announcements of the Plenary Talk1

中国图象图形学学会CSIG

0+阅读 · 2021年11月1日

【ICIG2021】Latest News & Announcements of the Industry Talk2

【ICIG2021】Latest News & Announcements of the Industry Talk2

中国图象图形学学会CSIG

0+阅读 · 2021年7月29日

【ICIG2021】Latest News & Announcements of the Industry Talk1

【ICIG2021】Latest News & Announcements of the Industry Talk1

中国图象图形学学会CSIG

0+阅读 · 2021年7月28日

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

Unsupervised Learning via Meta-Learning

Unsupervised Learning via Meta-Learning

CreateAMind

43+阅读 · 2019年1月3日

disentangled-representation-papers

disentangled-representation-papers

CreateAMind

26+阅读 · 2018年9月12日

【论文推荐】最新十篇目标跟踪相关论文—多帧光流跟踪、动态图学习、MV-YOLO、姿态估计、深度核相关滤波、Benchmark

【论文推荐】最新十篇目标跟踪相关论文—多帧光流跟踪、动态图学习、MV-YOLO、姿态估计、深度核相关滤波、Benchmark

专知

13+阅读 · 2018年5月26日

【论文推荐】最新六篇图像分割相关论文—控制、全卷积网络、子空间表示、多模态图像分割

【论文推荐】最新六篇图像分割相关论文—控制、全卷积网络、子空间表示、多模态图像分割

专知

25+阅读 · 2018年4月15日

【论文】变分推断（Variational inference)的总结

【论文】变分推断（Variational inference)的总结

机器学习研究会

39+阅读 · 2017年11月16日

温敏二嵌段胶束种子大分子RAFT试剂调介下的种子分散RAFT聚合

国家自然科学基金

0+阅读 · 2014年12月31日

常染色体隐性遗传小脑性共济失调新的致病基因CAX的功能研究

国家自然科学基金

0+阅读 · 2014年12月31日

玉米异染色质纽(Knob)形成的表观遗传机制及进化分析

国家自然科学基金

0+阅读 · 2013年12月31日

非编码RNA调控RNA聚合酶II转录的研究

国家自然科学基金

0+阅读 · 2013年12月31日

鱼类ADAR1剪接异构体基因的鉴定及其转录调控

国家自然科学基金

0+阅读 · 2012年12月31日

镧系硅氧氮化物荧光材料的晶体结构和发光特性

国家自然科学基金

0+阅读 · 2012年12月31日

抗肿瘤药物多功能纳米自组装载体逆转肿瘤多药耐药性及其机制研究

国家自然科学基金

0+阅读 · 2012年12月31日

HGF诱导NSCLC细胞对EGFR-TKIs耐药机制的研究。

国家自然科学基金

0+阅读 · 2011年12月31日

基于list-mode数据的快速SART真3D PET断层重建算法的研究

国家自然科学基金

0+阅读 · 2011年12月31日

PGK1诱导肿瘤基质成纤维细胞激活在前列腺癌发展、转移中作用的研究

国家自然科学基金

0+阅读 · 2009年12月31日

Learning 3D Object Shape and Layout without 3D Supervision

Arxiv

0+阅读 · 2022年6月14日

Object Scene Representation Transformer

Arxiv

0+阅读 · 2022年6月14日

A Survey of Automated Data Augmentation Algorithms for Deep Learning-based Image Classication Tasks

Arxiv

1+阅读 · 2022年6月14日

Looking Outside the Box to Ground Language in 3D Scenes

Arxiv

0+阅读 · 2022年6月13日

Learn2Augment: Learning to Composite Videos for Data Augmentation in Action Recognition

Arxiv

0+阅读 · 2022年6月9日

Transformers in Time Series: A Survey

Arxiv

34+阅读 · 2022年2月15日

A Survey on Data Augmentation for Text Classification

A Survey on Data Augmentation for Text Classification

Arxiv

16+阅读 · 2021年7月7日

Data Augmentation for Graph Neural Networks

Arxiv

38+阅读 · 2020年12月2日

Learning from Few Samples: A Survey

Learning from Few Samples: A Survey

Arxiv

77+阅读 · 2020年7月30日

Graph Convolutional Networks for Text Classification

Arxiv

11+阅读 · 2018年10月17日

VIP会员

文章信息

相关主题

state-of-the-art

估计/估计量

Neural Networks

相关VIP内容

ICLR 2022杰出论文公布：7篇论文获得，清华朱军课题组摘得

ICLR 2022杰出论文公布：7篇论文获得，清华朱军课题组摘得

专知会员服务

60+阅读 · 2022年4月22日

【CVPR 2022】基于粗粒度和细粒度特征匹配的视频描述评估，EMScore: Evaluating Video Captioning via Coarse-Grained and Fine-Grained Embedding Matching

【CVPR 2022】基于粗粒度和细粒度特征匹配的视频描述评估，EMScore: Evaluating Video Captioning via Coarse-Grained and Fine-Grained Embedding Matching

专知会员服务

10+阅读 · 2022年3月19日

最新《Transformers模型》教程，64页ppt

最新《Transformers模型》教程，64页ppt

专知会员服务

321+阅读 · 2020年11月26日

100+篇《自监督学习(Self-Supervised Learning)》论文最新合集

100+篇《自监督学习(Self-Supervised Learning)》论文最新合集

专知会员服务

165+阅读 · 2020年3月18日

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

专知会员服务

59+阅读 · 2019年10月17日

强化学习最新教程，17页pdf

强化学习最新教程，17页pdf

专知会员服务

182+阅读 · 2019年10月11日

[综述]深度学习下的场景文本检测与识别

[综述]深度学习下的场景文本检测与识别

专知会员服务

78+阅读 · 2019年10月10日

【人工智能在2019：一年回顾】反人工智能，AI in 2019: A Year in Review

【人工智能在2019：一年回顾】反人工智能，AI in 2019: A Year in Review

专知会员服务

79+阅读 · 2019年10月10日

【哈佛大学商学院课程Fall 2019】机器学习可解释性

【哈佛大学商学院课程Fall 2019】机器学习可解释性

专知会员服务

105+阅读 · 2019年10月9日

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

专知会员服务

41+阅读 · 2019年10月9日

热门VIP内容

开通专知VIP会员享更多权益服务

《战区安全决策课程体系》最新244页

《"无人机航母"原型平台》

任务规划与地形分析：现代复杂环境作战导航体系

《攻击场景描述形式化模型研究》

相关资讯

ACM MM 2022 Call for Papers

ACM MM 2022 Call for Papers

CCF多媒体专委会

5+阅读 · 2022年3月29日

【ICIG2021】Latest News & Announcements of the Plenary Talk1

【ICIG2021】Latest News & Announcements of the Plenary Talk1

中国图象图形学学会CSIG

0+阅读 · 2021年11月1日

【ICIG2021】Latest News & Announcements of the Industry Talk2

【ICIG2021】Latest News & Announcements of the Industry Talk2

中国图象图形学学会CSIG

0+阅读 · 2021年7月29日

【ICIG2021】Latest News & Announcements of the Industry Talk1

【ICIG2021】Latest News & Announcements of the Industry Talk1

中国图象图形学学会CSIG

0+阅读 · 2021年7月28日

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

Unsupervised Learning via Meta-Learning

Unsupervised Learning via Meta-Learning

CreateAMind

43+阅读 · 2019年1月3日

disentangled-representation-papers

disentangled-representation-papers

CreateAMind

26+阅读 · 2018年9月12日

【论文推荐】最新十篇目标跟踪相关论文—多帧光流跟踪、动态图学习、MV-YOLO、姿态估计、深度核相关滤波、Benchmark

【论文推荐】最新十篇目标跟踪相关论文—多帧光流跟踪、动态图学习、MV-YOLO、姿态估计、深度核相关滤波、Benchmark

专知

13+阅读 · 2018年5月26日

【论文推荐】最新六篇图像分割相关论文—控制、全卷积网络、子空间表示、多模态图像分割

【论文推荐】最新六篇图像分割相关论文—控制、全卷积网络、子空间表示、多模态图像分割

专知

25+阅读 · 2018年4月15日

【论文】变分推断（Variational inference)的总结

【论文】变分推断（Variational inference)的总结

机器学习研究会

39+阅读 · 2017年11月16日

相关论文

Learning 3D Object Shape and Layout without 3D Supervision

Arxiv

0+阅读 · 2022年6月14日

Object Scene Representation Transformer

Arxiv

0+阅读 · 2022年6月14日

A Survey of Automated Data Augmentation Algorithms for Deep Learning-based Image Classication Tasks

Arxiv

1+阅读 · 2022年6月14日

Looking Outside the Box to Ground Language in 3D Scenes

Arxiv

0+阅读 · 2022年6月13日

Learn2Augment: Learning to Composite Videos for Data Augmentation in Action Recognition

Arxiv

0+阅读 · 2022年6月9日

Transformers in Time Series: A Survey

Arxiv

34+阅读 · 2022年2月15日

A Survey on Data Augmentation for Text Classification

A Survey on Data Augmentation for Text Classification

Arxiv

16+阅读 · 2021年7月7日

Data Augmentation for Graph Neural Networks

Arxiv

38+阅读 · 2020年12月2日

Learning from Few Samples: A Survey

Learning from Few Samples: A Survey

Arxiv

77+阅读 · 2020年7月30日

Graph Convolutional Networks for Text Classification

Arxiv

11+阅读 · 2018年10月17日

相关基金

温敏二嵌段胶束种子大分子RAFT试剂调介下的种子分散RAFT聚合

国家自然科学基金

0+阅读 · 2014年12月31日

常染色体隐性遗传小脑性共济失调新的致病基因CAX的功能研究

国家自然科学基金

0+阅读 · 2014年12月31日

玉米异染色质纽(Knob)形成的表观遗传机制及进化分析

国家自然科学基金

0+阅读 · 2013年12月31日

非编码RNA调控RNA聚合酶II转录的研究

国家自然科学基金

0+阅读 · 2013年12月31日

鱼类ADAR1剪接异构体基因的鉴定及其转录调控

国家自然科学基金

0+阅读 · 2012年12月31日

镧系硅氧氮化物荧光材料的晶体结构和发光特性

国家自然科学基金

0+阅读 · 2012年12月31日

抗肿瘤药物多功能纳米自组装载体逆转肿瘤多药耐药性及其机制研究

国家自然科学基金

0+阅读 · 2012年12月31日

HGF诱导NSCLC细胞对EGFR-TKIs耐药机制的研究。

国家自然科学基金

0+阅读 · 2011年12月31日

基于list-mode数据的快速SART真3D PET断层重建算法的研究

国家自然科学基金

0+阅读 · 2011年12月31日

PGK1诱导肿瘤基质成纤维细胞激活在前列腺癌发展、转移中作用的研究

国家自然科学基金

0+阅读 · 2009年12月31日

微信扫码咨询专知VIP会员