<?xml version="1.0" encoding="utf-8" standalone="yes" ?>
<rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom">
  <channel>
    <title>数据集构建 | ViLab</title>
    <link>https://vilab.team/tag/%E6%95%B0%E6%8D%AE%E9%9B%86%E6%9E%84%E5%BB%BA/</link>
      <atom:link href="https://vilab.team/tag/%E6%95%B0%E6%8D%AE%E9%9B%86%E6%9E%84%E5%BB%BA/index.xml" rel="self" type="application/rss+xml" />
    <description>数据集构建</description>
    <generator>Hugo Blox Builder (https://hugoblox.com)</generator><language>en-us</language><lastBuildDate>Mon, 20 Apr 2026 00:00:00 +0000</lastBuildDate>
    <image>
      <url>https://vilab.team/media/icon_hu2896232876136423579.png</url>
      <title>数据集构建</title>
      <link>https://vilab.team/tag/%E6%95%B0%E6%8D%AE%E9%9B%86%E6%9E%84%E5%BB%BA/</link>
    </image>
    
    <item>
      <title>ReactID: Synchronizing Realistic Actions and Identity in Personalized Video Generation</title>
      <link>https://vilab.team/publication/reactid-synchronizing-realistic-actions-and-identity-in-pers/</link>
      <pubDate>Mon, 20 Apr 2026 00:00:00 +0000</pubDate>
      <guid>https://vilab.team/publication/reactid-synchronizing-realistic-actions-and-identity-in-pers/</guid>
      <description>&lt;p&gt;本文提出ReactID框架，旨在协调个性化视频生成中身份一致性与动作真实性的矛盾。针对主体-视频对齐不精确、训练不稳定、细粒度动作建模不足三大挑战，从数据、训练和动作建模三方面协同改进：构建高精度标注的ReactID-Data数据集；设计由易到难的渐进式训练课程；提出基于时间线的条件机制，通过主体感知交叉注意力和时间自适应RoPE，将子动作与特定主体绑定并嵌入时间坐标，从而生成更自然、可控的视频。&lt;/p&gt;
</description>
    </item>
    
    <item>
      <title>祝贺实验室科研成果发表于 ICLR 2026！</title>
      <link>https://vilab.team/event/%E7%A5%9D%E8%B4%BA%E5%AE%9E%E9%AA%8C%E5%AE%A4%E7%A7%91%E7%A0%94%E6%88%90%E6%9E%9C%E5%8F%91%E8%A1%A8%E4%BA%8E-iclr/</link>
      <pubDate>Mon, 20 Apr 2026 00:00:00 +0000</pubDate>
      <guid>https://vilab.team/event/%E7%A5%9D%E8%B4%BA%E5%AE%9E%E9%AA%8C%E5%AE%A4%E7%A7%91%E7%A0%94%E6%88%90%E6%9E%9C%E5%8F%91%E8%A1%A8%E4%BA%8E-iclr/</guid>
      <description>&lt;p&gt;热烈祝贺李蔚同学！论文《ReactID: Synchronizing Realistic Actions and Identity in Personalized Video Generation》已发表在 &lt;em&gt;ICLR 2026&lt;/em&gt;。&lt;/p&gt;
&lt;h2 id=&#34;reactid-synchronizing-realistic-actions-and-identity-in-personalized-video-generation&#34;&gt;ReactID: Synchronizing Realistic Actions and Identity in Personalized Video Generation&lt;/h2&gt;
&lt;p&gt;&lt;strong&gt;祝贺李蔚同学！该论文已发表在 &lt;em&gt;ICLR&lt;/em&gt;。&lt;/strong&gt;&lt;/p&gt;
&lt;p&gt;















&lt;figure  &gt;
  &lt;div class=&#34;d-flex justify-content-center&#34;&gt;
    &lt;div class=&#34;w-100&#34; &gt;&lt;img alt=&#34;ReactID: Synchronizing Realistic Actions and Identity in Personalized Video Generation&#34; srcset=&#34;
               /event/%E7%A5%9D%E8%B4%BA%E5%AE%9E%E9%AA%8C%E5%AE%A4%E7%A7%91%E7%A0%94%E6%88%90%E6%9E%9C%E5%8F%91%E8%A1%A8%E4%BA%8E-iclr/images/paper-01_hu10185920340004675066.webp 400w,
               /event/%E7%A5%9D%E8%B4%BA%E5%AE%9E%E9%AA%8C%E5%AE%A4%E7%A7%91%E7%A0%94%E6%88%90%E6%9E%9C%E5%8F%91%E8%A1%A8%E4%BA%8E-iclr/images/paper-01_hu15268550746155155384.webp 760w,
               /event/%E7%A5%9D%E8%B4%BA%E5%AE%9E%E9%AA%8C%E5%AE%A4%E7%A7%91%E7%A0%94%E6%88%90%E6%9E%9C%E5%8F%91%E8%A1%A8%E4%BA%8E-iclr/images/paper-01_hu4770338007721801293.webp 1200w&#34;
               src=&#34;https://vilab.team/event/%E7%A5%9D%E8%B4%BA%E5%AE%9E%E9%AA%8C%E5%AE%A4%E7%A7%91%E7%A0%94%E6%88%90%E6%9E%9C%E5%8F%91%E8%A1%A8%E4%BA%8E-iclr/images/paper-01_hu10185920340004675066.webp&#34;
               width=&#34;760&#34;
               height=&#34;288&#34;
               loading=&#34;lazy&#34; data-zoomable /&gt;&lt;/div&gt;
  &lt;/div&gt;&lt;/figure&gt;
&lt;/p&gt;
&lt;ul&gt;
&lt;li&gt;&lt;strong&gt;作者：&lt;/strong&gt; Wei Li、Yiheng Zhang、Fuchen Long、Zhaofan Qiu、Ting Yao、Xiaoyan Sun、Tao Mei&lt;/li&gt;
&lt;li&gt;&lt;strong&gt;发表载体：&lt;/strong&gt; &lt;em&gt;ICLR&lt;/em&gt;&lt;/li&gt;
&lt;li&gt;&lt;strong&gt;发表时间：&lt;/strong&gt; 2026年4月20日&lt;/li&gt;
&lt;li&gt;&lt;strong&gt;相关链接：&lt;/strong&gt; &lt;a href=&#34;https://proceedings.iclr.cc/paper_files/paper/2026/hash/6600458132d025c68a01e82081597b32-Abstract-Conference.html&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;论文链接&lt;/a&gt; · &lt;a href=&#34;https://proceedings.iclr.cc/paper_files/paper/2026/file/6600458132d025c68a01e82081597b32-Paper-Conference.pdf&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;PDF&lt;/a&gt;&lt;/li&gt;
&lt;/ul&gt;
&lt;h3 id=&#34;论文介绍&#34;&gt;论文介绍&lt;/h3&gt;
&lt;p&gt;本文提出ReactID框架，旨在协调个性化视频生成中身份一致性与动作真实性的矛盾。针对主体-视频对齐不精确、训练不稳定、细粒度动作建模不足三大挑战，从数据、训练和动作建模三方面协同改进：构建高精度标注的ReactID-Data数据集；设计由易到难的渐进式训练课程；提出基于时间线的条件机制，通过主体感知交叉注意力和时间自适应RoPE，将子动作与特定主体绑定并嵌入时间坐标，从而生成更自然、可控的视频。&lt;/p&gt;
</description>
    </item>
    
    <item>
      <title>Event-based head pose estimation: Benchmark and method</title>
      <link>https://vilab.team/publication/event-based-head-pose-estimation-benchmark-and-method/</link>
      <pubDate>Sun, 29 Sep 2024 00:00:00 +0000</pubDate>
      <guid>https://vilab.team/publication/event-based-head-pose-estimation-benchmark-and-method/</guid>
      <description>&lt;p&gt;本文针对传统RGB方法在剧烈运动和极端光照下头部姿态估计困难的问题，引入事件相机的高时间分辨率与高动态范围优势。作者构建了两个大规模事件头部姿态数据集，包含282个序列，覆盖不同分辨率与场景；并提出事件头部姿态估计网络EV-HPE，设计了事件时空融合模块和事件运动感知注意力模块，有效结合事件流时空信息，提升姿态估计精度与鲁棒性。&lt;/p&gt;
</description>
    </item>
    
    <item>
      <title>Panacea: Panoramic and controllable video generation for autonomous driving</title>
      <link>https://vilab.team/publication/panacea-panoramic-and-controllable-video-generation-for-auto/</link>
      <pubDate>Mon, 01 Jan 2024 00:00:00 +0000</pubDate>
      <guid>https://vilab.team/publication/panacea-panoramic-and-controllable-video-generation-for-auto/</guid>
      <description>&lt;p&gt;本文提出Panacea，一种面向自动驾驶场景的全景可控视频生成方法。该方法通过创新的4D注意力机制和两阶段生成流程，有效解决了生成视频中的时间与跨视角一致性问题；同时引入ControlNet框架，利用鸟瞰图（BEV）布局对生成内容进行精细控制。在nuScenes数据集上的实验表明，Panacea能够生成高质量的多视角驾驶视频，为BEV感知任务提供丰富的训练数据增强，推动自动驾驶感知技术的发展。&lt;/p&gt;
</description>
    </item>
    
  </channel>
</rss>
