◇ 今日论文 · AVA-Encoder: Towards Agent-Native Video Representation Learning

条目 ID:ASR-PAPER-20260816-2608-12313 | 类型:论文 | schema v0.3

人读卡片 · 五问

① 这是什么

今日论文《今日论文 · AVA-Encoder: Towards Agent-Native Video Representation Learning》(Chuyue Li、Jinpeng Yu、Haozhe Wang、Tian Xueyun、Zhijing Zhang、Bingnan Li),来自 HF Daily Papers 公开源。

② 值不值得用

进入 HF 每日精选说明有社区关注度;是否与你的问题相关需要读原文判断,本条目不评价研究质量。

③ 怎么开始

先读摘要,需要再读全文(https://huggingface.co/papers/2608.12313);对照论文 ID(2608.12313)可找实现与讨论。

④ 关键证据

  • 论文 ID 2608.12313
  • 作者 Chuyue Li、Jinpeng Yu、Haozhe Wang、Tian Xueyun、Zhijing Zhang、Bingnan Li

⑤ 注意事项与坑

摘要来自作者原文,未做同行评议级核验;引用请以正式发表版本为准。

打开来源 ↗
信任层
✅ 机检 通过 💰 升维成本 $0.0000 🧭 置信度 中 📌 schema 0.3 🛂 审核 agent reviewed

相关推荐(反茧房 · 应该知道的)

反馈

这条不对 / 我执行报错了?带 entry_id 一键提 Issue,回流后修订版本。

🐞 纠错 / 报错 → 提 Issue
AI 区块 · 机器可读(同一事实源)
{
  "schema_version": "0.3",
  "entry_id": "ASR-PAPER-20260816-2608-12313",
  "record_type": "paper",
  "title": "今日论文 · AVA-Encoder: Towards Agent-Native Video Representation Learning",
  "observed_at": "2026-08-16",
  "source": {
    "name": "huggingface.co/api/daily_papers",
    "url": "https://huggingface.co/api/daily_papers?limit=10",
    "snapshot_hash": "735b0882f0ad228e42c868385e901b1b1e78b57fb562534343256eb58fd33f9a",
    "fetched_at": "2026-08-16T10:45:59+00:00"
  },
  "confidence": "medium",
  "human_stream": {
    "summary": "Creative agents still lack an effective way to learn from high-quality human films, limiting their ability to produce cinematic-grade videos. A key challenge is the absence of a structured video representation that is both faithful to film content and directly usable for agentic reasoning and manipulation. To address the challenge, we propose the Agentic Video Auto-Encoder (AVA-Encoder), a framewo",
    "key_points": [
      "论文 ID 2608.12313",
      "作者 Chuyue Li、Jinpeng Yu、Haozhe Wang、Tian Xueyun、Zhijing Zhang、Bingnan Li",
      "发布 2026-08-12T00:00:00.000Z"
    ],
    "note": "来源为 HuggingFace daily papers 公开 API;完整阅读请走原文链接。",
    "what_it_is": "今日论文《今日论文 · AVA-Encoder: Towards Agent-Native Video Representation Learning》(Chuyue Li、Jinpeng Yu、Haozhe Wang、Tian Xueyun、Zhijing Zhang、Bingnan Li),来自 HF Daily Papers 公开源。",
    "worth_it": "进入 HF 每日精选说明有社区关注度;是否与你的问题相关需要读原文判断,本条目不评价研究质量。",
    "how_to_start": "先读摘要,需要再读全文(https://huggingface.co/papers/2608.12313);对照论文 ID(2608.12313)可找实现与讨论。",
    "evidence": [
      "论文 ID 2608.12313",
      "作者 Chuyue Li、Jinpeng Yu、Haozhe Wang、Tian Xueyun、Zhijing Zhang、Bingnan Li"
    ],
    "cautions": "摘要来自作者原文,未做同行评议级核验;引用请以正式发表版本为准。"
  },
  "ai_stream": {
    "structured": {
      "paper_id": "2608.12313",
      "title": "AVA-Encoder: Towards Agent-Native Video Representation Learning",
      "authors": [
        "Chuyue Li",
        "Jinpeng Yu",
        "Haozhe Wang",
        "Tian Xueyun",
        "Zhijing Zhang",
        "Bingnan Li",
        "Shuqi Gu",
        "Kan Ren",
        "Jiaming Liu",
        "Ruihua Hua"
      ],
      "published_at": "2026-08-12T00:00:00.000Z",
      "url": "https://huggingface.co/papers/2608.12313",
      "source_label": "hf-daily-papers"
    },
    "raw": [
      {
        "paper_id": "2608.12313",
        "title": "AVA-Encoder: Towards Agent-Native Video Representation Learning",
        "authors": [
          "Chuyue Li",
          "Jinpeng Yu",
          "Haozhe Wang",
          "Tian Xueyun",
          "Zhijing Zhang",
          "Bingnan Li",
          "Shuqi Gu",
          "Kan Ren",
          "Jiaming Liu",
          "Ruihua Hua"
        ],
        "summary": "Creative agents still lack an effective way to learn from high-quality human films, limiting their ability to produce cinematic-grade videos. A key challenge is the absence of a structured video representation that is both faithful to film content and directly usable for agentic reasoning and manipulation. To address the challenge, we propose the Agentic Video Auto-Encoder (AVA-Encoder), a framework for learning agent-native video representations via agentic auto-encoding.\n  AVA-Encoder transforms a video into a knowledge graph (KG) representation and then reconstructs it back into video. Its ",
        "published_at": "2026-08-12T00:00:00.000Z",
        "submitted_on_daily_at": null,
        "url": "https://huggingface.co/papers/2608.12313"
      }
    ]
  },
  "token_cost": {
    "total": 0.0,
    "currency": "USD",
    "breakdown": {
      "crawl": 0.0,
      "clean": 0.0,
      "elevate": 0.0,
      "verify": 0.0
    }
  },
  "machine_verified": true,
  "review": {
    "status": "agent_reviewed",
    "reviewer": "pipeline-validate",
    "reviewed_at": "2026-08-16T12:12:04+00:00",
    "comments": "机检通过(schema 0 error + 溯源一致 + 成本达标)"
  },
  "provenance": {
    "extracted_by": "codex",
    "extracted_at": "2026-08-16T10:45:59+00:00",
    "pipeline": "collect_hf_papers.py + build_entries.py v0.2(规则管线)",
    "access_urls": [
      "https://huggingface.co/api/daily_papers?limit=10"
    ],
    "card_generated_at": "2026-08-16T12:12:03+00:00",
    "card_pipeline": "enrich_human_stream.py v0.1(规则模板,零 Token)"
  }
}