Skip to content

Roadmap: photorealistic / video avatar visual layer #11

Description

@ykshv

Goal

Design and implement a replaceable visual layer for a more realistic avatar path.

Scope

  • Keep the current VRM path working.
  • Define the boundary for TTS audio -> audio/video renderer -> WebRTC video track.
  • Preserve session_id, turn_id, generation_id, branch_state, seq, and pts_ms semantics.
  • Document latency, consent, licensing, and GPU implications.

Acceptance

  • A design note or implementation plan exists.
  • Stale video/avatar frames cannot survive a generation change.
  • The README can clearly distinguish VRM mode from future photorealistic/video mode.

Metadata

Metadata

Assignees

No one assigned

    Labels

    enhancementNew feature or requestroadmapPlanned public roadmap workvisual-layerAvatar rendering, video avatar, and visual behavior

    Projects

    No projects

    Milestone

    No milestone

    Relationships

    None yet

    Development

    No branches or pull requests

    Issue actions