Skip to content

Roadmap: photorealistic / video avatar visual layer #11

Description

@ykshv

Goal

Design and implement a replaceable visual layer for a more realistic avatar path.

Scope

  • Keep the current VRM path working.
  • Define the boundary for TTS audio -> audio/video renderer -> WebRTC video track.
  • Preserve session_id, turn_id, generation_id, branch_state, seq, and pts_ms semantics.
  • Document latency, consent, licensing, and GPU implications.

Acceptance

  • A design note or implementation plan exists.
  • Stale video/avatar frames cannot survive a generation change.
  • The README can clearly distinguish VRM mode from future photorealistic/video mode.

Activity

Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Metadata

Metadata

Assignees

No one assigned

    Labels

    enhancementNew feature or requestroadmapPlanned public roadmap workvisual-layerAvatar rendering, video avatar, and visual behavior

    Projects

    No projects

      Milestone

      No milestone

      Relationships

      None yet

      Development

      No branches or pull requests

      Issue actions