Get in Touch

Course Outline

Hunyuan Multimodal Foundations and Lab Setup

  • Explore Hunyuan’s multimodal capabilities for image, 3D, and video scenarios
  • Identify relevant business use cases for creative, product, and content teams
  • Prepare the lab environment, sample assets, and model access credentials
  • Execute initial generation tasks and evaluate the results

Prompt Design and Workflow Patterns

  • Structure prompts to achieve consistent multimodal outcomes
  • Utilize text prompts, reference images, and fundamental input settings
  • Select appropriate workflows for generating images, videos, or 3D content
  • Refine prompts based on output quality and business objectives

Image Generation and Review Labs

  • Produce marketing, product, and concept imagery from text prompts
  • Adjust visual style, composition, and content consistency
  • Evaluate outputs for utility, quality, and brand alignment
  • Organize image assets for approval and subsequent use

Video Generation Labs

  • Create short video clips from prompts and prepared inputs
  • Manage style, scene direction, and output variation
  • Assess videos for clarity, continuity, and practical applicability
  • Prepare video assets for demos or content pipelines

3D Asset Creation Labs

  • Generate foundational 3D assets from text or image inputs
  • Verify geometry, texture quality, and asset usability
  • Export assets for visualization, prototyping, or content workflows
  • Determine when 3D generation is preferable to image or video approaches

Integration, Governance, and Next Steps

  • Publish generated assets via lightweight applications, services, or APIs
  • Link multimodal outputs to product, content, and review systems
  • Implement checks for quality, brand safety, copyright compliance, and responsible usage
  • Outline pilot use cases and strategic steps for internal adoption

Requirements

  • Fundamental knowledge of AI and generative AI principles
  • Practical experience with web applications, APIs, or standard developer tools
  • Basic proficiency in Python or scripting languages

Audience

  • Developers creating AI-driven product features
  • Technical product managers and solution architects
  • Innovation, media, and digital teams managing image, video, or 3D content
 14 Hours

Number of participants


Price per participant

Upcoming Courses

Related Categories