EpiDiff is a generative model based on Zero123 that takes an image of an object as a conditioning frame, and generates 16 multiviews of that object.
For usage instructions, please refer to
our EpiDiff GitHub repository.
We use renders from the LVIS dataset, utilizing
huanngzh/render-toolbox.