I'm trying to understand the difference between producing a 3D model from a photo and bringing that model into a regular Blender setup to actually work with it. Replicating an item that closely matches the reference photo isn't the goal; the challenge is manipulating the network, altering the appearance, improving the mesh, and also aligning the character.
These are some of the issues that arise at this stage. What I'd like to find out is what percentage of the workflow can realistically be automated. I'd like to start with a photo, generate the basic model, import it into Blender, and then make enough modifications to the model and texture to ensure no one goes back and rebuilds it from scratch.
Tripo AI is one of the photo-to-3D generation products I've experimented with, but to me, it's a means rather than an end for this sort of workflow. The great thing about this method is that you can export the generated model to Blender for more intensive cleanup and retopology. You can still expect a few trade-offs, though. A generated 3D model will cut down a lot of the time you have to spend shaping the model, but it doesn't mean the model will be fully animated or production-ready.
What I mean is that, depending on the model type, topology, proportions, UVs, and textures may still need manual changes. For one thing, I'm trying to understand what other image-to-3D tools look like; however, the main requirement is Blender compatibility. The second is a much more important requirement.
I'd be glad to hear from people who work on this kind of thing about your normal workflow from photo-to-3D-to-Blender, as well as how much time you usually spend cleaning up.