Thanks for sharing. I've taken a look on this record.
Indeed, it can hardly be tracked by Mocap Studio. The issues I see:
- Large distance between actors and sensors (about 4m). The quality of depth data provided by Kinect decreases with distance from the sensor. Obviously, the number of depth pixels per body is low at that distance.
- Actors mostly take poses where body parts are close to each other (and often to the floor as well). That is, depth pixels can hardly be distinguished to belong to a specific body part.
- Child's pixels are partially confused with background (floor) because the limbs are pretty thin.
I suppose that color cameras (PS Eyes) may be more appropriate for capturing in these conditions. Though I'm not completely sure about this. However, using color cameras imposes addition requirements on actor clothing (solid colors) and environment (bright even lighting).
Regarding using depth data from your recordings. I saw in your other post that you already own a license for Biomech add-on. Can't you just use point cloud exported via Biomech add-on as raw input for your purposes?