Loading Open Internet
    What are the biggest challenges in collecting high-quality speech and egocentric video datasets? [D]