Reachy Mini recognizes itself in the mirror
Remi plugged GPT-4o on Reachy Mini and it's pretty cool! Some new capabilities: - Image analysis: Reachy Mini can now look at a photo it just took and describe or reason about it - Face tracking: keeps eye contact and makes interactions feel much more natural - Motion fusion: [head wobble while speaking] + [face tracking] + [emotions or dances] can now run simultaneously - Face recognition: runs locally - Autonomous behaviors when idle: when nothing happens for a while, the model can decide to trigger context-based behaviors Questions for the community: • Earlier versions used flute sounds when playing emotions. This one speaks instead (for example the "olala" at the start is an emotion + voice). It completely changes how I perceive the robot (pet? human? kind alien?). Should we keep a toggle to switch between voice and flute sounds? • How do the response delays feel to you? Some limitations: No memory system yet No voice recognition yet Strategy in crowds still unclear: the VAD (voice activity detection) tends to activate too often, and we don’t like the keyword approach More details about Reachy Mini: https://huggingface.co/blog/reachy-mini