The finding here should make anyone shipping a voice agent pause: the retrieval tricks we add to improve multi-hop answers can make the system less robust once a speech recognizer sits in front of it.

If your RAG eval only ever sees clean, typed queries, youโ€™re grading the easy half of the problem. Worth reading against the HF paper page before you assume that more retrieval structure is strictly safer.

tags: [ rag ] [ conversational-ai ] [ knowledge-graphs ] [ research ]