Zhiyu Zoey Chen
226
posts
Then NLP is low-dim-input-and-output, and by sheer luck we also have oceans of data
🚀In our new work on RAG efficient search, we propose a hierarchical process reward to optimize each reasoning step during training by on-the-fly detection of suboptimal searches. We boosted the accuracy and slashed the over-search rate from over 27% to just 2.3%! 📄 Paper:
Our new benchmark has been accepted to #EMNLP2025 main conference! A well curated dataset to test LLM agents to reproduce LM research papers. ⬇️





