USTC · Cyber Science

Ruizhe Li

李睿哲

Master's student researching LLM agents, agent memory, and long-context language models — how agents retrieve, retain, and use information over long horizons.

University of Science and Technology of China · Hefei
School of Cyber Science
Cross-Media Intelligence Group (USTC-CMI)
Ruizhe Li in the mountains
01About

I'm Ruizhe Li (李睿哲), a master's student at USTC. My research focuses on LLM agents, agent memory, and long-context language models.

LLM Agents Agent Memory Long-Context LLMs
02News
03Education
2026.09 — Present

Master's Student

University of Science and Technology of China · School of Cyber Science

USTC-CMI · Hefei, China

Research interests: LLM agents, agent memory, and long-context language models.

2022.09 — 2026.06

B.S. in Data Science

University of Science and Technology of China · School of Artificial Intelligence and Data Science

Hefei, China · Top 10%

04Publications
Preprint · 2026

When Your Agent Opens the Chat App: Agent-Controlled Search over Raw Chat Logs Rivals Structured Memory

Ruizhe Li, Licheng Zhang, Benfeng Xu, Mingxuan Du, Zheren Fu, Weidong Chen

arXiv:2608.12888 · 2026

ReFind gives an agent iterative, controllable search over unmodified chat logs, combining lexical retrieval with session-aware ranking, local context, and temporal narrowing. It reaches the highest mean accuracy among compared systems on MemoryAgentBench without constructing summaries, trees, or knowledge graphs.

Preprint · 2026

Keep It InMind: Benchmarking the Implicit-Association Blind Spot in Agent Memory

Ruizhe Li, Mingxuan Du, Benfeng Xu, Zhendong Mao

arXiv:2607.24368 · 2026

InMind is a 125-task benchmark showing that a memory can be crucial even when it is not textually similar to the current query. Its paired controls isolate whether failures come from storage, missing world knowledge, or retrieval and routing.

Preprint · 2026

DeepResearch Bench II: Diagnosing Deep Research Agents via Rubrics from Expert Reports

Ruizhe Li*, Mingxuan Du*, Benfeng Xu, Chiwei Zhu, Xiaorui Wang, Zhendong Mao · *equal contribution

arXiv:2601.08536 · 2026

A benchmark of 132 expert-grounded tasks across 22 domains and 9,430 fully-verifiable binary rubrics, scoring deep-research reports on information recall, analysis, and presentation. Built with a four-stage LLM+human pipeline and 400+ expert-hours — even the strongest systems satisfy fewer than 50% of rubrics.

Findings of EMNLP 2025

Automated Creativity Evaluation for Large Language Models: A Reference-Based Approach

Ruizhe Li, Chiwei Zhu, Benfeng Xu, Xiaorui Wang, Zhendong Mao

Findings of EMNLP 2025 · arXiv:2504.15784

A reference-based method that scores the creativity of large-language-model outputs against human references, offering a more reliable and automatic alternative to costly human creativity ratings.

CIKM 2024

Reformulating Conversational Recommender Systems as Tri-Phase Offline Policy Learning

Gangyi Zhang, Chongming Gao, Hang Pan, Runzhe Teng, Ruizhe Li

CIKM 2024 · arXiv:2408.06809

Recasts conversational recommendation as a three-phase offline policy-learning problem, decoupling interaction-policy learning from costly online exploration.

05Projects
ReFind★ 9

Agent-controlled search over raw chat logs for long-term memory. ReFind combines iterative lexical retrieval with chat-native controls and rivals structured memory systems without building summaries, trees, or knowledge graphs.

Python · Agent MemoryRepo ↗
InMind★ 17

A 125-task benchmark for the implicit-association blind spot in long-term agent memory, with paired controls that isolate storage, world-knowledge, retrieval, and application failures.

Benchmark · DatasetRepo ↗
DeepResearch-Bench-II★ 83

Evaluation suite and leaderboard for deep-research agents, grounded in expert-written reports. Rubric extraction, batched LLM-as-judge scoring, and token accounting.

PythonRepo ↗
06Experience
2025.02 — 2025.08

LLM Algorithm Engineer Intern

Metastone Technology

Beijing, China

Designed the memory system for a role-play product.

07Service · Skills
  • President, Student Union School of AI & Data Science, USTC
  • Teaching Assistant, DS4001 Data Science course · 2025 Spring
  • Member, Tang Zhongying Charity Society Volunteering
PythonPyTorchLLM / AgentsRAG LaTeXGitHTML / CSS / JSC