AgentGraph: Toward Universal Dialogue Management With Structured Deep Reinforcement Learning

Top Reasons to Join SPS Today!

1. IEEE Signal Processing Magazine
2. Signal Processing Digital Library*
3. Inside Signal Processing Newsletter
4. SPS Resource Center
5. Career advancement & recognition
6. Discounts on conferences and publications
7. Professional networking
8. Communities for students, young professionals, and women
9. Volunteer opportunities
10. Coming soon! PDH/CEU credits
Click here to learn more.

TASLP Volume 27 Issue 9

AgentGraph: Toward Universal Dialogue Management With Structured Deep Reinforcement Learning

TASLPRO Featured Articles

By:

Lu Chen; Zhi Chen; Bowen Tan; Sishan Long; Milica Gašić; Kai Yu

Dialogue policy plays an important role in task-oriented spoken dialogue systems. It determines how to respond to users. The recently proposed deep reinforcement learning (DRL) approaches have been used for policy optimization. However, these deep models are still challenging for two reasons: first, many DRL-based policies are not sample efficient; and second, most models do not have the capability of policy transfer between different domains. In this paper, we propose a universal framework, AgentGraph , to tackle these two problems. The proposed AgentGraph is the combination of graph neural network (GNN) based architecture and DRL-based algorithm. It can be regarded as one of the multi-agent reinforcement learning approaches. Each agent corresponds to a node in a graph, which is defined according to the dialogue domain ontology. When making a decision, each agent can communicate with its neighbors on the graph. Under AgentGraph framework, we further propose dual GNN-based dialogue policy, which implicitly decomposes the decision in each turn into a high-level global decision and a low-level local decision. Experiments show that AgentGraph models significantly outperform traditional reinforcement learning approaches on most of the 18 tasks of the PyDial benchmark. Moreover, when transferred from the source task to a target task, these models not only have acceptable initial performance but also converge much faster on the target task.

Read on IEEE Xplore

Tags:

IEEE TASLP Article

SPS Social Media

IEEE SPS Facebook Page https://www.facebook.com/ieeeSPS
IEEE SPS X Page https://x.com/IEEEsps
IEEE SPS Instagram Page https://www.instagram.com/ieeesps/?hl=en
IEEE SPS LinkedIn Page https://www.linkedin.com/company/ieeesps/
IEEE SPS YouTube Channel https://www.youtube.com/ieeeSPS

IEEE SPS Educational Resources

IEEE SPS Resource Center

IEEE SPS YouTube Channel

© Copyright 2025 IEEE - All rights reserved. Use of this website signifies your agreement to the IEEE Terms and Conditions.
A public charity, IEEE is the world's largest technical professional organization dedicated to advancing technology for the benefit of humanity.

congratulations.jpg

Congratulations to Signal Processing Society Members Elevated to Senior Members!

MLSP-2027.jpg

2027 IEEE International Workshop on Machine Learning for Signal Processing (MLSP 2027)

ISPA-2025.jpg

2025 14th International Symposium on Image and Signal Processing and Analysis (ISPA)

What is Signal Processing?

Popular Pages

Today's:

All time:

Last viewed:

AgentGraph: Toward Universal Dialogue Management With Structured Deep Reinforcement Learning

Publications & Resources

For Authors

congratulations.jpg

CAI_2027_Call_for_Proposals.png

pod .png

Top Reasons to Join SPS Today!

AgentGraph: Toward Universal Dialogue Management With Structured Deep Reinforcement Learning

SPS Social Media

IEEE SPS Educational Resources

What is Signal Processing?

Popular Pages

Today's:

All time:

Last viewed:

AgentGraph: Toward Universal Dialogue Management With Structured Deep Reinforcement Learning

Search form

You are here

Publications & Resources

For Authors

Top Reasons to Join SPS Today!

AgentGraph: Toward Universal Dialogue Management With Structured Deep Reinforcement Learning

SPS Social Media

IEEE SPS Educational Resources