This disclosure relates generally to method and system for performing negotiation task using reinforcement learning agents. Performing negotiation on a task is a complex decision making process and to arrive at consensus on contents of a negotiation task is often expensive and time consuming due to the negotiation terms and the negotiation parties involved. The proposed technique trains reinforcement learning agents such as negotiating agent and an opposition agent. These agents are capable of performing the negotiation task on a plurality of clauses to agree on common terms between the agents involved. The system provides modelling of a selector agent on a plurality of behavioral models of a negotiating agent and the opposition agent to negotiate against each other and provides a reward signal based on the performance. This selector agent emulate human behavior provides scalability on selecting an optimal contract proposal during the performance of the negotiation task.
展开▼