İki taraflı müzakere için aktör-kritik pekiştirmeli öğrenme yaklaşımı
Bu tez size mi ait?
Bu kayıt toplu arşivden geldi. Sizinse profilinize bağlayın.
Özet (EN)
Designing an effective and intelligent bidding strategy is one of the most compelling research challenges in automated negotiation, where software agents negotiate with each other to find a mutual agreement when there is a conflict of interests. Instead of designing a hand-crafted decision-making module, this thesis proposes a novel bidding strategy adopting an actor-critic reinforcement learning approach, which learns what to offer in a bilateral negotiation. An entropy reinforcement learning framework called \acrfull{sac} is applied to the bidding problem, and a self-play approach is employed to train the model determining the target utility of the coming offer based on previous offer exchanges and remaining time. Furthermore, an imitation learning approach called behavior cloning is adopted to speed up the learning process. Also, a novel reward function is introduced that does not only take the agent's own utility, but also the opponent's utility at the end of the negotiation. The developed agent is empirically evaluated. Thus, a large number of negotiation sessions are run against a variety of opponents selected in different domains varying in size and opposition. The agent's performance is compared with its opponents and the performance of the baseline agents negotiating with the same opponents. The empirical results show that our agent successfully negotiates against challenging opponents in different negotiation scenarios without requiring any former information about the opponent or domain in advance. Furthermore, it achieves better results than the baseline agents regarding the received utility at the end of the successful negotiations.
Yazar
Furkan Arslan
Kurum
Bu Yayına Nasıl Atıf Yapılır
Furkan Arslan (Master Thesis). İki taraflı müzakere için aktör-kritik pekiştirmeli öğrenme yaklaşımı, 2021, Özyeğin University.
Anahtar Kelimeler
Lisans
Tüm Hakları Saklıdır
Bu eser belirtilen lisans koşulları altında paylaşılmaktadır.
Özyeğin University tezlerinden daha fazlası
- A metaheuristic approach for multiple-item economic lot sizing problem with inventory dependent demand(2023)
- İleri karmaşık olay işleme özellikli veri akışı yönetim sisteminin tasarım ve gerçeklemesi(2013)
- Biyolojik kendiliğinden iyileşen çimento esaslı harçların performansa dayalı değerlendirilmesi(2022)
- Effective remorse provisions for drug and stimulant substances crimes in the Turkish Penal Code(2023)
- Bina bölütlemesi ve yükseklik tahmini için görsel durum-uzayı tabanlı çoklu görevli öğrenme(2025)
- Tam ka-bant uydu haberleşmesi için çift dairesel kutuplamalı horn anten ve besleme ağı(2025)
