본문

서브메뉴

Towards Positive Outcomes in the AI Economy: Mitigating Algorithmic Collusion and Enabling Fair Recourse
Towards Positive Outcomes in the AI Economy: Mitigating Algorithmic Collusion and Enabling...
Towards Positive Outcomes in the AI Economy: Mitigating Algorithmic Collusion and Enabling Fair Recourse

상세정보

자료유형  
 학위논문 서양
최종처리일시  
20250211151451
ISBN  
9798382776545
DDC  
004
저자명  
Mibuari, Eric M.
서명/저자  
Towards Positive Outcomes in the AI Economy: Mitigating Algorithmic Collusion and Enabling Fair Recourse
발행사항  
[Sl] : Harvard University, 2024
발행사항  
Ann Arbor : ProQuest Dissertations & Theses, 2024
형태사항  
165 p
주기사항  
Source: Dissertations Abstracts International, Volume: 85-12, Section: A.
주기사항  
Advisor: Parkes, David.
학위논문주기  
Thesis (Ph.D.)--Harvard University, 2024.
초록/해제  
요약The rise of Artificial Intelligence (AI) promises to solve many important problems in the world. At the same time, awareness has been increasing about its potential and real harms. How can we extract maximum benefit from the promise of AI while minimizing present harms and mitigating future risks?In this thesis, I frame and answer this question from the perspective of enabling and promoting positive outcomes in the AI-enabled economy, where markets are facilitated using AI algorithms, including agent behavior, pricing, and matching and clearing. The goal of my research is to find and create the conditions under which the benefits of AI are preserved or even enhanced while the tendency to diminish welfare, perhaps to particular groups, is contained to the greatest possible extent.In lending domains, machine learning can be used to form a predictive model of the probability of default (a "risk score"), this driving loan decisions. For simple models, this brings the benefits of transparency and explainability, as well as guidance in regard to recourse. An alternative is to use policy learning, that is, learning a policy from borrower characteristics to loan decisions directly, and without explicit risk scoring. This emphasizes profit and can speed up learning, as a lender understands a borrower population, but with a concomitant loss of transparency. I introduce a risk-score based policy learning method, as well as a new metric of recourse effort fairness, and demonstrate that this risk-score based policy learning achieves optimal profits, explainability and transparency, as well as recourse effort fairness.There are a number of problems where economic actors follow sequential behaviors, for example, in making pricing adjustments over time on e-commerce platforms, or power trading and storage optimization through a typical day in an electricity market. In this thesis, I work with the Stackelberg POMDP framework (partially observable Markov decision process) to design interventions in these kinds of sequential settings, seeking to improve economic welfare. I am especially focused on concerns that can arise in regard to tacit collusion between AI-mediated pricing or trading and storage decisions by automated agents.In a first setting, I study algorithmic pricing on e-commerce platforms, where reinforcement learning (RL) algorithms have been shown to learn to set collusive prices with nothing more than profit feedback. This raises the question as to whether collusive pricing can be prevented through the design of suitable ``buy boxes," i.e., through the design of the rules that govern the promotion of particular products and prices to consumers. I show that RL can also be used by platforms to learn buy box rules that are effective in preventing collusion by RL sellers. For this, I adopt the Stackelberg POMDP framework, and demonstrate success in learning robust rules that provide high consumer welfare.In a second setting, I study trading and storage decisions by battery operators in electricity power markets. %requiring power storage are an example of a multi-agent dynamic storage optimization application in which artificial intelligence (AI) algorithms are actively applied by prosumers (battery operators who both produce and consume power). In this application, agents who correspond to battery operators buy and sell power in a market with producers and price-elastic consumers, and where power can be stored at a negligible cost up to an agent-specific capacity. The use of RL algorithms has been shown in this setting to lead to outcomes where battery operators arbitrage the market over time, and that correspond to tacit collusion. I again appeal to the Stackelberg POMDP framework, and demonstrate success in learning collusion-mitigation policies by a regulator, in particular through the use of a network-flow thresholding intervention.
일반주제명  
Computer science
키워드  
Algorithmic collusion
키워드  
Fair recourse
키워드  
Positive outcomes
키워드  
Economic platforms
키워드  
Stackelberg problem
기타저자  
Harvard University Engineering and Applied Sciences - Computer Science
기본자료저록  
Dissertations Abstracts International. 85-12A.
전자적 위치 및 접속  
로그인 후 원문을 볼 수 있습니다.

MARC

 008250123s2024        us                              c    eng  d
■001000017161831
■00520250211151451
■006m          o    d                
■007cr#unu||||||||
■020    ▼a9798382776545
■035    ▼a(MiAaPQ)AAI31296734
■040    ▼aMiAaPQ▼cMiAaPQ
■0820  ▼a004
■1001  ▼aMibuari,  Eric  M.▼0(orcid)0009-0007-5347-900X
■24510▼aTowards  Positive  Outcomes  in  the  AI  Economy:  Mitigating  Algorithmic  Collusion  and  Enabling  Fair  Recourse
■260    ▼a[Sl]▼bHarvard  University▼c2024
■260  1▼aAnn  Arbor▼bProQuest  Dissertations  &  Theses▼c2024
■300    ▼a165  p
■500    ▼aSource:  Dissertations  Abstracts  International,  Volume:  85-12,  Section:  A.
■500    ▼aAdvisor:  Parkes,  David.
■5021  ▼aThesis  (Ph.D.)--Harvard  University,  2024.
■520    ▼aThe  rise  of  Artificial  Intelligence  (AI)  promises  to  solve  many  important  problems  in  the  world.  At  the  same  time,  awareness  has  been  increasing  about  its  potential  and  real  harms.  How  can  we  extract  maximum  benefit  from  the  promise  of  AI  while  minimizing  present  harms  and  mitigating  future  risks?In  this  thesis,  I  frame  and  answer  this  question  from  the  perspective  of  enabling  and  promoting  positive  outcomes  in  the  AI-enabled  economy,  where  markets  are  facilitated  using  AI  algorithms,  including  agent  behavior,  pricing,  and  matching  and  clearing.  The  goal  of  my  research  is  to  find  and  create  the  conditions  under  which  the  benefits  of  AI  are  preserved  or  even  enhanced  while  the  tendency  to  diminish  welfare,  perhaps  to  particular  groups,  is  contained  to  the  greatest  possible  extent.In  lending  domains,  machine  learning  can  be  used  to  form  a  predictive  model  of  the  probability  of  default  (a  "risk  score"),  this  driving  loan  decisions.  For  simple  models,  this  brings  the  benefits  of  transparency  and  explainability,  as  well  as  guidance  in  regard  to  recourse.  An  alternative  is  to  use  policy  learning,  that  is,  learning  a  policy  from  borrower  characteristics  to  loan  decisions  directly,  and  without  explicit  risk  scoring.  This  emphasizes  profit  and  can  speed  up  learning,  as  a  lender  understands  a  borrower  population,  but  with  a  concomitant  loss  of  transparency.  I  introduce  a  risk-score  based  policy  learning  method,  as  well  as  a  new  metric  of  recourse  effort  fairness,  and  demonstrate  that  this  risk-score  based  policy  learning  achieves  optimal  profits,  explainability  and  transparency,  as  well  as  recourse  effort  fairness.There  are  a  number  of  problems  where  economic  actors  follow  sequential  behaviors,  for  example,  in  making  pricing  adjustments  over  time  on  e-commerce  platforms,  or  power  trading  and  storage  optimization  through  a  typical  day  in  an  electricity  market.  In  this  thesis,  I  work  with  the  Stackelberg  POMDP  framework  (partially  observable  Markov  decision  process)  to  design  interventions  in  these  kinds  of  sequential  settings,  seeking  to  improve  economic  welfare.  I  am  especially  focused  on  concerns  that  can  arise  in  regard  to  tacit  collusion  between  AI-mediated  pricing  or  trading  and  storage  decisions  by  automated  agents.In  a  first  setting,  I  study  algorithmic  pricing  on  e-commerce  platforms,  where  reinforcement  learning  (RL)  algorithms  have  been  shown  to  learn  to  set  collusive  prices  with  nothing  more  than  profit  feedback.  This  raises  the  question  as  to  whether  collusive  pricing  can  be  prevented  through  the  design  of  suitable  ``buy  boxes,"  i.e.,  through  the  design  of  the  rules  that  govern  the  promotion  of  particular  products  and  prices  to  consumers.  I  show  that  RL  can  also  be  used  by  platforms  to  learn  buy  box  rules  that  are  effective  in  preventing  collusion  by  RL  sellers.  For  this,  I  adopt  the  Stackelberg  POMDP  framework,  and  demonstrate  success  in  learning  robust  rules  that  provide  high  consumer  welfare.In  a  second  setting,  I  study  trading  and  storage  decisions  by  battery  operators  in  electricity  power  markets.  %requiring  power  storage  are  an  example  of  a  multi-agent  dynamic  storage  optimization  application  in  which  artificial  intelligence  (AI)  algorithms  are  actively  applied  by  prosumers  (battery  operators  who  both  produce  and  consume  power).  In  this  application,  agents  who  correspond  to  battery  operators  buy  and  sell  power  in  a  market  with  producers  and  price-elastic  consumers,  and  where  power  can  be  stored  at  a  negligible  cost  up  to  an  agent-specific  capacity.  The  use  of  RL  algorithms  has  been  shown  in  this  setting  to  lead  to  outcomes  where  battery  operators  arbitrage  the  market  over  time,  and  that  correspond  to  tacit  collusion.  I  again  appeal  to  the  Stackelberg  POMDP  framework,  and  demonstrate  success  in  learning  collusion-mitigation  policies  by  a  regulator,  in  particular  through  the  use  of  a  network-flow  thresholding  intervention.
■590    ▼aSchool  code:  0084.
■650  4▼aComputer  science
■653    ▼aAlgorithmic  collusion
■653    ▼aFair  recourse
■653    ▼aPositive  outcomes
■653    ▼aEconomic  platforms
■653    ▼aStackelberg  problem
■690    ▼a0984
■690    ▼a0800
■690    ▼a0501
■71020▼aHarvard  University▼bEngineering  and  Applied  Sciences  -  Computer  Science.
■7730  ▼tDissertations  Abstracts  International▼g85-12A.
■790    ▼a0084
■791    ▼aPh.D.
■792    ▼a2024
■793    ▼aEnglish
■85640▼uhttp://www.riss.kr/pdu/ddodLink.do?id=T17161831▼nKERIS▼z이  자료의  원문은  한국교육학술정보원에서  제공합니다.

미리보기

내보내기

chatGPT토론

Ai 추천 관련 도서


    신착도서 더보기
    최근 3년간 통계입니다.

    소장정보

    • 예약
    • 소재불명신고
    • 나의폴더
    • 우선정리요청
    • 비도서대출신청
    • 야간 도서대출신청
    소장자료
    등록번호 청구기호 소장처 대출가능여부 대출정보
    TF11513 전자도서 대출가능 마이폴더 부재도서신고 비도서대출신청 야간 도서대출신청

    * 대출중인 자료에 한하여 예약이 가능합니다. 예약을 원하시면 예약버튼을 클릭하십시오.

    해당 도서를 다른 이용자가 함께 대출한 도서

    관련 인기도서

    로그인 후 이용 가능합니다.