favs.blue
  1. 1

    🎉Our paper is out in JAIR!

    We tackle a key challenge in multi-objective reinforcement learning: how do you learn fair policies at scale when you have many conflicting objectives without requiring preference info upfront?

    www.jair.org/index.php/ja...

  2. 2

    Reply

    Καλά κάνετε, γιατί τα διπλοτσέκ έχουν οδηγήσει σε υποσημειώσεις (όπως αυτή που διευκρινίζει ότι τα δεδομένα είναι εκτιμήσεις).