🎉Our paper is out in JAIR!
We tackle a key challenge in multi-objective reinforcement learning: how do you learn fair policies at scale when you have many conflicting objectives without requiring preference info upfront?
www.jair.org/index.php/ja...
Καλά κάνετε, γιατί τα διπλοτσέκ έχουν οδηγήσει σε υποσημειώσεις (όπως αυτή που διευκρινίζει ότι τα δεδομένα είναι εκτιμήσεις).