Select a category to explore research frontiers
Loading categories...
Investigation of learning reward functions from human preference comparisons and ranking feedback, enabling alignment of agent behavior with human values without explicit reward specification.