Результаты поиска
Материал из MachineLearning.
По запросу «Preferences»
Страницы с названием «Preferences» не существует.
Для получения более подробной информации о поиске на страницах проекта, см. справочный раздел.
Ниже показаны 9 результатов, начиная с № 1.
Просмотреть (предыдущие 20) (следующие 20) (20 | 50 | 100 | 250 | 500)
Нет совпадений в названиях статей
Совпадения в текстах статей
- Experimental Economics and Machine Learning (workshop) (4869 байт)
7: ...h paying different amounts of money restricts the preferences of the subjects in experiments, the exclusive app... - Численные методы обучения по прецедентам (практика, В.В. Стрижов)/Группа 074, весна 2014 (40 609 байт)
400: ... million voting situations, consider all possible preferences deviations and calculate the coalitional manipula... - Рекомендации по доработке магистерской диссертации (18 015 байт)
66: ...ife started to beat faster, people switched their preferences from horizontal search to vertical." - Обучение с подкреплением из обратной связи человека (RLHF) (21 403 байта)
14: ...лавие=Deep Reinforcement Learning from Human Preferences |издание=Advances in Neural Information Pr...
161: ...лавие=Deep Reinforcement Learning from Human Preferences |издание=Advances in Neural Information Pr... - Супервыравнивание (25 861 байт)
78: ... P. et al. Deep Reinforcement Learning from Human Preferences // arXiv:1706.03741, 2017.</ref>
176: .... et al. ''Deep Reinforcement Learning from Human Preferences'' // arXiv:1706.03741, 2017. - Цивилизационная идеология как основа целеполагания в развитии ИИ. (44 587 байт)
92: ...references Deep Reinforcement Learning from Human Preferences] // Advances in Neural Information Processing Sys...
189: ...references Deep Reinforcement Learning from Human Preferences] // Advances in Neural Information Processing Sys... - Спецификация цели (31 477 байт)
45: ...modei D.'' Deep Reinforcement Learning from Human Preferences // Advances in Neural Information Processing Syst...
111: ...modei D.'' Deep Reinforcement Learning from Human Preferences // Advances in Neural Information Processing Syst... - Взлом вознаграждения (31 689 байт)
25: .... et al.'' Deep Reinforcement Learning from Human Preferences // NeurIPS 2017.</ref>. Этот случай —...
90: ...modei D.'' Deep Reinforcement Learning from Human Preferences // Advances in Neural Information Processing Syst... - Модель вознаграждения (17 576 байт)
82: ...лавие=Deep Reinforcement Learning from Human Preferences |издание=Advances in Neural Information Pr...
Просмотреть (предыдущие 20) (следующие 20) (20 | 50 | 100 | 250 | 500)

