Результаты поиска

Материал из MachineLearning.

По запросу «Preferences»

Перейти к: навигация, поиск

Страницы с названием «Preferences» не существует.

Для получения более подробной информации о поиске на страницах проекта, см. справочный раздел.

Ниже показаны 9 результатов, начиная с № 1.


Просмотреть (предыдущие 20) (следующие 20) (20 | 50 | 100 | 250 | 500)

Нет совпадений в названиях статей

Совпадения в текстах статей

  1. Experimental Economics and Machine Learning (workshop) (4869 байт)
    7: ...h paying different amounts of money restricts the preferences of the subjects in experiments, the exclusive app...
  2. Численные методы обучения по прецедентам (практика, В.В. Стрижов)/Группа 074, весна 2014 (40 609 байт)
    400: ... million voting situations, consider all possible preferences deviations and calculate the coalitional manipula...
  3. Рекомендации по доработке магистерской диссертации (18 015 байт)
    66: ...ife started to beat faster, people switched their preferences from horizontal search to vertical."
  4. Обучение с подкреплением из обратной связи человека (RLHF) (21 403 байта)
    14: ...лавие=Deep Reinforcement Learning from Human Preferences |издание=Advances in Neural Information Pr...
    161: ...лавие=Deep Reinforcement Learning from Human Preferences |издание=Advances in Neural Information Pr...
  5. Супервыравнивание (25 861 байт)
    78: ... P. et al. Deep Reinforcement Learning from Human Preferences // arXiv:1706.03741, 2017.</ref>
    176: .... et al. ''Deep Reinforcement Learning from Human Preferences'' // arXiv:1706.03741, 2017.
  6. Цивилизационная идеология как основа целеполагания в развитии ИИ. (44 587 байт)
    92: ...references Deep Reinforcement Learning from Human Preferences] // Advances in Neural Information Processing Sys...
    189: ...references Deep Reinforcement Learning from Human Preferences] // Advances in Neural Information Processing Sys...
  7. Спецификация цели (31 477 байт)
    45: ...modei D.'' Deep Reinforcement Learning from Human Preferences // Advances in Neural Information Processing Syst...
    111: ...modei D.'' Deep Reinforcement Learning from Human Preferences // Advances in Neural Information Processing Syst...
  8. Взлом вознаграждения (31 689 байт)
    25: .... et al.'' Deep Reinforcement Learning from Human Preferences // NeurIPS 2017.</ref>. Этот случай —...
    90: ...modei D.'' Deep Reinforcement Learning from Human Preferences // Advances in Neural Information Processing Syst...
  9. Модель вознаграждения (17 576 байт)
    82: ...лавие=Deep Reinforcement Learning from Human Preferences |издание=Advances in Neural Information Pr...

Просмотреть (предыдущие 20) (следующие 20) (20 | 50 | 100 | 250 | 500)



Искать в пространствах имён:

Показывать перенаправления
Искать
Личные инструменты