Publications
View By Topic:
All Topics
CI Causal Induction CD Cognitive Development CEIL Cultural Evolution and Iterated Learning DMRL Decision Making and Reinforcement Learning E Education F Foundations IB Inductive Biases NBM Nonparametric Bayesian Models P Perception PR Probabilistic Reasoning RPM Rational Process Models S&C Similarity and Categorization SC Social Cognition SML Statistical Models of Language
(Click on an author's name to view all papers by that author.)
Filter publications
By Fisac, J.SML Liang, K. , Hu, H. , Liu, R. , Griffiths, T. L. , & Fisac, J. F. (2026). RLHS: Mitigating misalignment in RLHF with hindsight simulation. Findings of the Association for Computational Linguistics: ACL 2026 , 11457-11483. (pdf)
SML Liang, K. , Hu, H. , Zhao, X. , Song, D. , Griffiths, T. L. , & Fisac, J. F. (2025) Machine bullshit: Characterizing the emergent disregard for truth in large language models. (preprint)
DMRL PR Fisac, J. F. , Gates, M. A. , Hamrick, J. B. , Liu, C. , Hadfield-Menell, D. , Palaniappan, M. , Malik, D. , Sastry, S. S. , Griffiths, T. L. , & Dragan, A. D. (2017). Pragmatic-Pedagogic Value Alignment. International Symposium on Robotics Research . (pdf)
PR Fisac, J. F. , Liu, C. , Hamrick, J. B. , Sastry, S. , Hedrick, J. K. , Griffiths, T. L. , & Dragan, A. D. (2016). Generating plans that predict themselves. In Proceedings of the 12th International Workshop on the Algorithmic Foundations of Robotics (WAFR 2016) . (pdf)
PR SC Liu, C. , Hamrick, J. B. , Fisac, J. F. , Dragan, A. D , Hendrick, J. K. , Sastry, S. S , & Griffiths, T. L. (2016). Goal inference improves objective and perceived performance in human-robot collaboration. Proceedings of the 2016 International Conference on Autonomous Agents & Multiagent Systems . (pdf)