Publication

Found 56 results
Author Title [ Type(Desc)] Year
Filters: Author is Katz, Boris  [Clear All Filters]
Conference Paper
Kuo, Y. - L., Katz, B. & Barbu, A. Encoding formulas as deep networks: Reinforcement learning for zero-shot execution of LTL formulas. 2020 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS) (2020). doi:10.1109/IROS45743.2020.9341325
Ross, C., Barbu, A., Berzak, Y., Myanganbayar, B. & Katz, B. Grounding language acquisition by training semantic parsersusing captioned videos. Proceedings of the 2018 Conference on Empirical Methods in Natural Language Processing (EMNLP 2018), (2018). at <http://aclweb.org/anthology/D18-1285>PDF icon Ross-et-al_ACL2018_Grounding language acquisition by training semantic parsing using caption videos.pdf (3.5 MB)
Tejwani, R. et al. Incorporating Rich Social Interactions Into MDPs. 2022 IEEE International Conference on Robotics and Automation (ICRA)2022 International Conference on Robotics and Automation (ICRA) (2022). doi:10.1109/ICRA46639.2022.9811991
Ross, C., Berzak, Y., Katz, B. & Barbu, A. Learning Language from Vision. Workshop on Visually Grounded Interaction and Language (ViGIL) at the Thirty-third Annual Conference on Neural Information Processing Systems (NeurIPS) (2019).
Morales, A., Premtoon, V., Avery, C., Felshin, S. & Katz, B. Learning to Answer Questions from Wikipedia Infoboxes. The 2016 Conference on Empirical Methods on Natural Language Processing (EMNLP 2016) (2016).PDF icon Morales-EMNLP2016.pdf (197.28 KB)
Schiatti, L. et al. Modeling Visual Impairments with Artificial Neural Networks: a Review. International Conference on Computer Vision 2023 (2023). at <https://openaccess.thecvf.com/content/ICCV2023W/ACVR/html/Schiatti_Modeling_Visual_Impairments_with_Artificial_Neural_Networks_a_Review_ICCVW_2023_paper.html>
Yaari, A. Uri et al. Multi-resolution modeling of a discrete stochastic process identifies causes of cancer. International Conference on Learning Representations (2021). at <https://openreview.net/forum?id=KtH8W3S_RE>
Myanganbayar, B. et al. Partially Occluded Hands: A challenging new dataset for single-image hand pose estimation. The 14th Asian Conference on Computer Vision (ACCV 2018) (2018). at <http://accv2018.net/>PDF icon partially-occluded-hands-6.pdf (8.29 MB)
Netanyahu, A., Shu, T., Katz, B., Barbu, A. & Tenenbaum, J. B. PHASE: PHysically-grounded Abstract Social Events for Machine Social Perception. AAAI-21 (2021).
Netanyahu, A., Shu, T., Katz, B., Barbu, A. & Tenenbaum, J. B. PHASE: PHysically-grounded Abstract Social Eventsfor Machine Social Perception. Shared Visual Representations in Human and Machine Intelligence (SVRHM) workshop at NeurIPS 2020 (2020). at <https://openreview.net/forum?id=_bokm801zhx>PDF icon phase_physically_grounded_abstract_social_events_for_machine_social_perception.pdf (2.49 MB)
Berzak, Y., Nakamura, C., Flynn, S. & Katz, B. Predicting Native Language from Gaze. Annual Meeting of the Association for Computational Linguistics (ACL 2017) (2017).
Palmer, I., Rouditchenko, A., Barbu, A., Katz, B. & Glass, J. Spoken ObjectNet: A Bias-Controlled Spoken Caption Dataset. Interspeech 2021 (2021). doi:10.21437/Interspeech.2021
Cheng, E., Kuo, Y. - L., Cases, I., Katz, B. & Barbu, A. Spontaneous sign emergence in humans and machines through an embodied communication game. JCoLE Workshop (2022).
Paul, R., Barbu, A., Felshin, S., Katz, B. & Roy, N. Temporal Grounding Graphs for Language Understanding with Accrued Visual-Linguistic Context. Proceedings of the Twenty-Sixth International Joint Conference on Artificial Intelligence (IJCAI 2017) (2017). at <c>
Kuo, Y. - L. et al. Trajectory Prediction with Linguistic Representations. 2022 IEEE International Conference on Robotics and Automation (ICRA) (2022). doi:10.1109/ICRA46639.2022.9811928
Subramaniam, V. et al. Using Multimodal DNNs to Study Vision-Language Integration in the Brain. ICLR 2023 (2023). at <https://openreview.net/pdf?id=OQQ1p0pFP4>
Tejwani, R. et al. Zero-shot linear combinations of grounded social interactions with Linear Social MDPs. Proceedings of the 37th AAAI Conference on Artificial Intelligence (AAAI) (2023).
Conference Post Doc/Student Spotlight Talk
Kuo, Y. - L., Katz, B. & Barbu, A. Deep Compositional Robotic Planners that Follow Natural Language Commands. Workshop on Visually Grounded Interaction and Language (ViGIL) at the Thirty-third Annual Conference on Neural Information Processing Systems (NeurIPS), (2019). at <https://vigilworkshop.github.io/>
Conference Proceedings
Berzak, Y., Katz, B. & Levy, R. Assessing Language Proficiency from Eye Movements in Reading. 16th Annual Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies (2018). at <http://naacl2018.org/>PDF icon 1804.07329.pdf (350.43 KB)
Barbu, A. et al. ObjectNet: A large-scale bias-controlled dataset for pushing the limits of object recognition models. Neural Information Processing Systems (NeurIPS 2019) (2019).PDF icon 9142-objectnet-a-large-scale-bias-controlled-dataset-for-pushing-the-limits-of-object-recognition-models.pdf (16.31 MB)

Pages