Publication
Learning a natural-language to LTL executable semantic parser for grounded robotics. (2020). doi:https://doi.org/10.48550/arXiv.2008.03277
CBMM-Memo-122.pdf (1.03 MB)
Learning Language from Vision. Workshop on Visually Grounded Interaction and Language (ViGIL) at the Thirty-third Annual Conference on Neural Information Processing Systems (NeurIPS) (2019).
Learning to Answer Questions from Wikipedia Infoboxes. The 2016 Conference on Empirical Methods on Natural Language Processing (EMNLP 2016) (2016).
Morales-EMNLP2016.pdf (197.28 KB)
A look back at the June 2016 BMM Workshop in Sestri Levante, Italy. (2016).
Sestri Levante Review (359.33 KB)
Measuring Social Biases in Grounded Vision and Language Embeddings. NAACL (Annual Conference of the North American Chapter of the Association for Computational Linguistics) (2021).
Measuring Social Biases in Grounded Vision and Language Embeddings. (2021).
CBMM-Memo-126.pdf (1.32 MB)
Modeling Visual Impairments with Artificial Neural Networks: a Review. International Conference on Computer Vision 2023 (2023). at <https://openaccess.thecvf.com/content/ICCV2023W/ACVR/html/Schiatti_Modeling_Visual_Impairments_with_Artificial_Neural_Networks_a_Review_ICCVW_2023_paper.html>
Multi-resolution modeling of a discrete stochastic process identifies causes of cancer. International Conference on Learning Representations (2021). at <https://openreview.net/forum?id=KtH8W3S_RE>
The Wiley Handbook of Human Computer Interaction 2, 539-559 (John Wiley & Sons , 2018).
ObjectNet: A large-scale bias-controlled dataset for pushing the limits of object recognition models. Neural Information Processing Systems (NeurIPS 2019) (2019).
9142-objectnet-a-large-scale-bias-controlled-dataset-for-pushing-the-limits-of-object-recognition-models.pdf (16.31 MB)
Partially Occluded Hands: A challenging new dataset for single-image hand pose estimation. (2018).
CBMM-Memo-097.pdf (8.53 MB)
Partially Occluded Hands: A challenging new dataset for single-image hand pose estimation. The 14th Asian Conference on Computer Vision (ACCV 2018) (2018). at <http://accv2018.net/>
partially-occluded-hands-6.pdf (8.29 MB)
PHASE: PHysically-grounded Abstract Social Events for Machine Social Perception. (2021).
CBMM-Memo-123.pdf (3.08 MB)
PHASE: PHysically-grounded Abstract Social Eventsfor Machine Social Perception. Shared Visual Representations in Human and Machine Intelligence (SVRHM) workshop at NeurIPS 2020 (2020). at <https://openreview.net/forum?id=_bokm801zhx>
phase_physically_grounded_abstract_social_events_for_machine_social_perception.pdf (2.49 MB)
Predicting Native Language from Gaze. Annual Meeting of the Association for Computational Linguistics (ACL 2017) (2017).
Social Interactions as Recursive MDPs. (2021).
CBMM-Memo-130.pdf (1.52 MB)
Spoken ObjectNet: A Bias-Controlled Spoken Caption Dataset. (2021).
CBMM-Memo-128.pdf (2.91 MB)
Spoken ObjectNet: A Bias-Controlled Spoken Caption Dataset. Interspeech 2021 (2021). doi:10.21437/Interspeech.2021
Spontaneous sign emergence in humans and machines through an embodied communication game. JCoLE Workshop (2022).
Temporal Grounding Graphs for Language Understanding with Accrued Visual-Linguistic Context. Proceedings of the Twenty-Sixth International Joint Conference on Artificial Intelligence (IJCAI 2017) (2017). at <c>
Towards a Programmer's Apprentice (Again). (2015).
CBMM-memo-030.pdf (294.27 KB)
]