Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Kate McCurdy

Differentiable Tree Operations Promote Compositional Generalization

Jun 01, 2023

Paul Soulos, Edward Hu, Kate McCurdy, Yunmo Chen, Roland Fernandez, Paul Smolensky, Jianfeng Gao

Abstract:In the context of structure-to-structure transformation tasks, learning sequences of discrete symbolic operations poses significant challenges due to their non-differentiability. To facilitate the learning of these symbolic sequences, we introduce a differentiable tree interpreter that compiles high-level symbolic tree operations into subsymbolic matrix operations on tensors. We present a novel Differentiable Tree Machine (DTM) architecture that integrates our interpreter with an external memory and an agent that learns to sequentially select tree operations to execute the target transformation in an end-to-end manner. With respect to out-of-distribution compositional generalization on synthetic semantic parsing and language generation tasks, DTM achieves 100% while existing baselines such as Transformer, Tree Transformer, LSTM, and Tree2Tree LSTM achieve less than 30%. DTM remains highly interpretable in addition to its perfect performance.

* ICML 2023. Code available at https://github.com/psoulos/dtm

Via

Access Paper or Ask Questions

Inflecting when there's no majority: Limitations of encoder-decoder neural networks as cognitive models for German plurals

May 18, 2020

Kate McCurdy, Sharon Goldwater, Adam Lopez

Figure 1 for Inflecting when there's no majority: Limitations of encoder-decoder neural networks as cognitive models for German plurals

Figure 2 for Inflecting when there's no majority: Limitations of encoder-decoder neural networks as cognitive models for German plurals

Figure 3 for Inflecting when there's no majority: Limitations of encoder-decoder neural networks as cognitive models for German plurals

Figure 4 for Inflecting when there's no majority: Limitations of encoder-decoder neural networks as cognitive models for German plurals

Abstract:Can artificial neural networks learn to represent inflectional morphology and generalize to new words as human speakers do? Kirov and Cotterell (2018) argue that the answer is yes: modern Encoder-Decoder (ED) architectures learn human-like behavior when inflecting English verbs, such as extending the regular past tense form -(e)d to novel words. However, their work does not address the criticism raised by Marcus et al. (1995): that neural models may learn to extend not the regular, but the most frequent class -- and thus fail on tasks like German number inflection, where infrequent suffixes like -s can still be productively generalized. To investigate this question, we first collect a new dataset from German speakers (production and ratings of plural forms for novel nouns) that is designed to avoid sources of information unavailable to the ED model. The speaker data show high variability, and two suffixes evince 'regular' behavior, appearing more often with phonologically atypical inputs. Encoder-decoder models do generalize the most frequently produced plural class, but do not show human-like variability or 'regular' extension of these other plural markers. We conclude that modern neural models may still struggle with minority-class generalization.

* To appear at ACL 2020

Via

Access Paper or Ask Questions