Abstract: Methods and apparatus for automated processing of natural language text is described. Received text can be preprocessed to produce language-space data that includes descriptive data elements for words. Source code that includes linguistic constraints, and that may be written in a programming language that is user-friendly to linguists, can be compiled to produce finite-state transducers and bi-machine transducers that are used by a language-processing virtual machine to process the language-space data. The language-processing virtual machine selects and executes code segments in accordance with path transitions of the transducers when applied on automatons to disambiguate meanings of words in the received text.