Many modern Artificial Intelligence (AI) models can produce meaningful text, but they often fail on complex structural and numerical data involving different units, formulas, information describing the data (i.e., metadata), and hierarchies. These failures are especially concerning in areas such as medicine, finance, defense, and space, where even small quantitative mistakes can lead to misleading conclusions and significant negative consequences. This project aims to address this problem by developing a new AI model focused on accurate comprehension of complex numerical and structured data rather than natural language text. The project will help make scientific knowledge more transparent and accessible, while also supporting education through new teaching materials, student research opportunities, and outreach activities that engage learners in data reasoning. By improving the ability of AI to work correctly with complex numerical and structured data, the project advances the progress of science, supports health and welfare, and strengthens the nation’s capacity for trustworthy data-driven discovery and decision-making. The project develops the Large Number Model (LNM), a hybrid neural-symbolic model for reliable reasoning over numbers, units, formulas, and complex tabular data. The research includes three main activities: creating scalable methods to extract numerical and structured information from documents, designing model architectures that represent quantities and t