MUMBAI, India, July 7 -- Intellectual Property India has published a patent application (202544130544 A) filed by Highradius Corporation on December 23, 2025, for Machine-Learning-Based System And Method For Automatically Extracting Fields From Documents.

Inventors include Debanwesh Bose; Pratyush Amrit; Sayantan Mondal; Sujatro Baral; Ansumun Gouda; Lohit Vankina; and Vamsi Ramkiran Bharadwaj Gandepally.

The application for the patent was published on July 03, 2026, under issue no. 27/2026.

Abstract: ABSTRACT MACHINE-LEARNING-BASED SYSTEM AND METHOD FOR AUTOMATICALLY EXTRACTING FIELDS FROM DOCUMENTS A machine-learning based (ML-based) system and method for automatically extracting one or more data fields from one or more documents, are disclosed. The ML-based system includes a document obtaining subsystem to obtain documents, a document pre-processing subsystem to generate pre-processed data, a field identifying subsystem to identify data fields using a trained ML model, and a field extracting subsystem to extract financial information. The ML-based system also comprises an output subsystem to deliver the extracted data to end users via user interfaces. The ML model is trained using historical documents, labelled data fields, and features such as distance-based features, direction-based features, dimension-based features, positional features, and value-based features. The M-based system employs hyperparameter optimization, noise removal, and accuracy assessment mechanisms to enhance performance. This ML-based system provides a scalable, accurate, and automated solution for financial information extraction, ensuring efficiency, adaptability, and seamless integration with enterprise systems. FIG. 2

Disclaimer: Curated by HT Syndication.