{"database":"biostudies-literature","file_versions":[],"scores":null,"additional":{"omics_type":["Unknown"],"volume":["11"],"submitter":["Feng Y"],"pubmed_abstract":["To address the growing demand for efficient public opinion analysis in water conservancy and related domains, as well as the inefficiencies and limited scalability of existing automated web data extraction algorithms for multi-source datasets, this research integrates advanced technologies including big data analytics, natural language processing, and deep learning. A novel, transferable web information extraction model based on deep learning (WIEM-DL) is proposed, leveraging knowledge graphs, machine learning, and ontology-based methods. This model is designed to adapt to varying website structures, enabling effective cross-website information extraction. By refining water conservancy-related online public opinion content and extracting key feature information from critical sentences, the"],"journal":["PeerJ. Computer science"],"pagination":["e2960"],"full_dataset_link":["https://www.ebi.ac.uk/biostudies/studies/S-EPMC12453640"],"repository":["biostudies-literature"],"pubmed_title":["A hybrid extraction model for semantic knowledge discovery of water conservancy big data."],"pmcid":["PMC12453640"],"pubmed_authors":["Zhang F","Wang P","Zhang Y","Feng Y","Dong J"],"additional_accession":[]},"is_claimable":false,"name":"A hybrid extraction model for semantic knowledge discovery of water conservancy big data.","description":"To address the growing demand for efficient public opinion analysis in water conservancy and related domains, as well as the inefficiencies and limited scalability of existing automated web data extraction algorithms for multi-source datasets, this research integrates advanced technologies including big data analytics, natural language processing, and deep learning. A novel, transferable web information extraction model based on deep learning (WIEM-DL) is proposed, leveraging knowledge graphs, machine learning, and ontology-based methods. This model is designed to adapt to varying website structures, enabling effective cross-website information extraction. By refining water conservancy-related online public opinion content and extracting key feature information from critical sentences, the","dates":{"release":"2025-01-01T00:00:00Z","publication":"2025","modification":"2026-06-03T19:19:06.378Z","creation":"2026-05-30T03:06:47.793Z"},"accession":"S-EPMC12453640","cross_references":{"pubmed":["40989349"],"doi":["10.7717/peerj-cs.2960"]}}