Related Experiment Video
Updated: Jan 10, 2026

An Experimental Analysis of Children's Ability to Provide a False Report about a Crime
Published on: May 3, 2016
Machine learning detection of manipulative environmental disclosures in corporate reports
Yuanzhe Li1,2,3, Junyuan Li4,5, Yutong Zheng5
1Carbon Neutrality Institute, China University of Mining and Technology, Xuzhou, 221116, China. yuanzhe001@e.ntu.edu.sg.
None:
Detecting manipulative environmental disclosures remains a critical yet unresolved challenge for regulators and investors. This study proposes a machine learning framework that integrates financial indicators, textual sentiment, and public attention data to identify potential manipulation among Chinese listed firms. A Random Forest model is trained using multi-source features derived from corporate reports and Baidu Index trends. The optimized model demonstrates strong discriminatory ability under severe class imbalance (ROC-AUC = 0.94, PR-AUC = 0.78, Balanced Accuracy = 0.86, MCC = 0.72), indicating robust and reliable performance across both majority and minority classes. Evaluation through balanced metrics further confirms the model's genuine predictive capacity rather than overfitting to training data. SHAP-based interpretation reveals that financial pressure, abnormal public attention, and sentiment deviation are the primary determinants of manipulation risk. Overall, the framework highlights how interpretable machine learning can strengthen data-driven environmental supervision. The findings are context-specific to the Chinese market due to reliance on Baidu-based indicators, warranting validation in other regulatory contexts in future research.
Related Concept Videos
Understanding Deception
Types of Reports I: Hands-off Report
Following are the key components and categories of hand-off reports:
Purpose and Process:
Stereotype Content Model
Detection of Gross Error: The Q Test
Self-Presentation: Self-Monitoring and Self-Handicapping
Quantifying and Rejecting Outliers: The Grubbs Test
