02
Research Projects
科研项目
AP Humanities Automated Scoring System
In progress
AP 人文科目自动评分系统
Building an automated scoring system for AP humanities essays — history, literature, and foreign languages. The goal is a scoring pipeline that is accurate, transparent, and equitable: refining evaluation criteria, exploring resource-efficient modeling, and expanding applicability across prompts and contexts.
Role Research contributorStatus Ongoing · UMD
Automated essay scoringLLMsEducational measurement
Multimodal Analysis of Online Social Behavior
In progress
社交媒体多模态行为分析
At Princeton University: machine learning on multimodal data to study human behavior and communication on social media — contributing annotated training data and identifying behavioral, linguistic, and contextual features for computational analyses of online discourse and social interaction.
Role Research AssistantStatus Ongoing · Princeton
Multimodal MLHuman behaviorComputational social science
Genomic Sequence Modeling & Molecular Phenotype Prediction
In progress
基因组序列建模与分子表型预测
At Westlake University: developing and evaluating deep-learning pipelines for biological data — genomic sequence modeling, molecular phenotype prediction, model generalization, and rigorous benchmarking — with large-scale datasets and GPU-based computational research environments.
Role ResearcherStatus Ongoing · Westlake
Computational genomicsSequence modelingDeep learning
Cost-efficient Generative-AI Summarization for Essay Scoring
Preprint · 2026
面向大规模评分的高性价比生成式摘要
Showed that generative-AI summarization can make automated essay scoring scalable and cost-efficient in educational assessment, reducing inference cost while preserving scoring quality.
Role Sole authorStatus arXiv:2607.15829
SummarizationCost efficiencyAES
LLM Scoring Rationales vs. Human Raters
Preprints · 2025
LLM 评分理由与人工评分员对比
Compared scoring rationales between large language models and human raters, explored how summarization by generative models supports scoring of long essays, and examined the utility of LLM rationales for enhancing automated essay scoring.
Role First authorVenue arXiv 2025 (×3)
LLM rationalesHuman–AI comparison
Generative-AI Text Detection
Published · 2024
生成式 AI 文本检测
Studied the impacts of tokenization and dataset size on identifying AI-generated text (Frontiers in Artificial Intelligence), and comparatively evaluated machine-learning and LLM approaches for detecting ChatGPT-written essays under revision conditions (IMPS 2024 Proceedings).
Role First authorVenue Frontiers in AI · IMPS Proceedings
LLM detectionNLPAcademic integrity