Publications
- ⭐: Co-first Author - 🚩: Corresponding Author - 💭: Under Review
Conference Papers (22)
Published in EMNLP, 2026-10
RLVR, Entropy Collapse, Policy Reheating, On-Policy Self-Distillation, Temperature Scaling, Continued Reinforcement Learning
Download Paper
Published in EMNLP, 2026-10
Multimodal Large Language Models, Long-Context Understanding, Evidence Chain, Scientific Literature, Benchmark
Download Paper
Published in EMNLP, 2026-10
Controllable Captioning, Omni-modal Intelligence, Plug-and-play Framework, Multimodal Large Language Models
Download Paper
Published in ACL, 2026-07
Probability-Guided Token Selection, High-Value Semantic Signals, Gradient Interference, Training Stability & Rapid Convergence
Download Paper
Published in ACL, 2026-07
AI-Generated Code Security, Repository-Level Benchmark, Large Language Models (LLMs), Common Vulnerabilities and Exposures (CVEs)
Download Paper
Published in ICML, 2026-06
GRPO, Policy-Level Diversity, Small-to-Large Policy Optimization, RLVR, Rollout Diversity, Progressive Annealing
Download Paper
Published in ICLR, 2026-04
Long-form music (full-song) generation, Audio Language Model, Music In-Context Learning, Controllability, human preference evaluation
Download Paper
Published in AAAI, 2026-01
Video Editing, Diffusion Models, Unified Control Signal, Object Distortion Control
Download Paper
Published in ACL, 2025-07
Multi-Paradigm, Large Language Model (LLM), Progressive Paradigm Training (PPT), Zero-shot Generalization
Download Paper
Published in ICLR, 2025-04
Large Multimodal Models (LMMs), Chart Understanding, Code Generation, Cross-modal Reasoning
Download Paper
Published in EMNLP, 2024-11
Screenwriting, Large Language Models (LLMs), Role Playing, Creative Generation, Multi-Agent Collaboration
Download Paper
Published in EMNLP, 2024-11
ToolBeHonest, Hallucination, LLM, Multi‑level Diagnostic
Download Paper
Published in SIGIR-AP, 2024-10
Massive Tool Retrieval (MTR), Query‑Tool Alignment (QTA), Massive Tool Retrieval Benchmark, DPO
Download Paper
Published in COLM, 2024-02
Structured Knowledge Grounding (SKG), Instruction tuning, Generalist model, Tables / Graphs / Databases
Download Paper
Published in SIGIR-AP, 2023-11
Multidimensional Ethics, Ethical Judgment, Large Multimodal Models (LMMs), Binary & Multi‑label Classification
Download Paper
Published in ACL, 2023-07
Information Extraction , Unified Across IE Tasks, Triaffine Attention, Span‑extractive Framework, Low‑resource Transferability
Download Paper
Published in ACL, 2023-07
System 1 & System 2, Cooperative Reasoning (CoRe),Stepwise Feedback, Math Word Problems
Download Paper
Published in CVPR, 2023-06
Multimodality, Uncertainty Modeling, Vision-Language Pre-training, Probability Distribution Encoder (PDE)
Download Paper
Published in EMNLP, 2022-10
Zero‑Shot Learning, Multiple Choice Format, Pre‑trained Masked Language Model (PMLM), Unified Multiple Choice model
Download Paper
Published in ACM MM, 2022-10
Multimedia Recommendation, Graph Fusion, Edge-wise Modulation, Graph Convolutional Network (GCN)
Download Paper
Published in EMNLP, 2021-10
Multimodal Interaction, Trilinear Transformer, Visual Question Answering (VQA), Two-Stage Workflow
Download Paper
Published in NTCIR, 2020-12
Dialogue Quality, BiLSTM + Attention), CNN (Convolutional Neural Network), Pre‑trained Language Model, MoE
Download Paper
Journal Articles (2)
Published in TMLR, 2026-03
Autonomous Driving, Reasoning, LLM/MLLM, Cognitive Hierarchy, Long-tail Scenarios, Social Game, Survey
Download Paper
Published in ACM TOIS, 2024-05
Named Entity Recognition (NER), Machine Reading Comprehension (MRC), Single-Stream Reasoner (SSR), Multi-choice Input Format
Download Paper
Arxiv Papers (15)
Published in Arxiv, 2026-08
Multimodal Geometry Reasoning, Credit Assignment, Code-CoT, CE-GRPO, Executable Perception Code, Event-Level Reinforcement Learning
Download Paper
Published in Arxiv, 2026-08
Autonomous Research, Multi-Agent Systems, Idea Generation, Evidence-Grounded Execution, Independent Verification, Scientific Discovery
Download Paper
Published in Arxiv, 2026-08
On-Policy Distillation, Warm-up, Chain-of-Thought, Teacher Compatibility, LoRA, Reasoning Models
Download Paper
Published in Arxiv, 2026-07
Vision-Language-Action, Mixture-of-Experts, Kinematics-Supervised Routing, Robot Manipulation, DIYRobot
Download Paper
Published in Arxiv, 2026-06
MLLM-Generated Web Artifacts, Requirement-Induced State Evaluation, Interaction Contract Graph, Web Generation Benchmark, DOM/Visual Assertion
Download Paper
Published in Arxiv, 2026-05
3D Asset Editing, Rectified Flow, Velocity-Space Editing, Training-Free Editing, Mask-Free Editing, TRELLIS
Download Paper
Published in Arxiv, 2026-05
Instruction Following, Rubric-Guided Reasoning, Rubric-as-Reward, Self-Generated Rubrics, Reinforcement Learning, Self-Consistency
Download Paper
Published in Arxiv, 2026-02
Latent Reasoning, Chain-of-Thought Compression, Visual Supervision, DeepSeek-OCR, Minimum Description Length, Reasoning Efficiency
Download Paper
Published in Arxiv, 2025-12
Object-Goal Navigation, Chain-of-Thought, Dual-Relation Reasoning, Similarity-Aware Memory
Download Paper
Published in Arxiv, 2025-08
Long-Chain Complexity, GUI Agents, Subtask-Level Verifiability, POMDP
Download Paper
Published in Arxiv, 2024-06
Knowledge‑intensive, Paired and Interleaved Documents, Multimodal Datasets, Data Format
Download Paper
Published in Arxiv, 2024-01
Multimodal Understanding, Subject Knowledge Reasoning , University Exam Questions, LMM Performance Evaluation
Download Paper
Published in Arxiv, 2023-10
Game Development, Multi-agent Collaboration, LLM, Redundancy
Download Paper
Published in Arxiv, 2022-09
Chinese Pre-trained Models, Foundation Models, Open-Source Ecosystem, Cognitive Intelligence
Download Paper
Published in Arxiv, 2022-08
Semantic Matching, Pre-trained Language Model (PLM), Propensity‑Corrected Loss (PCL), LUE Semantic Matching Challenge
Download Paper