# Ying Li
[[Peking University]] の研究者(
[email protected])。[[A Survey of AIOps in the Era of Large Language Models]] の責任著者の一人([[Tong Jia]] と並ぶ)。
- [[Lingzhe Zhang]]・[[Tong Jia]] とともに PKU の AIOps / LLM for SRE 研究クラスタを形成する。本 wiki ではこのグループの成果として LLM4AIOps サーベイ([[A Survey of AIOps in the Era of Large Language Models]])と [[MicroRemed]]([[@2025__arXiv__MicroRemed - Benchmarking LLMs in Microservices Remediation]])を取り込んでいる。
- [[@2026__arXiv__Towards Robust LLM Post-Training - Automatic Failure Management for Reinforcement Fine-Tuning]] の corresponding author の一人([[Tong Jia]] と並ぶ)。
- [[@2024__ESEM__Reducing Events to Augment Log-based Anomaly Detection Models - An Empirical Study]] の corresponding author。ログベース異常検知のイベント削減手法 [[LogCleaner]] を提案(ESEM 2024)。
- [[@2024__KDD__Multivariate Log-based Anomaly Detection for Distributed Database]](MultiLog, KDD 2024)および [[@2025__arXiv__LogDB - Multivariate Log-based Failure Diagnosis for Distributed Databases]](LogDB, J. ACM 2025)の corresponding author([[Tong Jia]] と並ぶ)。分散 DB 向けログ異常検知・障害診断の体系化を主導。
- [[@2025__IEEE TSC__Towards Close-To-Zero Runtime Collection Overhead - Raft-Based Anomaly Diagnosis on System Faults for Distributed Storage System]](RBAD, IEEE TSC 2024) の corresponding author([[Tong Jia]] と並ぶ)。Raft ログを異常診断に活用する初の研究で、MultiLog・LogDB のログ解析から「コンセンサスログ」という新たなデータ源開拓へ研究を拡張した。IBM での分散コンピューティング・IBM Master Inventor の経歴を持ち、30 件以上の米国/中国特許を保有する。60 件以上の国際論文発表実績。
- [[@2026__TDSC__Towards In-Depth Root Cause Localization for Microservices with Multi-Agent Recursion-of-Thought]](RCLAgent, IEEE TDSC 採録)の corresponding author の一人([[Tong Jia]] と並ぶ)。マイクロサービス根本原因特定にマルチエージェント Recursion-of-Thought を適用し、[[Huawei Theory Lab]] との共同研究として実施。
- [[@2026__arXiv__Bifrost - Empowering Pretrained Language Model with Fallibility Representation for Log-Based Fault Diagnosis]](Bifrost, ASE '26)の corresponding author の一人([[Tong Jia]] と並ぶ)。ログの多階層構造(実行フロー・イベント・コンポーネント)を fallibility representation として定式化し、自己教師あり対照学習で PLM に学習させる手法を提案。
- [[@2026__ICSE-SEIP__MagmaScope Identifying Root-Cause Changes for Emergency Incident in Large-Scale Cloud Infrastructure]](MagmaScope, ICSE-SEIP '26)の責任著者(corresponding author)。[[ByteDance]] との共同研究として [[MagmaScope]] を開発。IM グループチャットのマルチモーダルデータを活用した根本原因変更特定ハイブリッドシステムであり、PKU-ByteDance 協力プロジェクト・CPSF 特別研究員プログラムの支援を受ける。
## 関連
- 本ソース: [[A Survey of AIOps in the Era of Large Language Models]] / [[@2026__arXiv__Towards Robust LLM Post-Training - Automatic Failure Management for Reinforcement Fine-Tuning]] / [[@2024__ESEM__Reducing Events to Augment Log-based Anomaly Detection Models - An Empirical Study]] / [[@2024__KDD__Multivariate Log-based Anomaly Detection for Distributed Database]] / [[@2025__arXiv__LogDB - Multivariate Log-based Failure Diagnosis for Distributed Databases]] / [[@2025__IEEE TSC__Towards Close-To-Zero Runtime Collection Overhead - Raft-Based Anomaly Diagnosis on System Faults for Distributed Storage System]] / [[@2026__TDSC__Towards In-Depth Root Cause Localization for Microservices with Multi-Agent Recursion-of-Thought]] / [[@2026__arXiv__Bifrost - Empowering Pretrained Language Model with Fallibility Representation for Log-Based Fault Diagnosis]] / [[@2026__ICSE-SEIP__MagmaScope Identifying Root-Cause Changes for Emergency Incident in Large-Scale Cloud Infrastructure]]
- 所属: [[Peking University]]
- 共著者: [[Lingzhe Zhang]] / [[Tong Jia]] / [[Kangjin Wang]] / [[Minghua He]] / [[Chiming Duan]]
- 関連概念: [[AIOps]] / [[LLMによる根本原因分析]] / [[マルチエージェント協調]] / [[ログベース障害診断]]
- 関連 MOC: [[LLM4SRE - MOC]] / [[AIOps - Failure Detection - MOC]]