Skip to content
View TiezMind's full-sized avatar
๐ŸŽฏ
Focusing
๐ŸŽฏ
Focusing

Highlights

  • Pro

Block or report TiezMind

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please donโ€™t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this userโ€™s behavior. Learn more about reporting abuse.

Report abuse
TiezMind/README.md

Hi there, I'm Pengyu Li (ๆŽ้นๅฎ‡) ๐Ÿ‘‹

Typing SVG

Homepage Google Scholar Email GitHub


๐Ÿ™‹ About Me

I am a first-year M.S. student in Computer Science & Technology at Xi'an Jiaotong University. My research centers on large language models and multimodal / omni-modal foundation models, with a focus on the full post-training pipeline.

I have interned at ByteDance Seed, TikTok, and iFLYTEK, working on omni-modal (speech) foundation models, large-scale multimodal content understanding, and medical LLMs.

  • ๐Ÿ”ญ Currently working on the speech modality of an omni-modal foundation model @ ByteDance Seed
  • ๐ŸŒฑ Interested in unifying perception, reasoning, and generation across modalities โ€” efficiently
  • ๐Ÿ’ฌ Happy to chat about multimodal LLMs, omni-modal training, and RL post-training
  • ๐Ÿ“ซ Feel free to email me for any form of academic cooperation!

๐Ÿ”ฌ Research Interests

  • Multimodal & Omni-modal LLMs โ€” unifying text, vision, and speech into a single foundation model
  • LLM Continued Pre-training ยท Mid-training ยท Post-training โ€” data recipes, task composition & interference
  • Efficient & Unified Multimodal Reasoning โ€” token pruning, distillation, RL (GRPO) for reasoning

๐Ÿ“„ Selected Publications

โญ = first author ย ยทย  full list on Google Scholar

  • โญ AGTAO: Robust and Stabilized LLM Unlearning via Adversarial Gating Training with Adaptive Orthogonality
    ACL ย  Paper Code

  • โญ Visual-OPSD: Cross-Modal On-Policy Self-Distillation for Efficient Unified Multimodal Reasoning
    Review ย  Paper Code

  • โญ EMDFNet: Efficient Multi-scale and Diverse Feature Network for Traffic Sign Detection
    ICANN ย  Code

  • PhysPRM: A Generative Process Reward Model with Fine-grained Diagnosis for Physics Problem Solving
    ACL ย  Code

  • AERO: Autonomous Evolutionary Reasoning Optimization via Endogenous Dual-Loop Feedback
    COLM ย  Code

  • LogicGraph: Benchmarking Multi-Path Logical Reasoning via Neuro-Symbolic Generation and Verification
    EMNLP ย  Code

๐Ÿ’ผ Experience

Role Organization Period Focus
Intern ByteDance Seed 2026.02 โ€“ Present Omni-modal foundation models (speech)
Intern TikTok, ByteDance 2025.02 โ€“ 2025.09 Multimodal content understanding
Algorithm Intern iFLYTEK 2024.10 โ€“ 2025.01 Medical LLM (SFT + DPO) ยท national invention patent

๐Ÿ› ๏ธ Tech Stack


Profile Views
"Long-term correctness over short-term convenience."

Pinned Loading

  1. AGT-unlearning AGT-unlearning Public

    [ACL2026] AGTAO : Robust and Stabilized LLM Unlearning via Adversarial Gating Training with Adaptive Orthogonality

    Python 6 1

  2. EMDFNet EMDFNet Public

    [ICANN2024] EMDFNet: Efficient Multi-scale and Diverse Feature Network for Traffic Sign Detection

    Python 1

  3. Visual-OPSD Visual-OPSD Public

    [arxiv 2606] Visual-OPSD: Cross-Modal On-Policy Self-Distillation for Efficient Unified Multimodal Reasoning

    Python 57 1