Skip to content
View wang2226's full-sized avatar

Highlights

  • Pro

Block or report wang2226

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Content in all repositories owned by your account will be closed.
Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
wang2226/README.md

Hi there 👋, I'm Haoran

I'm an AI researcher working on Trustworthy AI, Large Language Models, and AI Agents.

Research Interests: LLM Safety · Privacy · Representation Steering · Decoding · RAG · AI Agents

💻 Tech Stack

🤖 AI & Machine Learning

Python PyTorch scikit-learn Pandas NumPy

🛠️ Tools

Git Docker VS Code Jupyter

📈 Contribution Timeline

Pinned Loading

  1. CI-Steering CI-Steering Public

    [COLM 2026] Do LLMs Know What Is Private Internally? Probing and Steering Contextual Privacy Norms in Large Language Model Representations

    Python 7 1

  2. PAD PAD Public

    [KDD 2026] Privacy-Aware Decoding: Mitigating Privacy Leakage of Large Language Models in Retrieval-Augmented Generation

    Python 8 3

  3. Awesome-LLM-Decoding Awesome-LLM-Decoding Public

    📜 Paper list on decoding methods for LLMs and LVLMs

    76 4

  4. mmcv-dataset/MMCV mmcv-dataset/MMCV Public

    [COLING 2025] Piecing It All Together: Verifying Multi-Hop Multimodal Claims.

    Python 10

  5. Trojan-Activation-Attack Trojan-Activation-Attack Public

    [CIKM 2024] Trojan Activation Attack: Attack Large Language Models using Activation Steering for Safety-Alignment.

    Python 30 3

  6. FOLK FOLK Public

    [EMNLP 2023] Explainable Claim Verification via Knowledge-Grounded Reasoning with Large Language Models

    Python 28 4