Skip to content
View xpxxx's full-sized avatar
🤯
I may be slow to respond.
🤯
I may be slow to respond.

Block or report xpxxx

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Content in all repositories owned by your account will be closed.
Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
xpxxx/README.md

Hi there, I'm Peixuan 👋

I hold an MSc in Computer Science from Nanjing University, and my research focuses on the intersection of Software Engineering (SE) and Artificial Intelligence.

🎓 Status: I have completed my Master's degree and am actively seeking PhD opportunities in AI evaluation, model auditing, and trustworthy systems.


🔬 Research Interests & Focus

My work focuses on empirical evaluation, benchmark validity, and auditing automated systems:

  • Model Auditing & Bias Probing: Developing black-box testing methodologies (e.g., Combinatorial Interaction Testing) to stress-test failure boundaries, behavioral inconsistencies, and systemic bias in foundation models.
  • Empirical Benchmarks: Designing robust, counterfactual, and task-driven benchmark suites to evaluate LLMs and agentic pipelines beyond surface-level metrics.
  • SE for AI / Reliable Tooling: Structuring agent workflows, Model Context Protocol (MCP) integrations, and verification mechanisms for reliable execution.

🛠️ Background & Applied Experience

  • LLM Evaluation: Studied empirical LLM benchmarks, failure modes, and RAG architectures as a Visiting Scholar at York University.
  • AI Testing: Investigated combinatorial black-box testing (CIT) to uncover multi-attribute failure modes in deep learning models.
  • Agent Frameworks: Contributed to open-source agent orchestration (ISEK) and authored a book chapter on intelligent agent development (From DeepSeek to Manus, Tsinghua University Press).

💬 Connect & Collaborate

  • Email: peixuanxia@gmail.com
  • Looking for: PhD opportunities / Research collaborations in LLM Auditing & Evaluation, and empirical SE.
  • Pronouns: She/Her
  • Beyond Code: Passionate about light hiking, crochet & sewing, intersectional feminism, and animal welfare.

Pinned Loading

  1. isekOS/ISEK isekOS/ISEK Public

    A decentralized agent network for building collaborative, LLM-powered agent-to-agent (A2A) systems.

    Python 568 40

  2. DefuzeX-AI/AgentBehaviorBench DefuzeX-AI/AgentBehaviorBench Public

    A benchmark suite for running and evaluating AI agents across frameworks

    Python 10 13