Skip to content
View Excelius-Wang's full-sized avatar

Highlights

  • Pro

Block or report Excelius-Wang

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Content in all repositories owned by your account will be closed.
Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Excelius-Wang/README.md

banner

Excelius

LLM post-training · alignment · evaluation · agents 🤖 — MSc @ BJTU

WebsiteEmail


About Me

  • 🎓 MSc in Software Engineering @ Beijing Jiaotong University; BEng @ Nantong University.
  • 🔬 I work on LLM post-training, alignment, and evaluation, and build LLM agents and developer tools.
  • 🛠️ I contribute fixes to open-source training, evaluation, and agent frameworks, with a focus on numerical stability, metric correctness, and reliable runtime behavior.
  • 📝 Co-authored a research paper.
  • ✍️ I write notes & blog posts at excelius.xyz.

Experience

  • ByteDance — Multimodal LLM Algorithm Intern · present
  • Baidu — LLM Post-training Algorithm Intern
  • Tsinghua University, Institute of Vehicle Power & Intelligent Energy — LLM Application & Full-stack Intern

Open Source Contributions

45 merged PRs across external open-source repositories, covering code fixes, tests, and documentation. Counts below include merged PRs only, as of September 13, 2026.

Other merged contributions include Atomic Agents (4), LiveKit Agents (2), MCP Servers (1), OpenAI Agents JS (1), and Axolotl (1). Examples: MCP resource templates · turn cancellation · UTF-8 file reads · hosted MCP outputs · activation checkpointing compatibility.

Tech Stack

  • Languages: Python, C++, TypeScript / JavaScript, Rust
  • LLM / Post-training: PyTorch, Hugging Face Transformers, LLaMA-Factory, ms-swift, vLLM
  • Alignment / Evaluation: SFT, LoRA, DPO, PPO, GRPO, reward modeling, LLM-as-a-Judge, EvalScope
  • Agents: LangChain, LangGraph, tool calling, ReAct, plan-and-execute workflows
  • Applications / Infra: React, Vue, FastAPI, Tauri, multi-node multi-GPU training, Git, Docker

Featured Projects

  • Repolane — A GitHub desktop workspace for browsing code, reviewing pull requests, and managing repository work, built with React, TypeScript, Rust, and Tauri. Repository: harbor; actively developing, with no packaged public release yet.
  • Self-DeepResearch — An iterative research agent built with LangGraph and Tavily: plan → search → review → report, with a Vue frontend and FastAPI streaming backend.
  • marginalia — A paper-reading skill for Claude Code / Codex that reconstructs a paper's reasoning, examines its assumptions, and publishes structured notes with figures and formulas to a Feishu knowledge base.
  • dive-into-transformer-pytorch — A Transformer language model implemented in PyTorch and trained on Dream of the Red Chamber, with DDP multi-GPU training, checkpoint saving, and training-curve visualization.
  • BERT_BiLSTM_CRF — Chinese named-entity recognition with BERT + BiLSTM + CRF, developed for BJTU NLP coursework.

📊 Dashboard

ghfind GitHub 评分卡

Pinned Loading

  1. modelscope/ms-swift modelscope/ms-swift Public

    Use PEFT or Full-parameter to CPT/SFT/DPO/GRPO 600+ LLMs (Qwen3.6, DeepSeek-V4, GLM-5.1, InternLM3, Llama4, ...) and 300+ MLLMs (Qwen3-VL, Qwen3-Omni, InternVL3.5, Ovis2.5, GLM4.5v, Gemma4, Llava, …

    Python 15.6k 1.7k

  2. openai/openai-agents-python openai/openai-agents-python Public

    A lightweight, powerful framework for multi-agent workflows

    Python 29.4k 4.7k

  3. modelcontextprotocol/servers modelcontextprotocol/servers Public

    Model Context Protocol Servers

    TypeScript 90.3k 11.6k

  4. OpenRLHF/OpenRLHF OpenRLHF/OpenRLHF Public

    An Easy-to-use, Scalable and High-performance Agentic RL Framework based on Ray (PPO & DAPO & REINFORCE++ & VLM & TIS & vLLM & Ray & Async RL)

    Python 10k 1k