Thanks to visit codestin.com
Credit goes to github.com

Skip to content
View AdrenalineY's full-sized avatar

Highlights

  • Pro

Block or report AdrenalineY

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Content in all repositories owned by your account will be closed.
Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
AdrenalineY/README.md

Yan Xiao | Hello World

email github profile views

M.S. in Software Engineering @ Nanjing University
Focus: LLM Post-training, Agent Systems, LLM for Software Engineering

Pac-Man Contribution Graph

pacman contribution graph

Profile

  • 求职方向:大模型研发工程师 / Agent 开发工程师
  • 研究与实践方向:大模型后训练、代码智能、Agent 上下文编排、工具系统与记忆系统
  • 技术关键词:SFT, DPO, FIM, ReAct, LangGraph, RAG, LLM-as-Judge, Pairwise Evaluation

Education

  • 南京大学软件学院 软件工程硕士,2025.09 - 2027.06
  • 南京大学软件学院 软件工程学士,2021.09 - 2025.06

Selected Work

中国电信广东研究院 | 大模型研发工程师实习(2024.11 - 2025.05)

  • 参与电信研发云内部企业级代码补全模型训练与优化,负责训练数据治理与训练算法调优
  • 结合 AST 分析开发者补全习惯,设计贴近真实使用场景的数据合成策略,构建百万条量级高质量训练数据集
  • 采用 SFT + DPO 两阶段训练流程,针对上下文重复等异常输出构造偏好数据;输出冗余长度下降 15%+,上下文重复率下降 60%+
  • 模型上线研发云平台,代码补全采纳率由 15% 提升至 25%+;核心方法投稿 EMSE(CCF-B,第二作者,送审)

每日简报与知识管理 Agent 系统 | 独立开发(2026.04 - 2026.06)

  • 独立设计并实现本地优先的 AI 信息简报与知识管理系统,形成“采集—筛选—沉淀—维护”闭环
  • 设计统一采集链路,串联去重、向量粗筛与 LLM 精筛,以较低成本生成来源可追溯的每日简报
  • 基于 LangGraph 构建受控 ReAct Agent,通过搜索、读取、修订、关联与审核提案工具辅助维护知识库
  • 设计分层记忆机制,将会话经验、长期摘要与待复核内容分层维护,并在运行时召回相关记忆辅助决策

腾讯犀牛鸟 NES 解释生成校企科研合作项目(2025.11 - 2026.04)

  • 面向下一步编辑建议(NES)解释生成场景,基于 Agent 执行轨迹构建数据合成与 LLM 评测链路
  • 搭建 NES 与解释数据的合成、筛选和质量标注流程,初步合成 1k+ 高质量 NES 数据
  • 构建 Pairwise LLM-as-Judge 评测体系,结合 ELO 与 Bradley-Terry 计算能力分数,并优化位置与长度偏差

仓库级代码补全检索模型 CARLCoder 研究(2025.06 - 2025.09)

  • 研究多段上下文召回对仓库级代码补全效果的影响,提出 RRL 指标量化复合上下文有效性
  • 设计基于强化学习的 CARLCoder 检索模型,EM 与 ES 相较此前方法分别提升 43.8% 和 15.6%
  • 相关成果投稿 EMNLP(CCF-B,学生二作,送审)

Stack

  • 编程与训练:Python, PyTorch, SFT, DPO, 强化学习
  • Agent 与检索:LangGraph, ReAct, Qdrant, BM25, Embedding, LLM API
  • 工程开发:TypeScript/JavaScript, Node.js, Vue 3, Java, Spring Boot
  • 技能证书:华为 HCIA-AI 人工智能工程师(2025.12)

python pytorch transformers llm training rag llm judge

3D Contribution Profile

3d contribution profile

Generated SVG Dashboard

metrics base

languages.indepth habits
notable isocalendar.fullyear

stars calendar.full

Contact

Pinned Loading

  1. AdrenalineY AdrenalineY Public

    2

  2. AITravelPlannn AITravelPlannn Public

    TypeScript

  3. LLM4SE-ImageWatermark2 LLM4SE-ImageWatermark2 Public

    再探 vibe coding

    Python

  4. wanghaao/ImpAPTr wanghaao/ImpAPTr Public

    A Tool For Identifying The Clues To Online Service Anomalies

    Python 13 12