- π Now: AI Native Architect @ Bluesmart Digital β Shipping intelligence from prototype to production
- π§ Focus: FDE Β· AI Native Architecture Β· Agent Workflow Deployment Β· RAG / GraphRAG
- π Research: Education Agent β SCI Q1 journal under review Β· 1 ETNCC conference paper
- π± Community: Silver-tier organizer at Datawhale open-source AI community
- π¬ Reach me: threethreeliu@163.com
π§ My world revolves around Mercury β Mercury At The Center Of My World
Turning LLM capabilities into shippable, iterable, production-grade engineering systems β this is the through-line of everything I do today.
| Track | Core Capabilities | Toolchain / Keywords |
|---|---|---|
| 𧬠Model Training & Fine-tuning |
SFT Β· LoRA Β· Parameter-Efficient Fine-tuning (PEFT) RLHF Β· DPO Β· PPO Β· GRPO β alignment & reinforcement learning Model internals & attention mechanism analysis Distributed training, GPU memory & throughput optimization |
PyTorch Β· DeepSpeed Β· CUDA Β· PEFT Β· Transformers Β· TensorFlowTRL Β· trl-lib Β· vLLM Β· Megatron-LMβ GPU resources optimized 25% Β· dev velocity +40% |
| π€ Agent Engineering |
Planning-Execution-Feedback orchestration Multi-agent collaboration Β· role-based task division Tool calling / state machine / context management |
LangGraph Β· LangChain Β· AutoGen Β· ReAct Β· Function Callingβ Solvely.ai Β· Operations Hive β HiveMote |
| π§ͺ PromptOps & Evaluation |
Automated failure-sample feedback loop (AutoFeedbackLoop) Template perturbation Β· multi-model adapter layer Β· fallback strategy Badcase governance & effect alignment |
PromptOps Β· A/B Eval Β· LLM-as-Judge Β· Regression Suiteβ Significant stability gains on long-horizon tasks |
| π RAG / GraphRAG / Retrieval |
Document parsing Β· knowledge extraction Β· Knowledge Map Vector retrieval Β· hybrid retrieval Β· Reranking Citation tracing Β· multimodal QA Β· multi-document conversation |
RAG Β· GraphRAG Β· Vector DB Β· BM25 Β· Embedding Β· Schemaβ Solvely Long-Document AI Content Engine |
| π Production Engineering |
Streaming output Β· explicit caching Β· context state machine Token Mask Β· inference cost governance Canary release Β· auto-scaling Β· monitoring & alerting |
Streaming Β· Cache Β· Docker Β· K8s Β· Linux Β· Aliyunβ TTFT 35.7s β 17.4s (-51%) Β· Token cost -90% Β· throughput +91% |
π "Mercury At The Center Of My World" β β From threethreeliu.top

