I build LLM-powered tools, agentic workflows and data pipelines โ and ship them end to end.
Independently designed, built, tested and shipped โ from product idea to portable release.
Video โ timeline-annotated Markdown notes. A local-first service: feed it a Bilibili / Douyin / YouTube link or a local file, it pulls platform subtitles (incl. Bilibili AI subtitles) or transcribes offline with faster-whisper, then an LLM writes structured, time-stamped notes.
Smaller tools and notes โ open by default.
A modern desktop client for the DeepSeek Harness (DSH) plugin ecosystem โ everything is a plugin, including the desktop itself.
pi extension: native vision tool with multi-backend fallback (Step-3.7-Flash โ GLM โ Qwen) for image understanding.
Notes on the Hot 100 problems I got wrong, over-thought, or missed the key insight on.
From quantitative finance to LLM-driven data engineering.
Built an automated pipeline over 4M+ multi-source records across 8 categories: LLM semantic classification โ Elasticsearch vector retrieval โ industry-chain knowledge-graph mapping โ MySQL. Multi-model cross-validation (ChatGPT / GLM / DeepSeek) with prompt engineering; Python multithreading + checkpoint-based retries for month-long batch jobs; local deployment of BGE-M3 / GPT-OSS to cut commercial API costs.
Developed & tested government-bond / policy-bank-bond yield-curve models (2022โ2024) and Z-Spread pricing; MA / ARIMA / Holt-Winters time-series models with adaptive tuning, extended across multiple bond types.
GPA 3.98 / 4.3 ยท TA for Data Mining (Python) ยท Academic Scholarship (First & Second Class)
Econometrics 97 ยท Machine Learning 94 ยท Mathematical Statistics 90
Improved PageRank over a 5,600+ node musician-influence network; PCA + cosine-similarity genre similarity matrix.
Logistic credit-risk default model + nonlinear programming for SME lending terms.
Tools I reach for daily.