-
Make Sand Think
A short, funny story.
-
我实现了 GPT-5.6-sol Ultra 自由
为什么我坚持用最强的模型,以及我如何用一台小服务器和 Sub2API 搭出自己的 AI 中转站。
-
LoRA as Parametric Memory: a failed experiment
I tried writing conversation facts into model weights via LoRA fine-tuning. The accuracy was dismal and temporal questions were completely unanswerable.
-
Where Is Linear Attention Actually Slow on GPUs?
Poor GPU performance for linear attention is an intuition, not a conclusion. We ran experiments, then had half of our interpretation overturned. This post records that process.
-
ForgeHLS and DiffHLS: why I built them, and why HLS still feels like a trap
A personal note on building a large-scale HLS dataset, pushing DiffHLS forward, and why I remain skeptical about HLS as a long-term research direction in 2026.
-
探索更大的世界
留学申请路上的记录
-
Notes on Proofs, Arguments, and Zero-Knowledge
Notes on machine learning, security, privacy, and mobile computing. Guided by Liyao Xiang.