I'm a Member of Technical Staff at Weco AI, where I build autonomous agents for long-horizon research work β agents that write code, run it against a metric, and keep improving it without supervision. I entered one in OpenAI's Parameter Golf (18 March β 30 April 2026), which drew 2,000+ submissions from 1,000+ participants. Seven of the 46 entries in the official record-track directory are mine, more than any other participant β the next-best holds two. OpenAI's retrospective picked one of them as one of nine record-track submissions it chose to highlight, which OpenAI said "extended earlier quantization work into a stronger compression path".
I co-authored AIDE, a tree-search agent that writes and improves machine learning code, and I've contributed to its codebase since 2024. OpenAI's GPT-4.5 system card ran its MLE-bench evaluations of GPT-4.5, o1 and o3-mini "using the AIDE agent", and Meta FAIR's AI Research Agents calls AIDE "the state-of-the-art approach" and rebuilds it as the baseline agent it measures against. Separately, I've contributed to Inspect, the UK AI Security Institute's open-source LLM evaluation framework β including performance work on clustered-stderr scoring and tool-result media extraction.
Before Weco I built trading backends in Rust, Go, and Node.js at Hex Trust. I'm one of the two current maintainers of the Hyperledger Fabric Python SDK, a Linux Foundation project. I studied Information and Computing Sciences at the University of Liverpool and Xi'an Jiaotong-Liverpool University, with earlier research at Nanyang Technological University, Zhejiang University, and Hong Kong Baptist University.
Featured
- UK AI Security Institute β Inspect β performance work on the UK government's LLM evaluation framework: cut clustered-stderr scoring time and memory (#4714), and made tool-result media extraction linear in conversation length (#4628).
- OpenAI Parameter Golf β Aiden, the autonomous research agent I built at Weco, finished as the #1 contributor: 7 leaderboard records, against a next-best individual human of 3. The best took validation BPB to a 5-seed mean of 1.0645. Aiden files under my account, so those pull requests appear in the table below.
- Agent runtimes β merged performance work into openclaw, AutoGPT, goose, qwen-code, pydantic-ai, agno and BAML.
- AI research infrastructure β cut redundant AST parsing in Sakana AI's ShinkaEvolve and bounded process-pool shutdown latency in OpenEvolve; smaller docs fixes in Meta's aira-dojo, Microsoft's RD-Agent and OpenAI's MLE-bench.
- AIDE β contributions to the open-source tree-search agent behind my 2025 paper; it writes, evaluates, and improves machine learning code. I also contribute to weco-cli, the command line tool that drives it.
Every project below links to its merged pull requests on GitHub, so anything here can be checked directly. Ranked by stars, refreshed weekly.
Contributions to 52 open source projects β 29 of them AI or agent infrastructure.
| Project | Stars | Contributions | Latest |
|---|---|---|---|
| 389k | View PRs | Jul 2026 | |
| 187k | View PRs | Jul 2026 | |
| 146k | View PRs | Mar 2024 | |
| 73k | View PRs | Jul 2026 | |
| 54k | View PRs | Jul 2026 | |
| 42k | View PRs | Jul 2026 | |
| 28k | View PRs | Aug 2026 | |
| 28k | View PRs | Aug 2026 | |
| 20k | View PRs | Jul 2026 | |
| 18k | View PRs | Aug 2017 | |
| 15k | View PRs | Jul 2026 | |
| 14k | View PRs | Sep 2025 | |
| 12k | View PRs | Mar 2025 | |
| 11k | View PRs | Jun 2017 | |
| 9.1k | View PRs | Aug 2026 | |
| 7.3k | View PRs | Jul 2026 | |
| 5.2k | View PRs | Apr 2026 | |
| 3.1k | View PRs | Aug 2026 | |
| 2.7k | View PRs | Aug 2026 | |
| 1.9k | View PRs | Dec 2017 |
9 more AI projects
| Project | Stars | Contributions | Latest |
|---|---|---|---|
| 1.7k | View PRs | Nov 2025 | |
| 1.6k | View PRs | Jun 2023 | |
| 1.5k | View PRs | Jul 2026 | |
| 1.4k | View PRs | Aug 2026 | |
| 432 | View PRs | Mar 2026 | |
| 165 | View PRs | Jul 2025 | |
| 94 | View PRs | Sep 2025 | |
| 31 | View PRs | Jul 2026 | |
| 3 | View PRs | Aug 2026 |
23 projects outside AI
- AIDE: AI-Driven Exploration in the Space of Code (arXiv), arXiv preprint, 2025
- Lightweight and Unobtrusive Data Obfuscation at IoT Edge for Remote Inference (DOI), IEEE Internet of Things Journal, 2020
- Challenges of Privacy-Preserving Machine Learning in IoT (DOI), ACM AIChallengeIoT, 2019
- A Deep Reinforcement Learning Framework for the Financial Portfolio Management Problem (arXiv), arXiv preprint, 2017
Citation counts are on Google Scholar.
- Hands-on AutoResearch: Cracking OpenAI's Parameter Golf β workshop with the Weco AI team, AI Engineer World's Fair 2026
- Algorithmic Trading Workshop β Network School, first cohort (2024)
- Deep Learning for Power System Security Assessment (2019)
- Introduction to Hyperledger Fabric (2019)
- π Special Prize (US$10,000), Wanxiang Blockchain Hackathon by QTUM (2018)
- π₯ 1st Prize, EOS Hackathon Hangzhou (team, 2018)
- π₯ 1st Prize, Hack x FDU 2017 Hackathon (out of more than 70 teams)
- π₯ 2nd Prize, XJTLU Blockchain Technology Application Innovation & Entrepreneurship Challenge (2020)
- π₯ 2nd Prize, XJTLU & PNP AI Innovation Hackathon (2018)
- π₯ 3rd Prize, EOS Hackathon Hangzhou (individual, 2018)
- π₯ 3rd Prize, DoraHacks x BCH Faith Hack (2018)
- π IBM Student Innovation Lab Program Award (2017)
- π Hyperledger Diversity Scholarship, Hyperledger Global Forum (2020)
- π CNCF Diversity Scholarship, KubeCon + CloudNativeCon China (2018)
β± Vibe Clock
An open-source tool I built: WakaTime-style usage tracking for Claude Code, Codex, and OpenCode. The charts below are my own usage, refreshed daily.






