Past Experience

02

Data Science

Jiritsu Network

ProjectAutomated Token-Issuer Screening Infrastructure with Scalable ETL, Ranking Models & Counterparty Shortlisting

  • Representative example: re-implemented GSIA negative screening and best-in-class selection as a two-stage “screen out bad, then select the best” workflow, moving from 18,000 CoinGecko tokens to an 800-issuer universe and a 30-issuer shortlist for commercial, risk, and governance review
  • Engineered a Python ETL with Prefect, DuckDB, and Parquet to ingest, standardize, and refresh large-scale token data for issuer analysis
  • Built negative screens covering liquidity and pricing validity, data completeness, and basic governance and disclosure red flags to remove unworkable names early
  • Added business relevance filters to focus on RWA, stablecoin, serious DeFi, and infrastructure issuers, excluding meme, gambling, and NSFW categories
  • De-duplicated multi-chain and wrapped assets and mapped tokens to issuer-level entities to produce a partnership-ready counterparty list
  • Trained a LightGBM pairwise ranker for “go for good,” using uncertainty sampling to concentrate labeling on the hardest boundary decisions
  • Used Leiden clustering to surface look-alike issuers and improve coverage and diversity beyond the top-ranked candidates
  • Led governance and disclosure engagements with shortlisted issuers, aligning on reporting cadence, reserve transparency, and inputs needed for on-chain verification
  • Structured 6 tokenization pilots with lower-risk counterparties, tying selection to clear disclosure commitments and verification readiness
  • Made the pipeline repeatable and monitorable with scheduled runs, retries on failure, and Parquet-based outputs for fast reruns and backfills

Euromonitor International

ProjectLarge-Scale E-Commerce Data Standardization via Automated SKU ETL & Market-Entry Scenario Modeling

  • Owned the China workstream for a global green home appliance market-entry engagement for a European brand, translating energy-efficiency standards and subsidies into launch timing, pricing, and positioning scenarios for cross-regional strategy discussions
  • Built category- and city-tier scenarios for timing and price positioning, accounting for subsidy volatility, tightening efficiency standards, and consumer trade-up in mid- to high-price bands
  • Represented China in global stakeholder calls by aligning definitions of green products, calibrating expectations on policy rollout pace, and flagging approval, certification, and channel-execution risks
  • Standardized 50,000 e-commerce SKUs across 5 platforms and 7 major appliance categories through an automated Python ETL pipeline, enabling consistent comparison of efficiency labels, price points, promotions, feature sets, and basic review signals
  • Packaged the ETL logic into a reusable template for other APAC markets, doubling processing throughput and reducing duplicated manual data cleaning
  • Visualized competition and shelf density in Tableau to surface under-served combinations of high efficiency and mid- to high-price positioning, especially in tier-1 and tier-2 city demand
  • Representative example: separated national programs from local pilot subsidies and mapped transitions from old to new efficiency standards to identify categories with stable requirements versus those likely to tighten again; applied the same approach across multiple categories and market scenarios

China Construction Bank Asia

ProjectQuantitative Issuer Scoring System using PCA-Based Universe Construction & Methodology Governance

  • Applied principal component analysis to 10 sustainability indicators across 2,000 issuers, constructing a ranked investment universe that highlighted the top 28 percent for follow-on credit and equity analysis
  • Specified scoring logic, data sources, and edge cases in a concise methodology note, clarifying model behavior for product managers and risk officers evaluating the scoring platform’s adoption