Projects / Selected systems

Structured evidence,
made operational.

Four public case studies trace my work across AI safety assurance, multilingual NLP, quantitative modelling, and production operations. Each page records the problem, role, system, methods, and public scope.

See the research context
Featured / Public case studies

The portfolio spans research frameworks, applied AI, statistical models, and deployed software—showing how analysis moves into usable systems.

01AI safety assuranceSep.–Dec. 2025

AI Safety Testing Framework for LLM and Agentic Systems

A four-layer evaluation framework that translates public AI risk guidance into testing requirements, evidence structures, deployment decisions, and mitigation checks.

ArchitectureFour evaluation layersAssurance modelSeven trustworthiness dimensions
View public case study
02Multilingual content safetyAug.–Dec. 2024

AI-Driven Media Analysis for Toxic and Xenophobic Language Detection

A multilingual NLP system combining GPT-4 and a fine-tuned BERT model to support structured review of xenophobic language, toxicity, misinformation, and biased framing in media content.

SystemThree integrated analysis modulesInputsPDF, Word, and plain text
View public case study
03Quantitative researchDec. 2022–Jul. 2023

Forecasting the Term Structure of Interest Rates with Dynamic SV Models

A quantitative research project using ten years of Chinese bond-market data and six time-series models to analyze and forecast interest-rate behavior.

Data horizonTen yearsModel setSix time-series models
View public case study
04Product & systems deliveryJan. 2025–present

B2B Digital Operations Platform

End-to-end product, workflow, data, backend, and deployment ownership for a B2B platform spanning a WeChat Mini Program, back-office system, and company website.

Product surfaceMini Program, back office, websiteOperationsFive connected business domains
View public case study
Portfolio index / Roles & evidence

What each project made public.

These are case-study records, not an article library. Research papers, private datasets, working files, and proprietary system details remain off the site.

01
Evaluation framework

AI Safety Testing Framework for LLM and Agentic Systems

Role

Independent Researcher and Project Lead

Evidence

Four evaluation layers · Seven trustworthiness dimensions · Three case-study systems · Five practical artifact types

02
Applied NLP system

AI-Driven Media Analysis for Toxic and Xenophobic Language Detection

Role

Independent Researcher · UNICC-Sponsored NYU Applied Project

Evidence

Three integrated analysis modules · PDF, Word, and plain text · GPT-4 + fine-tuned BERT · Streamlit web application

03
Time-series modelling

Forecasting the Term Structure of Interest Rates with Dynamic SV Models

Role

First Author

Evidence

Ten years · Six time-series models · Chinese bond market · First author

04
Operational software

B2B Digital Operations Platform

Role

Product and Operations Manager

Evidence

Mini Program, back office, website · Five connected business domains · 34-page Mini Program · Requirements through adoption

Open channel / Collaboration

Working on agent safety,
evaluation, or multilingual AI?