← Articles
CW

Chen Wei

Author

LLM fine-tuning specialist at a Beijing AI lab, focused on domain-specific adaptation of foundation models for legal and medical text. Built the RLHF pipeline that outperformed GPT-4 on Chinese legal benchmarks.

aillmperformancebackend

Articles by Chen Wei(3)

Building a classification pipeline with zero-shot and few-shot prompting

Zero-shot and few-shot prompting turn LLMs into flexible classifiers without fine-tuning. This guide covers how to design, evaluate, and productionise a classification pipeline that handles multi-label, hierarchical, and confidence-scored tasks

llmclassificationfew-shotzero-shotprompt-engineering

16 July 2026

LLM latency optimisation: batching, caching, and model selection

A deep-dive into the techniques that actually move the needle on LLM response times in production: request batching, prompt caching, model routing, and infrastructure choices that compound to cut p99 latency by 60% or more

llmlatencyperformancecachingbatching

16 July 2026

Chen Wei — ANN Tech — ANN Tech