Wei Yang

Full-Stack Software Engineer at PiSrc

Email: hey.weiyang@gmail.com

Phone: +1 551 260 0541

Web: inscribedeeper.github.io

Professional Summary

Full-stack software engineer and data science practitioner with experience in production generative AI, enterprise search, natural language processing, recommendation systems, and data integration. Work spans secure RAG applications, multilingual retrieval pipelines, API-based enterprise data synchronization, cloud infrastructure, and applied NLP research. Experienced in translating machine learning methods into reliable systems for industrial and business platforms.

Areas of Expertise

Generative AI & Information Retrieval: Retrieval-augmented generation, conversational AI, agentic workflows, semantic search, hybrid retrieval, reranking, and AI safety controls

Enterprise Platforms & Data: Multilingual indexing, knowledge management, API integration, Elasticsearch, Solr, Weaviate, Redis, MuleSoft, Adobe Experience Manager, Azure, and AWS

Software Engineering: Python, Java, JavaScript, SQL, REST APIs, distributed caching, load balancing, Docker, Nginx, monitoring, and performance optimization

Machine Learning & Analytics: PyTorch, Hugging Face Transformers, scikit-learn, Pandas, NLP, recommendation systems, text classification, and statistical modeling

Selected Engineering Projects

Enterprise RAG Platform

Azure OpenAI, Redis, Weaviate

  • Led full-cycle development of a production conversational AI application supporting 1K+ daily active users
  • Designed multi-turn memory, context management, function-calling tools, sticky-session routing, load balancing, content filtering, and PII masking

Multilingual Hybrid Search

Enterprise Knowledge Retrieval

  • Built scheduled indexing pipelines spanning 10+ heterogeneous data sources
  • Implemented keyword and semantic retrieval, reranking, query expansion, multi-channel routing, iterative retrieval, caching, reporting, and user-feedback workflows

Enterprise Data Integration

MuleSoft, Solr, Elasticsearch

  • Designed 7+ API workflows for incremental synchronization of partner account and location data
  • Integrated multi-source records into search layers and optimized cache behavior for responsive retrieval

Recommendation Systems

Personalization and Offline ML

  • Developed user-to-item and item-to-item recommendation pipelines with rolling-cache and CDN integration
  • Built offline machine learning workflows for sales-funnel and marketing-campaign optimization

Professional Experience

PiSrc

Full-Stack Software Engineer

February 2022 - Present

Lead the development of production AI, enterprise search, data integration, and personalization capabilities for industrial digital platforms.

  • AI Chatbot & RAG Systems:
    • Led full-cycle development of a production-grade RAG conversational AI chatbot using Azure OpenAI, Redis, and Weaviate vector database; architected sticky session routing and load balancing to support 1K+ DAUs
    • Designed multi-turn conversational memory with Redis persistence and context window management; built agentic workflows with function calling to orchestrate custom tools and external APIs
    • Implemented security guardrails including input sanitization, content filtering, and PII masking to ensure safe and compliant AI interactions
    • Delivered real-time AI Overview and autosuggest powered by live user queries with caching layer for low-latency responses; built scheduled report pipelines and user feedback loops to continuously improve relevance and quality
    • Engineered scheduled multilingual (I18N) indexing pipelines across 10+ heterogeneous data sources, combining keyword-based and semantic hybrid search with semantic reranking, multi-channel query routing, query expansion, and iterative retrieval
  • Infrastructure & Platform Engineering:
    • Architected scalable full-stack infrastructure: VM provisioning, runtime orchestration, and offline pipelines for knowledge base synchronization and cache optimization
    • Maintained high reliability and low-latency performance through proactive monitoring and tuning
    • Applied data-driven insights to evolve CMS architecture and scale web platforms
  • Data Integration & Search Optimization:
    • Designed 7+ MuleSoft API integration workflows for incremental delta updates of partner accounts and locations
    • Integrated multi-source data into a Solr and Elasticsearch–based search layer with cache optimization
  • Personalization & Machine Learning Pipelines:
  • Content Platform (AEM):
    • Developed licensable software modules on Adobe Experience Manager to streamline content authoring and multi-channel publishing

Stevens Institute of Technology

Research Assistant/Teaching Assistant, Deep Learning and Web Analytics

June 2020 - December 2021

Language features pattern detection with Deep Learning models, data parsing and statistics modeling

  • Cleaned and structured 94,581 earnings call transcripts to explore 96 language factors influencing stock return
  • Achieved 72% accuracy on text classification with a domain-adapted BERT language model using PyTorch
  • Conducted text analysis, sentiment analysis, and text mining utilizing NLP techniques with SpaCy and NLTK
  • Set up and managed a remote GPU environment on Ubuntu for Machine Learning and Deep Learning tasks
  • Developed tutorials on implementing deep learning models using related Python packages

Education

Monroe University

Master of Business Administration (MBA), In Progress

2025 - Present

Selected Coursework: Strategic Marketing & Data Mining, Software System Design, Computer Networks, Research & Statistics for Managerial Decision-Making, Organizational Behavior & Leadership in the 21st Century, Managing in the Global Environment

Stevens Institute of Technology

MSc in Data Science (GPA 3.8/4.0)

September 2019 - December 2021

Relevant Coursework: Statistical Methods, Statistical Inference, Advanced Optimization Methods, Advanced Data Analytics & Machine Learning, Deep Learning, Natural Language Processing, Web Analytics, Database Management Systems, Web Programming, Data Structures & Algorithms

Guangzhou University

BSc in Mathematics and Applied Mathematics

September 2014 - May 2018

Relevant Coursework: Probability and Mathematical Statistics, Operational Research, Numerical Analysis, Advanced Algebra, Mathematical Analysis, Real Function Theory, Functional analysis, Ordinary Differential Equations, Partial Differential Equations

Awards & Honors

Provost's Scholarship

Stevens Institute of Technology

2019

Awarded the Provost’s Scholarship during the MSc in Data Science program.

National Mathematical Modeling Contest

Second Prize

2015

Ranked in the top 6.3% out of 25,558 teams.

Extracurricular Activities

UBS Quant Hackathon

2020

UBS Pitch Competition

2020

Stevens HealthTech Hackathon

2019

Academic Projects

MyPlace Web Development

Develop a web application for furniture and rental information exchange

  • Led team of four students to design and develop a web application with Node.js and Express
  • Designed document schema on MongoDB and wrapped CRUD operations as RESTful APIs
  • Implemented Login and user-specific functions such as account signup and authentication system
  • Developed features for rental and furniture information exchange such as comments, search, dashboard

E-commerce Recommender System

Develop an e-commerce recommender system using user, item, and interaction data

  • Created and implemented recommender engine with users, items, and interaction records from JD.com
  • Integrated multiple memory-based and model-based collaborative filtering algorithms to make recommendations
  • Simulated on 7,000 pre-defined user-item interaction samples and attained 77% Top-10 Accuracy

Fintech Pitch Competition - 6th Position

Develop a mathematical model based on public data and metrics to measure and predict the vibrancy of the city in the U.S.

  • Constructed a vibrancy index to interpret and predict the prosperity trend of cities in the U.S.
  • Explored, collected, and blended data from Google POIs, Instagram, Zillow, Bureau of Labor Statistics with Pandas
  • Performed feature engineering on panel data, and fine-tuned models for prediction with XGBoost and Keras

Quantify the AI impacts on Jobs skills

Data scraping, parsing and information extraction, topic modeling with clustering algorithms and neural networks

  • Scraped, cleaned and structured Job Descriptions textual data from INDEED.com
  • Filtered noisy data by selecting clusters from K-means algorithm and visualized skillsets distribution
  • Categorized, analyzed skillsets trend after Topic Modeling and Aspect Extraction Deep Neural Network

Analysis on Factors Influencing Bitcoin

Feature engineering, data analysis and modeling

  • Performed data cleaning and sentiment analysis for over 20 million related Tweets in second granularity
  • Delivered feature engineering for financial indicators like MACD and RSI etc
  • Analyzed, finetuned, and back-tested for over five types of deep learning model to increase the portfolio return

China Undergraduate Mathematical Contest in Modeling (CUMCM)

Guangzhou, China

September 2015

  • Implemented dilation and erosion algorithms and the Moving Average Method to extract information from images
  • Constructed a geolocation model based on shadow changes by modeling the solar altitude angle; solved the optimization problem with a Genetic Algorithm
  • Won the national second prize in the 2015 CUMCM in China