Product Classification API Part 2: Data Preparation
Part 2 of a series on building a product classification API, focusing on data cleaning, preparation, and measuring data purity for machine learning.
Eugene Yan is a Principal Applied Scientist at Amazon, building AI-powered recommendation systems and experiences. He shares insights on RecSys, LLMs, and applied machine learning, while mentoring and investing in ML startups.
188 articles from this blog
Part 2 of a series on building a product classification API, focusing on data cleaning, preparation, and measuring data purity for machine learning.
A presentation on Lazada's machine learning framework for ranking products in catalog and search results to improve user experience.
Announcing the launch of a product image classification API for fashion, built with deep learning, with details on performance and usage.
Author shares their acceptance into Georgia Tech's affordable online Master's in Computer Science program and their application essay.
First post in a series on building a product classification API, covering the process of sourcing and formatting open-source Amazon product data for machine learning.
A data scientist reviews Martin Odersky's Functional Programming in Scala Coursera course, covering key learnings and its practical application.
DataKind Singapore's Project Accelerator connects volunteer data scientists with nonprofits to solve data challenges, like analyzing water consumption data.
A summary of a talk on achieving top 3% in a Kaggle competition, covering validation, feature engineering, and ensemble techniques.