Posted last month
The Data Understanding team creates high‑quality datasets and quantized representations for OpenAI, focusing on data synthesis, VQ representations, and processing to improve model training. The role involves researching data selection, curation, and experimental validation to drive scalable data pipelines.