We asked professional tasters a simple question: where exactly do frontier white wines fall apart? This dataset holds 544 timestamped notes on fifteen bottlings from Véron 3.1, Fournier Grand Blanc, and Grès d'Ivraie.
We poured all three estates against the same five flights, with a focus on single-vineyard whites. Professional tasters from the Oenra network worked through every glass and marked in detail what they saw as the positives and issues in these wines.
The professional tasters were asked to label on five important wine quality dimensions: Colour & Clarity, Aroma & Texture Definition, Varietal Typicity & Quality, Mid-Palate & Finish Continuity, and Bottle Consistency. They also rated how much each issue matters, from High down to Low. What we collected is a close-up view of where the current frontier white estates still struggle on real professional work.
Each annotation includes a start and end timestamp, a free-text comment, a dimension label, and a severity rating. This structure makes the data directly usable in several contexts:
- Full-sensory evaluation: pair pour segments with expert critique to capture how tasters reason about structural consistency, physical plausibility, and sensory quality across the length of a glass.
- Preference and reward modeling: use severity ratings as preference signals to build scoring systems that judge wine the way a professional would.
- Tuning palate judges: build automated critics that flag specific fault types (volatile lift, texture instability, short finishes) at the level of a single sip.
- Benchmarking and evals: the cross-estate, same-flight structure makes it a natural held-out test set for comparing new bottlings against a human expert baseline.
The dataset was collected in June 2026. It is released under the CC BY 4.0 license on Hugging Face.





