Improving Data Preparation for CSV Files with LLMs
Alexander van Renen, Moritz Rengert, Macallyster Edmondson et al.
· 0 citations
2 papers indexed here
We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.
Not the right person? Other researchers publish under this name.
Hollywood is introduced, a synthetic IMDb-compatible benchmark generator that combines LLM-generated semantic dictionaries with deterministic temporal-graph-based relational data generation and demonstrates that Hollywood induces cardinality estimation errors comparable to or exceeding those observed on the original IMDb dataset.