We investigate the problem of building least squares regression models over training datasets defined by arbitrary join queries on database tables. Our key observation is that joins entail a high degree of redundancy in both computation and data representation, which is not required for the end-to-end solution to learning over joins.
No takes yet. Share an insight, caveat, or question.
Schleich et al. (2016) studied this question.
Synapse has enriched one closely related paper. Consider it for comparative context: