I have several point data sets representing: A. high school students who bicycle to school; B. high school students who don't bicycle to school; C. features of the road network that discourage travel by bicycle, and D. features that encourage cycling.
When mapped, the above features appear to be related, ie. there are more students who bicycle where there are more encouraging/less discouraging features and more students who don't bicycle where there are more discouraging/less encouraging features (see attached image).
Legend - Light blue = students who cycle
- light brown = students who don't cycle
- red/pink = discouraging features
- green = encouraging features
My question is how do I prove this statistically?