Logo Questions Linux Laravel Mysql Ubuntu Git Menu
 

Sample data generation from existing data - Algorithm?

I'm facing a what-am-I-looking-for problem again. I've got tracking data, meaning I tracked myself riding a bike a couple of times on the same track in order to have some test data. So I've got time-distance pairs. I want even more different ones, but want to generate them. I want the virtual test drivers to be both faster and slower. I don't want it to be a linear summing up algorithm.

Basically I just want a hint, what I could search for. Can somebody help out?

like image 546
rdoubleui Avatar asked Sep 24 '26 10:09

rdoubleui


2 Answers

You could check out Markov Chains for generating random data although it depends on how complex you want your distribution of data to be. For a simpler approach, Reed Copsey's solution is going to be easier to implement.

like image 57
Adamski Avatar answered Sep 28 '26 13:09

Adamski


Try multiplying each data point's time value by a constant. If it's less than one you'll be creating a faster test drive, and if it's more than one you have a slower test drive. You can also fudge the position similarly if you want.

like image 40
redtuna Avatar answered Sep 28 '26 14:09

redtuna



Donate For Us

If you love us? You can donate to us via Paypal or buy me a coffee so we can maintain and grow! Thank you!