Well you won't have to worry what to do when they come for X: cos you are first![]()
Of course it makes uses of the information practical. This is AUTOMATED. I do not need to analyze the data personally.If this was addressed to me I think I did not make myself clear. The examples that you give about what you can say about an individual seem to make my point: of course we are predictable in our behaviour. But that says nothing at all about what we might do in odd circumstances.
For example it is absolutely true that if my shopping pattern changes something has changed. So what? There are loads of reason that will happen and there are loads of people at any given time it will happen to.
And yes, if you want to you can tell my politics from my posting. So what? They are not a secret. If they were I woud not be posting about them, would I?
You don't need to analyse my shopping records to find out who my friends are: we have CCTV for that. I object to CCTV but the vast majority of people in this country seem to actively welcome it. But again, I have school records and employment records and all sorts of things: It would not be hard to find out who I hang out with if you wanted to: and that is no different from any period in living memory.
What I do not see is any evidence that all this increase in information makes the use of that informaion practical: if anything it has the opposite effect. I am fairly sure that much more was known about people in rural and feudal societies than at any time since.
For instance, if people have been working a certain amount of overtime and are X% less likely to vote for each hour of overtime they've worked in the preceding two weeks leading up to the election (total information, remember?
Hi INRM:GreyICE,
There need to be laws and treaties to regulate this.
Agreed, but either nobody seems to care, the law hasn't caught up with the technology, and ethical standards as a result are not being kept in line with scientific advancement.
Science without morality, as anybody with a brain knows, is extremely dangerous.
Why would I care if they knew my call durations?
He's basically right, except society hasn't caught up to the point we need to be where google search info won't matter if it's private.Just remember as far as Google's CEO is concerned if you have concerns about your on-line privacy it's because you are up to no good!
"If you have something that you don't want anyone to know, maybe you shouldn't be doing it in the first place, if you really need that kind of privacy, the reality is search engines, including Google, do retain that information." - Schimdt
Hi Dancing David,
One word: cookies. Google tracks you using a cookie they sent to your browser. Every site that subscribes to Google Analytics (and there are a LOT of them) causes your browser to make a connection to Google so that Analytics can update their information. And the unique cookie that Google assigned to your browser, which is permanent unless you clear it, gets sent with that request. The cookie is independent of your current IP address, and it remains across internet connects, browser restarts, and system reboots. (But if you use two different browsers, they will have two different cookies.)(much stuff snipped)
But here is the thing, when you use the Internet, a lot of data is labeled and passed around, and it is all data, Google is a service, they track what people do, first off that is what makes Google so strong, their searches are sorted by the data they gather and the order on the page reflects the aggregate usage. Second they get your IP address and will track what that specific IP address it doing. Now if you use Google Desktop they will get a whole lot more information about you.
That's true of ADSL. I'm using a cable provider, and I have had the same IP address for over a year.But here is the thing, because of DHCP you basically gets a new IP address each time you log on to your machine, so it is hard for them to know exactly who you are. Which is why tracking cookies and spy ware and important to keep off your machine, because they do a much better job.
I'm on Linux, and prefer using mtr
He's basically right, except society hasn't caught up to the point we need to be where google search info won't matter if it's private.
the thing about embarrassing google search information is that it only holds power over you if society maintains a puritanical double standard. For as long as we pretend that porn is for the other guy or that my interest in categorizing poop smells is shameful, we will have a need for private info.
Hi Dancing David,
That's a good overview of IP. I want to comment on a couple of points.
One word: cookies. Google tracks you using a cookie they sent to your browser. Every site that subscribes to Google Analytics (and there are a LOT of them) causes your browser to make a connection to Google so that Analytics can update their information. And the unique cookie that Google assigned to your browser, which is permanent unless you clear it, gets sent with that request. The cookie is independent of your current IP address, and it remains across internet connects, browser restarts, and system reboots. (But if you use two different browsers, they will have two different cookies.)
That's true of ADSL. I'm using a cable provider, and I have had the same IP address for over a year.
I'm on Linux, and prefer using mtr![]()
Really? Why can one not calculate X for each voter? One merely needs sufficient amounts of information, and sufficient processor time. Neither is a scarce resource. The standard deviation is hardly as huge as you think - people are reasonably predictable, and we have a good thousand of them or so to play with. That's more than enough for standard dev to be minimal.But that value of X is just an average (even in 2025). One cannot calculate X for each voter. Furthermore the standard deviation on that distribution will be huge. Lastly given that the data is based on only a few data points for each person, the amount of noise in that statistical function will be so large as to make the calculation meaningless when trying to apply the average value of X to a single individual or even a small group of individuals. Lastly, trying to take action so that a specific person is required to work overtime is absurdly complicated. You have not convinced me.
Or you are trying to plan a surprise get together for someone one, or looking for advice on how to save your marriage or...
There are many, many things that many people would wish to be private that do not fall into the either the social taboo nor illegal camp.
Really? Why can one not calculate X for each voter? One merely needs sufficient amounts of information, and sufficient processor time. Neither is a scarce resource. The standard deviation is hardly as huge as you think - people are reasonably predictable, and we have a good thousand of them or so to play with. That's more than enough for standard dev to be minimal.
Finally, your entire assumption is based around limitations - you assume that the action is limited to one single thing.
You assume that the number of data points is limited.
You assume the data gathered is too poor to have a good reliability.
You assume that no one has the processor time to compute this.
You are going about this the wrong way. You are insisting that we must act like it's the 20th century, as if computers had not been invented yet.Yes, all one needs is sufficient amounts of information. The catch is that people vote once every four years. Assume your target is 33 years old. He has had the chance to vote in in four elections. Say he voted in three of them and missed one of them. That's already public record and easily obtainable. The reason he didn't vote in that one election is not deducible no matter how much electronic information you have. Maybe he had a sore throat that day, maybe he simply forgot, maybe he drove there and the line was too long, maybe he got a flat tire, maybe inclement weather kept him away, maybe he saw the difference in polls going into election day and decided his vote wasn't needed, maybe it really was the amount of overtime he worked in the past seven days (although, I am unsure how unlimited internet tracking will allow a third party to determine how much overtime was worked in the past week). I'll agree that if you tracked his every movement on the internet, you could infer with a certain reliability who he would vote for, but determining if he is going to vote in an upcoming election requires more than four data points spread out over 16 years.
Do you see what is happening here? Overtime is not the ONE DECIDING factor. On this we both agree. But it is a factor. On this we most likely probably agree (if your argument is that people decide to vote by rolling dice, I suppose we'll have to agree to disagree - but I would be agreeing to be right, and you would not be).No, I don't. You were the one who asserted that the amount of overtime worked was the deciding factor.
The fact that the number of data points related to election behavior is limited is beyond question.
No, I do not. I will agree that processor time is not the limiting factor.
You are going about this the wrong way. You are insisting that we must act like it's the 20th century, as if computers had not been invented yet.
This is nonsense. Last presidential election, we got exit polls. Assuming a nice spread of data points, we most likely got anywhere from 50,000 to 200,000 data points. Improved surveying methods might drive this up.
We in fact have very good data about how a 26 year old male who voted democratic, worked 55 hours a week on average, was college educated with a liberal arts major, and regularly followed the news voted. We also have demographics of the average 26 year old male, the number of hours a week they worked, their education level, etc.
In fact, we can isolate many factors. The number of recent positive versus negative articles they remember about their candidate. The effect increased work hours has on that demographic. The effect that different issues had.
You act like this is impossible. It is already occurring. Our methods simply, in terms of sophistication, are primitive. You are staring at a sharp rock, thinking about metalworked blades, and determine that no one would dig up a bunch of rock, melt it, purify out different metals, and forge it just to get something sharper than that sharp rock (it would take years, if not decades for someone to make themselves this new sharper rock (call it a knife).
Of course you are correct. No one would go through all that trouble for one knife. But the fact that you are correct does not mean that you have entirely thought the matter through.
Do you see what is happening here? Overtime is not the ONE DECIDING factor. On this we both agree. But it is a factor. On this we most likely probably agree (if your argument is that people decide to vote by rolling dice, I suppose we'll have to agree to disagree - but I would be agreeing to be right, and you would not be).
You further agree that overtime's effect on voting can probably be analyzed. Finally you note that processor time is not a limiting factor.
But we are at a quandary here. If election behaviors are based on a number of factors, then the factors can be analyzed. All we need is good data. You have agreed to this. You have limited the number of data points to one per year (rather than one per person per election year, which seems a tad more likely).
Then you have concluded the data is poor because we only get one point per year rather, than, say, 250 million or so.
It's somewhere in-between the two. Each year it gets closer to the latter than the former. So the quality and granularity of the data is improving.
Do we agree or disagree on this basic premise?
But why? It's rather useless to make sure that a specific voter does not turn up. It's rather a tad more useful to make sure that five hundred, or one thousand voters do not turn up. If an industry employs, say, one thousand workers, and you can calculate they are 70% likely to vote and 80% likely to vote for your opponent, you can use that for informed decision making.
If one can identify 3,000 voters who are likely to turn up and make them 3,000 voters who are unlikely to turn up, that's a very big margin of change. And there is beginning to be enough information out there that a computer can automatedly generate the highest probability methods of making sure those people do not vote - tailored for each individual.