You probably going thats what i've been saying the whole time.
I think you are right that we aren't entirely in disagreement. I think you are also correct that my communication of what I am seeing in my experiment is lacking.
I'll try again this time hopefully with better clarity even though I know my observations depart from concepts you have provided documentation for and claim Google is using.
Google is sorting my experiment pages which test a secret keyword in different semantic nodes of HTML. There is a definitive order that is not random coming from Google that slowly changes over time.
I do not change my experiment pages yet the sort order of them coming from google slowly changes over time. If Google is Tag agnostic I would rarely expect change in the order of unchanging documents.
Upon swapping better ranking semantic tags, which is what I can alter in my source, as deemed by my experiment I am seeing substantial and continued gains for the 15 aged sites I manage. The reasons for this aren't entirely understood but I am inclined to think the content changes I made based on the experiment are somehow related, but I can't prove that at this point even though my observations are very suggestive.
I like the idea of tag agnostic content scoring but I believe that when Google starts sorting documents by relevance that it ultimately is telling you something about the "learned reputation" of the semantic usage, all the way down to the specific tag being used. I postulate this because my secret word being tested is imaginary and has no context or definition and isn't used anywhere else on the internet yet google still finds a way to sort the test pages consistently and somehow slowly change the sort logic over time. The only difference in many experiment cases are the HTML tag names used. If google truly believed all tags are equal then my results wouldn't be reproducible over time and across the two other similar experiments I have access to, which use different secret words and W3C HTML Tags.
The tags do not appear in alphabetical order and changes would not occur the way they do.
The order would change wildly and frequently if cross system hashcodes were in play. Hashcodes values do not remain consistent from server to server and runtime to runtime. I would expect randomness from hashcode order. This does not happen.
The order is not random.
The test is reproducible.
I can only conclude that how they determine relevance and how they define web spam ultimately yields a base weight effect of the semantic tags being used. It might not be intended or purposeful, but I think the order is an effect of their methods. I make no claims to intended or unintended by Google. I don't care about Google's intentions or claims. I care about their actuals.
I understand that my observations & opinions are inconsistent with claims made by Google and it's patents and that for reason of practicality there is only so far I can go in proving my case. I don't think Google or it's patents are being completely open and transparent about how they operate in actuality.
Given Google's claims, Google's patent's claims, and field observations. I am more inclined to consider field observations over the others. This is because in "my experiences with" and "my opinions of" Google it has a history and reputation of lies of omission and waging plausible deniability campaigns regarding their algorithms and rules. I would not put it past them to file decoy patents or make false statements regarding their methods.
Matt Cutts is Google's "Cancer Man" of X-files fame in my opinion. The official smoke and mirrors man who gets the cover up jobs to keep the world in the dark and to keep everyone guessing.
If it works for me and IF it turns out to work for others, I'm not going to let Google or academics tell me it doesn't work. That isn't scientific method either.
Tag order is one part of my system. My system works well enough that I actually project my SEO revenue growth into the future and continue to meet or exceed it. My SEO projections are actually a part of the company books and the results aren't random.
My system is built upon looking at my keyword space macroscopically and microscopically and building systems based on the observations and trends I'm seeing in the data.
My theories are often wrong, but my results are surprisingly stable, effective, and recurring. For whatever explanation there may ultimately be. Acting on empirical observations in SEO is a formula for success.
I know I am still not align with the Math and the claims. The only other way I can put it is that the math and the claims don't explain the observations.
Either the observations are bad or there is something happening at the semantic level of the source over at Google. I'm still inclined to believe the later as silly as it sounds because I have not encountered an even marginally plausible theory that produces these results that is to the contrary. The rub is that it isn't a theory if it can't be tested and there are a whole lot of practical barriers to what can be tested in terms of SEO insights.
So limited to experiments that can be performed. How can I test to the contrary. I'm not opposed. I just don't know how best to do that.