Actually, yes. We have been working with the music tech program at our local college for some time (we conduct our loudspeaker tests on campus), and we've been discussing a variety of ways to incorporate some of their music tech students into the BGM review process. Setting up some double blind testing is one of the ideas on the table. We have a lot of cool things planned, but it's a lot of work doing the level of testing that we are currently doing, and it's not all that easy to find the time to work more things in (and to do so in an intelligent, "connected" fashion which will work well in the context of a review).
That being said, I do hope to incorporate some group double blind testing in over the next year or so.
I wouldn't bother with this myself. First, double blind is probably not appropriate (or even possible) for these types of questions. First, you guys (the testers) aren't biased to make one head or cab or whatever win. And, there would need to be much interaction with the head by whoever is turning the knobs anyway, since testing a head at only one EQ setting would give a very limited and possible misleading result, since a different mix of non-linearly related components could give you a very different result.
Also, since any test of a bass, amplifier, cab or whatever is an interactive and highly non-linear system, trying to isolate one simple linear impact (i.e., maximum volume in this example) is a losing game. Depending on tone dialed in, interaction of voicing with cabinet, the player (technique, etc.), the room, etc., etc., you can get quite different result, even if you have a single blind rating panel with a reasonable number of members. Even holding all this constant (one player, one room, one cab, one bass, etc., etc.) only results in a very limited and very possibly misleading result.
Possibly, doing a 'difference test' on some subtle issues like cables or whatever might be interesting, but even then, the signal source (strings, player, tone, attack, etc.) can interact with a cable or whatever to skew the results.
Much better to rely on a pattern of a large number of real world and seemingly relatively unbiased TB reviews (if you remove endorser responses, etc.) for that type of information IMO.