Mathman, when you break the experiment down like that, it looks less confusing... but infinitely more crazy.
If they're going to split the judges, just have some judges specialize in GOE and some specialize in PCS.
They tried that as an experiment at Nebelhorn a number of years ago but decided not to adopt that split. What I heard was that the judges who were judging GOEs-only were bored.
It looks like this breakdown is an attempt to give at least some of them something else to do as well. And to give the PCS-only judges less to do so they can be more analytical about the components that they are judging, plus they won't all be tied to the Skating Skills mark since some judges aren't judging Skating Skills at all.
There's no guarantee this experiment will be deemed successful and its approach officially adopted. And if it is, the breakdown of who does what would have to be flexible since not all competitions would be able to bring in 12 judges plus a tech panel.
I think we are on the right track, I'm looking forward to see if it works well or the judges'll find it too complicated even if I still believe that 1 tech specialist, 6 GOEs' judges and 6 PCS' judges is the easiest and best way.
This could be something like what Mathman suggested above.
The tech specialist could be responsible for identifying how many elements were executed and getting the codes for each one input into the computer. If they make a calling mistake, the data entry operator (if there is one) or the referee or any of the judges could alert them (did you really mean toe loop? it looked like a flip to us!) to fix it at the end of the program.
Element judges could identify jump errors that reduce the base value (but might not agree, so with this year's rules and five judges, the base value for a flip or lutz with questionable rotation and questionable takeoff edge could be different for every judge).
These judges would also subtract quality/GOE reductions for other errors and add points for good quality and for difficulty features. How those pluses and minuses would be displayed on the protocol would need to be determined.
Or keep the levels and 3-person tech panel system already in existence, but split the judges so that some ("technical judges") evaluate elements and Skating Skills and Transitions, and others ("performance judges") judge Performance Execution, Choreography, and Interpretation, with separate appointments and training for those two different roles, although any individual is welcome to pursue both.
Actually, one thing I would like to see different is the PCS range. While I can clearly identify a -3 from a -2 or a +1 from a +2 on a technical element, I can't understand the difference between a 6.25 and a 6.50 or 7.75 from a 8.25 skater. I've always found this range too big, with too many options inside that honestly don't give you the real idea of a skater. A 1 to 5 range would be more immediate to understand (for example 1 is poor on that element, 2 is sufficient, 3 is decent, 4 is good, 5 is majestic). Or if they want to diversify more the skaters they can still use 1 to 10 but without decimals.
It's a dilemma. The problem that any scoring system must deal with is that a single scale must be able to accommodate skaters at all levels from beginners to world champions.
Yup.
When starting out learning IJS, it's best to focus on the whole numbers. See the
Program Components Overview (linked at the bottom of the page).
The numbers are defined as
Outstanding 9-10
Very Good 8
Good 7
Above Average 6
Average 5
Fair 4
Weak 3
Poor 2
Very Poor 1
Extremely Poor <1
For Skating Skills especially, these correlate with technical skill levels. I think of 10 as being a great performance by an all-time great skater, 0-1 as being a beginner with very little one-foot skating or identifiable edges. So what is Average/5? Seeing how judges have been using the numbers, I think of it as acceptable senior quality, nothing special in senior competition -- but a strong score at lower levels, more exceptional the lower you go.
If 5 is basic senior quality, then scores lower than that will be more common at lower levels, and scores higher than that will be more common at the international competitions. Judges who are experienced at judging all levels will quickly be able to peg skaters to a general range (e.g., 5s or 7s) on the full scale.
Fans who only watch seniors, or only elite seniors, will see a narrower range of skills and might think of 5 as a very low score. But watching a lot and analyzing the criteria for the various components could allow even fans who don't know much about skating technique to predict whether a skater they've never seen before (i.e., no reputation judging) will likely earn 6s or 7s or 8s.
The decimal places allow for finer distinctions among skaters who are basically in the same skill range. That's where judges might start thinking comparatively between skaters in the same event, even though strictly speaking they're not supposed to. They also allow judges to balance out the various criteria on the same component, in case a skater is notably stronger at some criteria and weaker at others.
At that level of distinction, there aren't really right or wrong answers.
The other components don't need to be directly tied to the Skating Skills skill level, although some of the criteria (e.g., Difficulty and Quality of Transitions, Carriage under Performance/Execution, Pattern and ice coverage under Choreography, Effortless movement under Interpretation) do rely at least to some degree on control of the technique.
So each judge needs to develop a mental standard of what is "Weak" or "Average" or "Very Good" performance or interpretation, across the full range of skaters from beginners to world champions. Here's where I think more detailed guidelines and training would be useful. But judges do develop a consensus of what they consider average, etc., by judging with each other, reading protocols to compare their marks to the whole panel, discussing in the judges' room afterward what they liked and didn't like.
Fans can develop a sense of above-average performance or very good interpretation too, at least at the whole number level. Since these don't rely so much on technical skating knowledge, fans' evaluations in these could be just as valid as judges', especially for fans with performing arts backgrounds. But fitting their evaluations to the 10-point scale means understanding the range of the scale, having a sense of what to expect from non-elite skaters as well as the elites.
6.0 had the same challenge, hence the need for decimals. If you didn't have enough divisions then every elite skater would automatically deserve the highest mark, in comparison to all the skaters in the world.
Yes.
The other reason 6.0 needed decimals, along with tiebreakers, was to have enough room to rank the skaters in large fields. If you only have 10 numbers available for each mark, you would have to use the full range of numbers for every competition regardless of the skill level of the skaters. There would not even be a rough correspondence between scores and skill level -- they would be nothing but place holders. And judges still might run out of numbers/get "boxed in" pretty quickly. Deductions (as in short programs) could not be taken from the actual scores but just considered mentally by each judge when deciding which placeholder scores to put up to rank the skater appropriately.
(With compulsories, which only received one mark for each figure or for each dance until the early 1990s, it would be impossible to distinguish 30 skaters with only 10 possible numbers.)
1 technical specialist, 6 GOE judges and 6 PCS judges would be the best option for me, since there is another problem: training the judges, since "normal" judges (I think) are not trained in all the small details required to assign level/e or UR calls, so creating "technical judges" who evaluate both levels and GOEs would cost a lot...
Yes, any major reassignment of duties would require more training.
If the tech panel as it now exists were to be abolished and individual judges were assigned to independently score difficulty along with quality for each element, I think it might make sense to get rid of the current definition of levels and just allow element judges to give extra points for extra difficulty as well as for extra quality. But they would still need to be retrained.
I think the only "automatic" mark is SS, which is obvious higher for elite skater than beginners and you don't need to watch the skater closely to evaluate it.
I don't think you need to watch the skater closely to get a sense of what general range they belong in. But two skaters with similar power and edge depth, for example, may show different levels of mastery of multidirectional skating and one-foot skating. So paying close attention allows judges to reward skaters who actually show more skills even if the first impression is similar.
But sometimes it happens that, for example, a junior can give a better interpretation than a senior skater. So I don't see any problem if the first get a 7 and the second one get a 5, I don't think it has to do with their category or level.
Yup.
Maybe a 1 to 10 scale is appropriate and perfectly understandable when 6 is the sufficiency, I can also accept a 0.5 mark... but having so many decimals for PCS... it's not like in school when you get a mark based on the number of errors, how can you say a performance was 0.25 better than another one?
As I mentioned, strictly speaking judges aren't supposed to be thinking "I already gave Skater P 6.5 on this component, and Skater Q was a little better, so I'll give her 6.75." Although I wouldn't be surprised if some do think that way.
What it's really for is for a judge to say to himself something like "Skater Q was better on this component than just Above Average, but not quite Good. Halfway in between? No, I'd say Q was closer to Good in my mind, almost there, but those couple little problems/weaknesses won't let me go all the way to 7 on this score."
Imagine that you're judging numerous skaters who are all very close in overall ability on this component. How much room should be available to reflect slight differences? If they're all pretty much average and nothing special, should they all earn 5.0? Or can the ones who have average skill and are having a good day earn 5.5 or even 6.0, and the ones having a bad day earn 4.5 or even 4.0?