Forums

Bloodstock & Breeding

Welcome to Live View – Take the tour to learn more
Start Tour
There is currently 1 person viewing this thread.
Airmail Special Gr3
24 Sep 11 01:00
Joined:
Date Joined: 17 Nov 08
| Topic/replies: 170 | Blogger: Airmail Special Gr3's blog
Editor=s note: Bill Oppenheim delivered this speech at
the Pedigree/Genetics/Performance Conference in
Lexington, KY, September 7, and it is being reprinted
here with the permission of the organizers.
From TDN 21st September.

Most of you know me as a columnist for the
Thoroughbred Daily News, so I'd like to briefly outline
my credentials to speak here today. When people ask
me what I do, I tell them, "I'm an analyst in the
breeding and auction sales sector of the Thoroughbred
horse industry." The usual reaction to that is, "say
what?"
As an analyst, one of my jobs is a journalist, and
that's how I'm best known. But another of my jobs is as
a pedigree consultant, advising on matings and
purchases at sales. I've also done this job for about 30
years, ever since John Ed Anthony hired me to help him
select mares to go to Cox's Ridge when that horse
retired to Claiborne Farm, I think in 1980. I would
advise on perhaps 200-300 matings a year and have
been fortunate to have been involved with three major
operations over that time, which means the advisor is
working with at least some top mares and sires, which
of course enhances your chances. I've been involved in
advising on the matings of 13 Classic winners since
1992, including two winners of the Epsom Derby, three
consecutive European champion 2-year-olds, two Arc
winners, and the winners of the Epsom Oaks in 2007,
2008, and 2010. I report this simply to say this is a
field I have worked in for 30 years with some success,
so I do bring some practical experience to the table.
I also want to thank two people who have been
instrumental in putting together this presentation (I'm
very low-tech, myself): my now long-standing (I almost
said long-suffering) associate of five years, Brianne
Stanley; and Dr. Emily Plant, an Assistant Professor at
the University of Montana, who reviewed, critiqued, and
helped us assemble the presentation.
SECTION 1. THE CONDITIONS WE WORK WITHIN
There are a few things I really like about how the
Thoroughbred business, especially as what I call the
breeding and auction sales sector of it (as distinct from
the racing sector) has evolved. I like that it is a truly
international business, a 'business without borders;'
sure, there are a lot of wholly national characteristics of
the industry in each country, but when we talk about,
say, bloodstock agents who travel to and buy at the
major international auctions, we mean Kentucky,
Newmarket, Goffs in Ireland, Deauville, at least, but the
buyers are coming from Australia, Japan, the Middle
East, South America, South Africa, Russia, India, Korea,
Turkey, and many other countries and regions as well.
In a very important respect, the 'breeding and sales
sector,', anyways is a business without borders.
Second, we are fortunate to be working in an
essentially unregulated 'free market' economy. Of
course, there are rules and regulations, but there are not
artificial, externally imposed barriers to, or conditions of,
trade. It makes it more feasible to analyze how certain
principles operate, such as one of the most important--
the law of supply and demand. Also, this has become a
big-money business. Thoroughbred horses can be worth
really big money, so much so that it has created its own
economy. This is because Thoroughbreds are not just
racehorses and breeding stock; Thoroughbred horses are
also financial instruments, like, say, commodity options,
or derivatives. The difference is these financial
instruments are live animals, too.
Third, at the end of the day-- I mean, at the end of
the race--there is a finish line, and a result. Horse racing
itself has to be a meritocracy--the best horse wins. As
an aside, that's what I personally dislike about handicap
racing, because the intention is for the best horse not to
win, but, fortunately, in the really important races from
the long-term perspective of the breeding standpoint,
the best horse wins. And, fortunately for many of us, in
the last 30 years, the business itself has become much
more of a meritocracy. If you have the best idea, you
will be heard, maybe even funded. And the wealth in
the business is spread out much more democratically; if
you come up with the right horse, you will get paid, no
matter who you are, because the best horse wins.
That, as I see it, is the essential landscape we are
working in: these are the conditions in which we all
operate--and I, for one, am grateful that we can work in
a business, a real business, too, in which you don't
have to put a tie on to go to work every day. Being a
big-money business, there is a fair amount of greed and
chicanery to be found, but there is also a very serious
demand for advice and assistance in the creation or
discovery of the superior Thoroughbred. That's where
we all--the speakers at this symposium--come in.
SECTION 2. WHAT WE DO
So far today you've listened to a fair bit of science,
but this afternoon I'd like to change the focus a little bit,
to what we might call more traditional methods of
looking at pedigrees, and deciding on matings. When I
refer, in the title of my talk, to "the limitations of the
data that we use," I'm talking more about us boffins
with our books than the data the scientists use; I'll have
to leave them to comment on their data, that issue isn't
one I'm really competent to comment on.
There is one assumption we share, though, and I
emphasize the word "assumption" because I think it's
vital that we understand the assumptions from which
our research and conclusions emanate. And the
assumption that scientists and more traditional pedigree
students share is this: we want to replicate the
pedigrees of good horses. Students of DNA will look to
replicate the patterns they find in top horses; pedigree
students looks for patterns in the pedigrees of good
horses. But we're all working from the same
assumption: that which has happened before is more
likely to happen again than that which has not happened
before; we're trying to replicate past successes,
however we define those successes. The $64-million
dollar question is: what is it that we're trying to
replicate; what has succeeded before?
Let me cite three examples which I think challenge
some of the conventional wisdom which prevails among
pedigree students today.
First there is the APitfall of the Small Sample Size,@ in
which a few early successes lead breeders to flock to
the hot new cross, only to discover the law of averages
tends to even things out the more occurrences there
are. The greatest example I have ever seen of this was
the case of Storm Bird over Secretariat mares. Storm
Bird has two named foals out of Secretariat mares in his
first crop, in 1983, and they were both graded-stakes
winners: Grade I winner Storm Cat, and a Grade
III-winning filly Storm Star. The crops weren't so big
then as they are now, as you know, and by the 1987
crop, there had been only seven more named foals by
Storm Bird out of Secretariat mares. One of the four
foals of 1987, though, was Summer Squall; in 1991
there were 5 named foals bred on this cross, and in the
next four years there were 9, 9, 16, and 12,
respectively.
If you'd taken a snapshot of Storm Bird over
Secretariat after his first eight crops--and evidently
plenty did--you'd have made the staggering discovery
that the cross of Storm Bird out of Secretariat mares
had thus far produced 36 percent stakes winners, and
29% graded-stakes winners. What would you have
done if you had a Secretariat mare?
Exactly; in his last 10 crops, Storm Bird had 60
named foals out of Secretariat mares. They included one
black-type winner, and no graded SW.
First 8 crops:36%; last 10 crops, 2%.
Why? Did the cross suddenly become less potent?
No. Breeders were fooled by a few good early results
from a sample size which was of no statistical
significance. Anything could have happened. When it
was all totaled up, Storm Bird sired 9% black-type
winners from foals, and 5% graded SW from foals. Out
of Secretariat mares, his career totals were 8% and
5%, respectively; so really, it was just an average cross
for the sire, which happened to have all its stars arrive
in the first half of his career.
So, beware of the small sample size, I say. Here's
another nugget of conventional wisdom which the facts
disprove, and which actually applies to the Storm Bird--
Secretariat example as well. Conventional wisdom
believes when you find a winning formula, a pattern
that works, you keep repeating it; but, I maintain, as
soon as we know something works, it is outdated; it is
on the way out. Here's why: there is a four- or five-year
lag between when we decide on a mating and the
resulting produce gets to the racetrack. In that time
there can be significant shifts in the population; certain
lines become numerically stronger or weaker within the
population. A line which has become stronger was
probably under-represented proportionally five years
ago, and vice versa. When you did a mating five years
ago, you were working with a different population.
Here's another example of a cross which was
outdated the minute it was discovered: Sadler's Wells
over Darshaan. There were no foals by Sadler's Wells
out of Darshaan mares until 1992; both had their first
foals the same year, 1986. There were two graded SWs
bred on that cross in 1997 from his 1994 crop, Crimson
Tide and Ebadiyla. From the 1998 crop came three
black-type winners from five foals out of Darshaan
mares, including 2001 St. Leger winner Milan; and out
of the eight named foals by Sadler's Wells out of
Darshaan mares in the 1999 crop came High Chaparral,
Islington, and Quarter Moon in 2002.
From the 12 crops foaled from 1992-2003, there
were a total of 66 foals bred on this cross. Seventeen
of them (27.5%) were black-type winners, 12 of those
(18.2%) group winners.
Once it was obvious this was a very powerful cross
(by 2003), the numbers jumped up, and the strike-rate
declined. In his five crops foaled between 2004-2008,
there were a further 97 named foals bred on this cross.
The numbers weren't terrible--11.3% black-type
winners, 8.2% group winners--but the cross worked
less than half as well after it was clear this was a
sensational cross than it had before.
These are myths, these conventional wisdoms that
we live by. It is a myth that a small number of good
horses from an insignificant sample is a reliable
predictor. It is a myth that once we discover a great
cross, it will continue to be just as great a cross as it
was before we discovered it. These myths crash on the
rocks of the laws of probability; in these cases, the
greater the numbers, the less the success.
The third myth I should like to address has to do with
the question: "what class does a horse need to be
before we want to replicate its pedigree?" This is an
absolutely essential question. The further you go down
the class scale, the more 'random' the pedigrees are.
Conventional wisdom is that the pedigrees of all blacktype
winners are worth collecting and studying. Jack
Werk used all black-type winners when he created his
database--I believe that is correct--and True Nicks has
followed suit, introducing the supposed 'improvement'
of being able to measure opportunity. I believe I am
right about that, but please correct me if I'm wrong
[correction from the audience: he did not use restricted
black-type winners]. But it sounds quite sensible,
doesn't it? All black-type winners.
The trouble is, according to my research, all blacktype
ain't equal. We have plenty of anecdotal evidence
of this in the number of Europeans who come to the
American sales and stare blankly at the catalogue
pages, trying to figure out what in the hell does all this
black-type really mean? There's a reason why.
In Europe, there are two kinds of black-type races:
group races, and listed races. But in North America
there are three kinds: graded races--same as group;
listed races; and 'other' black-type races--I think, now,
these are any black-type races with a gross below
$75,000. Yes, the percentage of black-type races in
Europe is higher than in North America--though, in
2008, the number of races run in Britain, Ireland, and
France combined was less than a quarter of the number
of races run in North America.
In any case, it sort of works. I think it would make a
lot more sense to call listed races Grade 4, or Group 4,
because that's really what they are. But here's a vital
statistic: in 2008, there were 1,890 black-type races in
North America, according to The Jockey Club Fact
Book. Of those, 742, or 39%, were listed or graded
races, and 1,148, or 61%, were not. In Britain, Ireland,
and France, in the same year, there were 630 blacktype
races; none were below listed class. 61%; 0%:
something's out of whack.
As some of you may know, I use a somewhat
different way of identifying class, which is earningsbased,
and identifies a group called 'A Runners,= which
are the top 2% of earners in designated countries each
year. I mention this only to say I once did a study which
told me the percentage of A Runners in the population is
very close to the percentage of listed plus graded and
group winners.
The point is, I have learned, using my method, that
only the pedigrees of A Runners are worth studying; the
pedigrees of B Runners, C Runners, everything else, are
increasingly random. So I infer that black-type winners
below listed standard are the equivalent of B Runners
and C Runners in my system: not worth studying.
My conclusion: if we look at the graph here on Pitfall
#3, we can see that, in 2008, 45% of the black-type
winners in the five countries combined were below
listed class; in other words, in my opinion, nearly half
these databases are junked up with horses which
shouldn't be there. I'd say it's pretty likely the
conclusions could be riddled with 'false positives.'
Byron Rogers, one of the organizers of this
conference, is a very smart guy--one of the smartest
I've met in this business. One of the big selling points of
TrueNicks over Jack Werk's material, I remember him
telling me, was their ability to "measure opportunity"--
not just has something worked (given my caveat about
both groups' databases), but how many times has it
really worked compared to how many times it's been
tried?
It sounds good on paper, but the more I've thought
about it, the less I like it. There's eminent danger of the
small sample size problem in almost any case--oh, this
has been tried 16 times, this has been tried 24 times,
and so on. When you say, well, something's better than
nothing, you have to understand, in these cases, no, it's
not.
Somewhere between 2 and 3% of horses, depending
on the study and the class parameters used, are actually
good horses, horses the pedigrees of which are worth
trying to replicate. Let's call it 2.5%--one out of 40. So
39 out of 40 fail to make the cut. With such an
incredibly high failure rate, I really have to question
whether measuring opportunity can be truly significant.
Most horses fail--the vast majority of horses fail. If
something works three out of 40 times, and something
else works one out of 40 times, does that mean we
should we doing the three-out-of-40 every time? Maybe
not.
Then, if you want to go one step up in the pedigree;
in other words, not sire on damsire, but something like
grandsire and sons on damsire, and so forth, another
problem arises: where do you make the break? Is there
such a thing as a sire line, and if so, where do you
break it? Ribot was always Ribot, even if it is Albertus
Maximus, by Albert the Great, by Go for Gin, by
Cormorant, by His Majesty, by Ribot; five generations
back. The identity and classification of sire lines is
subjective, not objective; everybody is eligible to use
something different.
Measuring opportunity could be significant if you were
looking at enough cases; but how should they be
properly grouped? There's no agreement on this. All
taken together--I used to be quite interested in
measuring opportunity. I'm not at all any more; I'm
really much more interested in accurately identifying
success.
SECTION 3. THE IMPLICATIONS OF UNCERTAINTY
Every time a breeder signs on the dotted line to breed
his or her mare to Distorted Humor, or Unbridled's
Song, or Giant's Causeway--and a few others--it's near
enough a $100,000 decision
. It should make a
difference to those of us who work in the field of
evaluating and recommending pedigrees, where
percentage plays are so important, that we are speaking
from a vantage point of uncertainty, not of certainty.
Even from the point of view of the scientists, we're
entitled to ask "how do you even know what to
measure?" and "how can we know that what you do
measure is significant?" Whatever we do, we're taking
a shot here: the odds in any given case are against
success.
It bothers me when I hear someone assert "this or
that is absolutely the way to go;" it bothers me when I
say it myself. At best it is absolutely the way to go,
given the circumstances. I try not to be too militant,
above all to be open-minded, to use inductive as well as
deductive reason--let the evidence tell you, not the
other way around.
An example is the application of dosage to matings
planning. Dr. Roman, who will be speaking tomorrow,
will be able to explain it much better than me. I don't
have a problem with the theory of dosage, but with the
application of it. Here's what I mean. As I understand it,
dosage is a sort of pedigree map, showing the important
names in pedigrees and the aptitudes of those named
sires. A calculation is then done which shows the
distribution of influences in the existing pedigree, from
which the idea is that you can predict aptitude, and add
certain elements to balance out the pedigree; in some
way to produce a good horse.
It seems to me the problem here is an assumption
the assumption that the individual horse on the ground,
or the prospective individual horse on the ground will
resemble the designed pedigree. My experience is that
most horses throw back to certain influences in their
pedigrees, except, given the genetic law that
populations tend to breed back to the middle, you might
find a horse "throws back" to a certain sire, but isn't as
good.
If horses throw back to particular influences in their
pedigrees, surely a map of a horse's dosage tells us
simply about the opportunity, not the fact, of influence.
Like I say, my issue is not with the theory of dosage,
but the application of it: the pedigree may be that of a
two-mile stayer, but if the horse on the ground looks
like a sprinter and stops after five furlongs, I'd be
looking for the names in that pedigree of five-furlong
horses, and start working from the assumption that's
the horse we're really dealing with.
There are two corollaries to approaching pedigree
assessment with an open mind: one is the recognition
that, even in science, the development of hypotheses
requires a creative 'spark'--somebody has to think up
the right question to ask, the right hypothesis to test.
The other corollary is what is called to "think outside
the box.@ Let me give you an example that utilizes both.
At the end of 1985, I was hired to make
recommendations for the American mares of one of the
major international operations. There was a particular
mare, a Halo mare, who I was keen to see go to Mr.
Prospector. His first foals had raced in 1978, but until
then--1985--he had yet to sire a good horse out of a
mare from the Hail to Reason sire line. But, expanding
the parameters one level, I knew that Mr. Prospector's
sire, Raise a Native--and, in fact, virtually all Native
Dancer--really liked Hail to Reason-line mares. So I
explained the situation, explained why I thought it
would still be worth trying, and they rolled the dice.
They got Machiavellian. From then on, I became a lot
less rigid about insisting that this sire had to have
worked with that damsire.
The point I'm ultimately driving at is that we are
operating in a universe, this universe of Thoroughbred
racing and breeding, which is more chaotic than
ordered. There is a natural tendency to want things to
make sense and be in order. But that path leads to
overrating the true, applicable value of the data we are
using and the conclusions we reach from that data. It's
the search for the one great formula, or set of formulas:
the Philosopher's Stone, the Magic Wand, the Secret
Elixir. I don't think they exist, not in this pursuit,
breeding racehorses.

Then what?
Then, we have to recognize--to concede--that the
pedigree work is only part of the equation; it's
dangerous to make all this that we discuss too big a
factor in the decision-making process; at the end of the
day, it's an individual bred to an individual. Let me just
reiterate the reservations I have about the assumptions
which result in conventional wisdom:

1. Jumping to conclusions based on too small a
sample size;

2. By the time we discover a particular cross is really
good, it is losing its potency;

3. The assumption that all black-type should count the
same as an indicator of class is flawed;

4. The value of measuring opportunity is open to
question;

5. A pedigree is a map of opportunity, not a statement
of fact. Horses are not blends;

6. Generalizing 'up one level' is often more productive
than trying to be too specific;

A few final thoughts: the first thing we have to do is:
eliminate obvious errors. Then, get ahead of the curve;
and do art as well as science. Finally--it's blindingly
obvious that the people who use the services we're all
working to develop want the silver bullet, the holy grail,
the magic formula. Yes, let's use all the tools we have
and can develop, but let's also try very hard to make
them understand, as we should, the limitations of what
we are doing. Thank you.
Pause Switch to Standard View THE MYTHS WE LIVE BY: THE DATA WE USE...
Show More
Loading...
Report potentialmillionaire September 26, 2011 8:21 PM BST
Thanks for that Airmail.

I always enjoy Bill Oppenheim articles. Knowledgable reality!
Report jmc27 September 27, 2011 5:59 PM BST
Interesting read!
Post Your Reply
<CTRL+Enter> to submit
Please login to post a reply.

Wonder

Instance ID: 13539
www.betfair.com