February 2016 - Heavy Topspin

Two New Ways to Chart Tennis Matches

Readers of this site are probably already aware of the Match Charting Project, my effort to coordinate volunteer contributions to build a massive shot-by-shot database of professional tennis. If this is the first you’ve heard of it, I encourage you to check out the detailed match- and player-level data we’ve gathered already.

In the last week, two developers have released GUIs to make charting easier and more engaging. When I first started the project, I put together an excel spreadsheet that tracks all the user input and keeps score. I’ve used that spreadsheet for the hundreds of matches I’ve charted, but I recognize that it’s not the most intuitive system for some people.

The first new interface is thanks to Stephanie Kovalchik, who writes the tennis blog On the T. (And who has contributed to the MCP in the past.) Her GUI is entirely click-based, which means you don’t have to learn the various letter- and number-codes that are required for the traditional MCP spreadsheet.

While it’s web-based, it has some of the look and feel of a modern handheld app. It’s probably the easiest way to get started contributing to the project.

(Which reminds me, Brian Hrebec wrote an Android app for the project almost two years ago, and I haven’t given it the attention it deserves. It also makes getting started relatively easy, especially if you’d like to chart on an Android device. [Update, December 2019: Unfortunately, it appears this app is no longer available.])

[Update, December 2019: The other GUI referred to in the title has bugs that render the output unusable. I recommend sticking with the standard MCP spreadsheet, or with Stephanie’s GUI described above.]

With four ways to chart matches and add to the Match Charting Project database, there are even fewer excuses not to contribute. If you’re still not convinced, I have even more reasons for you to consider. And if you’re ready to jump in, just click over to one of the new GUIs, or click here for my Quick Start guide.

New at TennisAbstract: Weekly Elo Reports

Starting today, you can find weekly Elo ranking reports on the home page of Tennis Abstract. Here are the men’s ratings, and here are the women’s ratings.

Elo is a rating system originally designed for chess, and now used across a wide range of sports. It awards points based on who you beat, not when you beat them. That’s in direct contrast to the official ATP and WTA ranking systems, which award points based on tournament and round, regardless of whether you play a qualifier or the number one player in the world.

As such, there are some notable differences between Elo-based rankings and the official lists. In addition to some rearrangement in the top ten, ATP Elo ratings place last week’s champion Roberto Bautista Agut up at #12 (compared to #17 in the official ranking) and Jack Sock at #13 (instead of #23).

The shuffling is even more dramatic on the women’s side. Belinda Bencic, still outside the top ten in the official WTA ranking, is up to #5 by Elo. After her Fed Cup heroics last weekend, Bencic is a single Elo point away from drawing equal with #4 Angelique Kerber.

These new Elo reports also show peaks for every player. That way, you can see how close each player is to his or her career best. You can also spot which players–like Bencic and Bautista Agut–are currently at their peak.

Like any rating system, Elo isn’t perfect. In this simple form, it doesn’t consider surface at all. I haven’t factored Challenger, ITF, or qualifying results into these calculations, either. Elo also doesn’t make any adjustments when a player misses considerable time to injury; a player just re-assumes his or her old rating when they return.

That said, Elo is a more reliable way of comparing players and predicting match outcomes than the official ranking system. And now, you can check in on each player’s rating every week.

What Happens After an Unsuccessful First Serve Challenge?

Italian translation at settesei.it

A lot of first serves miss, so every player has a well-established routine between the first and second serve. So much so that, traditionally, if something disrupts that routine, the receiver may grant the server another first serve.

Hawkeye has changed all that. If the server doubts the line call, he or she may challenge it. That results in a lengthy wait, usually some crowd noise, and a general wreckage of that between-serves routine.

The conventional wisdom seems to be that the long pause is harmful to the server: that if the challenge fails, the server is less likely to put the second serve in the box. And if the second serve does go in, it’s weaker than average, so the server is less likely to win the point.

My analysis of over 200 first-serve challenges casts doubt on the conventional wisdom. It’s another triumph for the null hypothesis, the only force in tennis as dominant as Novak Djokovic.

As I’ve charted matches for the Match Charting Project, I’ve noted each challenge, the type of challenge, and whether it was successful. I’ve accumulated 116 ATP and 89 WTA instances in which a player unsuccessfully challenged the call on his own first serve. For each of these challenges, I also calculated some match-level stats for that server: how often s/he made the second serve, and how often s/he won second serve points.

Of the 116 unsuccessful ATP challenges, players made 106 of their second serves. Based on their overall rates in those matches, we’d expect them to make 106.6 of them. They won exactly half–58–of those points, and their performance in those matches suggests that they “should” have won 58.2 of them.

In other words, players are recovering from the disruption and performing almost exactly as they normally do.

For WTAers, it’s a similar story. Players made 77 of their 89 second serves. If they landed second serves at the same rate they did in the rest of those matches, they’d have made 77.1. They won 38 of the 89 points, compared to an expected 40 points. That last difference, of five percent, is the only one that is more than a rounding error. Even if the effect is real–which is doubtful, given the conflicting ATP number and the small sample size–it’s a small one.

Of course, the potential benefit of challenging the call on your first serve is big: If you’re right, you either win the point or get another first serve. Of the challenges I’ve tracked, men were successful 38% of the time on their first serves, and women were right 32% of the time.

There’s no evidence here that players are harmed by appealing to Hawkeye on their own first serves. Apart from the small risk of running out of challenges, it’s all upside. Tennis pros adore routine, but in this case, they perform just as well when the routine is disrupted.

First and Second Serves: Another ATP Info-miss

Breaking news, everybody: First serves are better than second serves!

That’s what I learned, anyway, from the latest article in the “Infosys ATP Beyond the Numbers” series:

When you average out the Top 10 players in the 2015 season, they are saving break points 72 per cent of the time when making a first serve. On average, that drops to 53 per cent with second serves. That 19 per cent difference is one of the most important, hidden metrics in our sport.

Is the difference between first and second serves “important?” Definitely. Is it in any way “hidden?” Not so much.

The melodramatic phrasing here suggests that break points are different from regular points, perhaps with a much larger spread between first and second serve winning percentages. But no, that’s not the case.

Last year, top ten players won 75.6% of first-serve points and 55.4% of second-serve points. Combined with the Infosys numbers–which I can’t verify, because the ATP doesn’t make the necessary raw data available–that means that top ten players win 5% less often when making a first serve on break point, and 5% less often when missing their first serve on break point.

At the risk of belaboring this: When it comes to the importance of making your first serve, break points are no different than other points.

Even that 5% difference is less meaningful that it looks. Break points don’t occur at random–better opponents generate more break opportunities. If you play two matches, one against Novak Djokovic and one against Jerzy Janowicz, you’re likely to face far more break points against Novak than against Jerzy … and of course, you’re less likely to win them.

Pundits tend to focus on break points, and in part, they are right to do so, because this small subset of points have an outsized effect on match outcomes. However, because of the small sample, it’s easy–and far too common–to read too much into break point results. My research has repeatedly shown that, once you control for opponent quality, most players win break points about as often as they do non-break points.

The ATP is sitting on a wealth of information. If we’re going to learn anything meaningful when they go “beyond the numbers,” it would be nice if they took advantage of more of their data and offered up more sophisticated analysis.

Match Charting Project February Update

At the beginning of the year, I announced an ambitious goal: to double the number of matches in the Match Charting Project dataset. That’s a target of 1,617 new matches in 2016–about 135 per month, or 4.5 per day.

So far, so good! In January, ten contributors combined to add 162 new matches to the total. Our biggest heroes were Edo, with 35 matches, including many Grand Slam finals; Isaac, with 33; and Edged, whose 22 included some of the dramatic late-round men’s matches from Melbourne.

As we close in on the 1,800-match mark, I’m excited to announce a new addition to the stats and reports available on Tennis Abstract. Now, for every player with at least two charted matches in the database, there’s a dedicated player page with hundreds of aggregate data points for that player.

Here’s Novak Djokovic’s page, and here’s Angelique Kerber’s. I’m still working on integrating these pages into the rest of Tennis Abstract, but for now, you’ll be able to access them by clicking on the match totals next to every player’s name on the Match Charting home page.

These pages each feature four charts, which compare the player’s typical rally length, shot selection, winner types, and unforced error types to tour average. The other links on each page take you to tables very similar to those on the MCP match reports. Move your cursor over any rate to see the relevant tour average, as well as that player’s rates on each surface.

I hope you like this new addition, which owes so much to the amazing efforts of so many volunteer charters.

I hope, too, that you’ll be inspired to contribute to the project as well. When you’re ready to try your hand at charting, start here. As always, the more matches we have, the more valuable the project becomes.