RSS Amplifier

The Visual In the Noise · Mar 18, 2021

No data? No problem

0
Sign in to vote or save

JP Hwang · The Visual In the Noise

Hi!

The last few weeks/months have been pretty hectic; I’d been working on a few really exciting projects. I can’t share too much yet, but I did have a few adjacent thoughts. Since this is a data and visualisation blog, let’s start with a few (moving) pictures.

Can you tell what this is?

What on earth is this, you ask?

It’s a skeleton extracted from a YouTube video of a ballet dancer. (Extra kudos to you if you spotted it immediately!)

Ballet - but make it AI (Original Video: YouTube)

Here is another, capturing a (ahem) slightly less orthodox form of dance.

Does it make you feel like dancing?

Here is the actual footage with the overlay.

Yup, still looks bats*** insane (Original Video: YouTube)

You would probably not be surprised by the fact that these were generated by an AI model. More specifically, they were generated by a “pose estimation” model going through each YouTube video and identifying key points on the person’s body. These points are then superimposed onto the image, lines are drawn where appropriate, and voila - we have our own dancing skeletons.

That’s not all. Maybe I want to copy Serena’s service motion, or learn how to shoot free throws like Steve Nash:

One of the best ever (Original Video: YouTube)

This gives us more than just visuals. Because the model predicts the location of the joints, I can extract from the videos these experts’ body motion.

No, doesn’t mean that I can immediately copy them, especially given how uncoordinated I am at times. But it sure helps me recognise and understand what’s going on. I could even film myself shooting free throws and compare it quantitatively against anyone else’s. It’s really exciting.

But the applicability of this model isn’t what I wanted to focus on.

What did surprise me were these: a) I could pick up a pre-trained, off-the-shelf model to do this for free, and b) generate these in real-time on my MacBook, without a fancy GPU.

In other words, machine learning is more accessible than ever. Literally anyone can stand on the shoulders of giants as well as to extract the shape made by those very shoulders.

Add that to the ever-growig availability of high-quality datasets, the days of data availability being the bottleneck is fast disappearing.

Just to be clear, I am not arguing that we have all the data we need (far from it - as an example, look up things like the gender data gap). I am simply pointing out that there is more data going around than we know how to deal with, and it’s never been easier to generate more.

People can scrape websites (*its legality appears to be still quite uncertain, despite a famous legal victory), download government data, and take advantage of data from data competitions. And now that machine learning models are more accessible than ever, you can generate your own custom data at little to no cost.

The real value, as always, is in identifying what the data can teach us. Pose estimation by itself is a nice gimmick, like extracting the skeleton of the dancers. But applied to a coaching setting, or sports analytics, there’s going to be great deal of value added. It’s the same with language processing, object detection or whatever. These are just tools. But these tools are clearly getting to a stage where they’re more accessible than ever, both in cost, compute required and ease of use (although there’s quite a lot of work still required there imo).

Get out there and take a look at what’s possible. I was amazed at what I found, and you will probably be too.

FYI: I did the above using the MediaPipe library by google. This repo is a good place to start looking around also.

As a PSA - The pandemic burnout is real. For various personal and pandemic reasons I wasn’t at 100%, and I took on probably too much work at once given the circumstances. I was wrong to think that I could just power through it all and not suffer for it. Look after yourselves!

What a portfolio

X avatar for @adolfux

Adolfo Arranz@adolfux

Today, ten years ago, @SCMPNews started to publish full-back pages infographics. Here you go, almost three hundred graphics. #INFOGRAPHIC #HongKong #newsdesign @SCMPgraphics

3:38 AM · Mar 12, 2021

127 Reposts · 492 Likes

When someone else’s cool DataViz is used to imbue meaning to your word salad.

X avatar for @aaref

Aaref Hilaly@aaref

Data from basketball, but could apply to venture capital. Today it's all about (growth) slamdunks or (early) 3-pointers

10:18 PM · Mar 9, 2021

312 Reposts · 1.62K Likes

@canzhiye said it best:

X avatar for @canzhiye

Canzhi@canzhiye

these analogies are so dumb because if the viewer doesnt know anything about basketball, they dont get it, and if the viewer knows about basketball they just think you must be stupid because the analogy doesnt actually work

6:02 AM · Mar 10, 2021

18 Likes

You can do so much with so little.

X avatar for @jschwabish

Jon Schwabish@jschwabish

Check out this cool animation @alvinwendt made of the list of preattentive attributes shown in Chapter 1 of #BetterDataVisualizations. Thanks, Alvin!

3:26 PM · Mar 3, 2021

145 Reposts · 627 Likes

If you liked the post, you can share it here:

Share

And please remember to subscribe if you haven’t:

No posts

Read the original on visualnoise.substack.com

Comments

Nothing yet. Say the first thing.

    Sign in to join the conversation.