WEBVTT

00:00:00.000 --> 00:00:03.486 align:middle line:90%
[MUSIC PLAYING]

00:00:03.486 --> 00:00:15.950 align:middle line:90%


00:00:15.950 --> 00:00:19.440 align:middle line:84%
Hello, and welcome to a quick
walkthrough of the Textbook

00:00:19.440 --> 00:00:20.624 align:middle line:90%
Data Set using STATA.

00:00:20.624 --> 00:00:22.040 align:middle line:84%
And of course this
is the data set

00:00:22.040 --> 00:00:23.930 align:middle line:84%
that was collected for
Quantitative Research

00:00:23.930 --> 00:00:26.060 align:middle line:84%
Methods for
Communication-- a Hands-On

00:00:26.060 --> 00:00:28.100 align:middle line:90%
Approach, the Fourth Edition.

00:00:28.100 --> 00:00:30.320 align:middle line:84%
So when you open up STATA,
one of the first things

00:00:30.320 --> 00:00:32.903 align:middle line:84%
you're going to see is you have
the big STATA page right here.

00:00:32.903 --> 00:00:35.520 align:middle line:84%
And it's going to let you know
what currently you have open.

00:00:35.520 --> 00:00:39.110 align:middle line:84%
So you can see here that I have
open Textbook Data Set.dta,

00:00:39.110 --> 00:00:41.570 align:middle line:90%
which is the STATA file format.

00:00:41.570 --> 00:00:43.470 align:middle line:84%
Over here on the
right-hand side,

00:00:43.470 --> 00:00:45.572 align:middle line:90%
you have your Variables column.

00:00:45.572 --> 00:00:47.780 align:middle line:84%
Now, this is going to show
you all the variables that

00:00:47.780 --> 00:00:51.570 align:middle line:84%
are currently listed in your
data set and, of course,

00:00:51.570 --> 00:00:53.730 align:middle line:84%
if there's any
labels next to them.

00:00:53.730 --> 00:00:56.210 align:middle line:84%
One of the things I
always do is I always

00:00:56.210 --> 00:01:00.020 align:middle line:84%
label the first item in
an individual measure

00:01:00.020 --> 00:01:02.260 align:middle line:84%
so that I know what
those items go with.

00:01:02.260 --> 00:01:04.849 align:middle line:84%
For example, you'll see
here this is the PRCA1.

00:01:04.849 --> 00:01:08.919 align:middle line:84%
That's the personal report for
communication apprehension 24.

00:01:08.919 --> 00:01:11.210 align:middle line:84%
And of course that's the very
first item in that scale,

00:01:11.210 --> 00:01:13.251 align:middle line:84%
and then you have the
second one all the way down

00:01:13.251 --> 00:01:14.499 align:middle line:90%
to the 24th item.

00:01:14.499 --> 00:01:16.040 align:middle line:84%
And then the next
one right there, it

00:01:16.040 --> 00:01:18.800 align:middle line:84%
goes to ethno1, which is our
measure of ethnocentrism.

00:01:18.800 --> 00:01:21.020 align:middle line:84%
And so you can kind of
scroll through all the way

00:01:21.020 --> 00:01:22.062 align:middle line:90%
to the bottom.

00:01:22.062 --> 00:01:24.270 align:middle line:84%
And then you'll start seeing
some more-- specifically

00:01:24.270 --> 00:01:26.390 align:middle line:84%
it was like biological
sex, politics,

00:01:26.390 --> 00:01:28.040 align:middle line:90%
school classification.

00:01:28.040 --> 00:01:29.840 align:middle line:84%
And then we get to
these right here,

00:01:29.840 --> 00:01:32.300 align:middle line:90%
which are the summed totals.

00:01:32.300 --> 00:01:34.586 align:middle line:84%
This is where we take all
of the items for that.

00:01:34.586 --> 00:01:36.710 align:middle line:84%
For example, in this case,
it's group communication

00:01:36.710 --> 00:01:38.641 align:middle line:84%
apprehension, so we
have all the items

00:01:38.641 --> 00:01:40.640 align:middle line:84%
that are designed to
measure group communication

00:01:40.640 --> 00:01:41.181 align:middle line:90%
apprehension.

00:01:41.181 --> 00:01:44.850 align:middle line:84%
We sum those together, and that
gives us that summed total.

00:01:44.850 --> 00:01:47.330 align:middle line:84%
So this is what it looks
like in your Variable view.

00:01:47.330 --> 00:01:49.970 align:middle line:84%
Now, we can look at it
in a different view.

00:01:49.970 --> 00:01:51.930 align:middle line:84%
So I'm going to come
up here to Data,

00:01:51.930 --> 00:01:55.501 align:middle line:84%
and I'm going to come
down here to Data Editor.

00:01:55.501 --> 00:01:57.500 align:middle line:84%
And I'm only going to
click on Browse, because I

00:01:57.500 --> 00:01:58.340 align:middle line:90%
don't want to edit anything.

00:01:58.340 --> 00:01:59.460 align:middle line:90%
I just want to show you this.

00:01:59.460 --> 00:02:02.126 align:middle line:84%
So I'm going to click on Browse,
and then I'm going to maximize.

00:02:02.126 --> 00:02:05.660 align:middle line:84%
So what this pulls up is it
pulls up a giant spreadsheet

00:02:05.660 --> 00:02:07.000 align:middle line:90%
version of your data set.

00:02:07.000 --> 00:02:09.166 align:middle line:84%
Now, you'll still have the
Variable column over here

00:02:09.166 --> 00:02:10.550 align:middle line:90%
on the right-hand side.

00:02:10.550 --> 00:02:13.130 align:middle line:84%
But then you see basically
how the data is input

00:02:13.130 --> 00:02:16.770 align:middle line:84%
and what it looks like if you
go to any other stats program.

00:02:16.770 --> 00:02:18.920 align:middle line:84%
So here's how you
can understand this.

00:02:18.920 --> 00:02:22.470 align:middle line:84%
The columns themselves
represent individual items.

00:02:22.470 --> 00:02:24.290 align:middle line:84%
So if you go and
look at the survey

00:02:24.290 --> 00:02:26.660 align:middle line:84%
that we collected for
Quantitative Research

00:02:26.660 --> 00:02:28.699 align:middle line:84%
Methods for Communication--
a Hands-On Approach,

00:02:28.699 --> 00:02:30.490 align:middle line:84%
you'll see that there
are individual items.

00:02:30.490 --> 00:02:35.840 align:middle line:84%
Again, here was the PRCA-24
item 1, item 2, item 3,

00:02:35.840 --> 00:02:39.000 align:middle line:84%
all the way over until
we hit right over here,

00:02:39.000 --> 00:02:41.170 align:middle line:90%
which is the 24th item.

00:02:41.170 --> 00:02:42.680 align:middle line:84%
And so those are
just the items that

00:02:42.680 --> 00:02:45.320 align:middle line:84%
are designed to
measure the PRCA-24.

00:02:45.320 --> 00:02:47.180 align:middle line:84%
The rows, on the
other hand, these

00:02:47.180 --> 00:02:49.310 align:middle line:84%
indicate individuals
who actually

00:02:49.310 --> 00:02:50.970 align:middle line:90%
participated in our study.

00:02:50.970 --> 00:02:53.730 align:middle line:84%
So this was our first person,
this is our second person,

00:02:53.730 --> 00:02:56.010 align:middle line:90%
third person, and so on.

00:02:56.010 --> 00:02:58.640 align:middle line:84%
So if I wanted to know what
the first person's score was

00:02:58.640 --> 00:03:00.590 align:middle line:84%
for the very first
item, I can just

00:03:00.590 --> 00:03:03.290 align:middle line:84%
look here at the
PRCA1 with person 1,

00:03:03.290 --> 00:03:05.840 align:middle line:84%
and I can see right there
that their score that they

00:03:05.840 --> 00:03:08.600 align:middle line:90%
put for that item was a 2.

00:03:08.600 --> 00:03:11.230 align:middle line:84%
So that is how you
can actually use this.

00:03:11.230 --> 00:03:13.910 align:middle line:84%
I am going to scroll all the
way over to the right-hand side

00:03:13.910 --> 00:03:15.940 align:middle line:84%
so I can show you some
of those summed totals.

00:03:15.940 --> 00:03:18.140 align:middle line:84%
Now, one of the nice
things about STATA

00:03:18.140 --> 00:03:20.750 align:middle line:84%
is when you get to some of
those nominal variables,

00:03:20.750 --> 00:03:22.430 align:middle line:84%
it actually makes
them nice and blue.

00:03:22.430 --> 00:03:24.900 align:middle line:84%
So it's a little bit easier
for them to jump out.

00:03:24.900 --> 00:03:26.810 align:middle line:84%
So in this case,
it has sex, and it

00:03:26.810 --> 00:03:28.310 align:middle line:84%
lets us know what
those labels are

00:03:28.310 --> 00:03:31.040 align:middle line:84%
that we have for them-- for
example, male versus female.

00:03:31.040 --> 00:03:32.930 align:middle line:84%
Over here we have
Democrat versus other,

00:03:32.930 --> 00:03:35.660 align:middle line:84%
versus Republican, versus
not registered to vote,

00:03:35.660 --> 00:03:36.900 align:middle line:90%
all the way down.

00:03:36.900 --> 00:03:38.090 align:middle line:90%
We have our classification.

00:03:38.090 --> 00:03:40.124 align:middle line:84%
We have how much time
they spent online,

00:03:40.124 --> 00:03:41.540 align:middle line:84%
their age, and
then of course what

00:03:41.540 --> 00:03:43.150 align:middle line:90%
edition this was collected for.

00:03:43.150 --> 00:03:45.410 align:middle line:84%
And then next to
that, we actually

00:03:45.410 --> 00:03:48.750 align:middle line:84%
have the summed totals
for all of those measures.

00:03:48.750 --> 00:03:50.570 align:middle line:84%
For example, this
one right here is

00:03:50.570 --> 00:03:54.710 align:middle line:84%
the what I call big CA, which
is the sum total for all 24

00:03:54.710 --> 00:03:57.560 align:middle line:84%
items on that personal
report for communication

00:03:57.560 --> 00:03:59.480 align:middle line:90%
apprehension-24.

00:03:59.480 --> 00:04:03.060 align:middle line:84%
So that's how you can understand
the Data Editor Browse

00:04:03.060 --> 00:04:07.880 align:middle line:84%
mode is it's just going to show
you what the giant spreadsheet

00:04:07.880 --> 00:04:10.470 align:middle line:84%
version of this data
actually looks like.

00:04:10.470 --> 00:04:12.220 align:middle line:84%
And so that's how you
can understand that.

00:04:12.220 --> 00:04:13.790 align:middle line:84%
I do want to show
you one other thing.

00:04:13.790 --> 00:04:15.290 align:middle line:84%
I just want to come
up here to open.

00:04:15.290 --> 00:04:17.714 align:middle line:84%
I had this other data set
in here called shortened.

00:04:17.714 --> 00:04:19.130 align:middle line:84%
All that I've done
with that one--

00:04:19.130 --> 00:04:21.019 align:middle line:84%
I'm going to go ahead and
open that one-- is instead

00:04:21.019 --> 00:04:22.970 align:middle line:84%
of starting with all
those individual items,

00:04:22.970 --> 00:04:24.920 align:middle line:84%
you'll see that it starts
with biological sex.

00:04:24.920 --> 00:04:27.170 align:middle line:84%
So I've kind of
truncated this version.

00:04:27.170 --> 00:04:30.020 align:middle line:84%
I'll also show you that
over here in Data Editor

00:04:30.020 --> 00:04:31.500 align:middle line:90%
so that we can look at that.

00:04:31.500 --> 00:04:32.810 align:middle line:90%
So again, it starts right here.

00:04:32.810 --> 00:04:34.601 align:middle line:84%
It doesn't have all
those individual items.

00:04:34.601 --> 00:04:37.190 align:middle line:84%
It just makes for
a smaller data set,

00:04:37.190 --> 00:04:39.830 align:middle line:84%
and it's a little
bit easier to use,

00:04:39.830 --> 00:04:43.336 align:middle line:84%
because there's not as much
information in this data set.

00:04:43.336 --> 00:04:44.960 align:middle line:84%
So it's always good,
and periodically I

00:04:44.960 --> 00:04:47.600 align:middle line:84%
recommend using this one
just because it does make

00:04:47.600 --> 00:04:49.020 align:middle line:90%
your life a little bit easier.

00:04:49.020 --> 00:04:51.140 align:middle line:84%
So that is how
you can understand

00:04:51.140 --> 00:04:53.650 align:middle line:90%
your data set in STATA.

00:04:53.650 --> 00:04:56.700 align:middle line:90%
[MUSIC PLAYING]

00:04:56.700 --> 00:05:07.756 align:middle line:90%